2026 Benchmarks Updated

Miralas Voice Model Arena

Compare Miralas against leading voice AI systems using verified model capabilities — then plug in your own measured audio samples and internal benchmarks.

Miralas Baseline
500M
Chatterbox Multilingual V3
Baseline Languages
23+
Official Chatterbox multilingual model
Audio I/O
Realtime
Miralas pipeline + voice samples
Uzbek Track
Native
Miralas proprietary training direction

Miralas Core Architecture

Real-Time Audio Intelligence Demo

Interactive Voice Arena

Test and compare how different models handle complex user prompts in real-time conversation. Based on Full-Duplex-Bench-v3 scenarios.

User Prompt • Strategic Reasoning

“I'm considering a 900-square-foot indie coffee shop beside a commuter rail station. Give me a strategic pre-mortem: if this fails after a year, what probably happened?”

Miralas

Chatterbox Multilingual V3 baseline + Miralas training

internal
0:00 / 0:12

GPT-Realtime

OpenAI

official
0:00 / 0:12

Gemini 3.1 Flash Live

Google

official
0:00 / 0:12

Grok Voice

xAI

official
0:00 / 0:12

Chatterbox Multilingual V3 Engine

Miralas uses Chatterbox Multilingual V3 as its open-source TTS baseline, then focuses its own training and evaluation work on languages and voices that matter to our users. The official V3 model is listed at 500M parameters and 23+ supported languages.

English

en • reference sample

Spanish

es • reference sample

Chinese

zh • reference sample

Hindi

hi • reference sample

Arabic

ar • reference sample

Japanese

ja • reference sample

Custom Voice Clones & Fine-Tuning

Train your proprietary voice models using 5 minutes of clean audio data with emotion preservation and speaker embedding stability.

Shahzoda — Custom Voice Pipeline

Dataset: 45 min • Status: Internal evaluation • Add measured MOS after your test run

Corporate Brand Voice — Acme Corp

Dataset: 120 min • Status: Fine-tuning • Metrics shown after internal evaluation

Live Performance Benchmarks

Continuous accuracy and latency testing against leading industry solutions. Data sourced from Artificial Analysis Speech-to-Speech Index, Full-Duplex-Bench-v3, and Trelis Research (2026).

Verified Capability Comparison

These are product/model capabilities documented by the vendors, not fabricated cross-vendor benchmark scores. For latency, MOS, WER and reasoning, run the same script, prompt, hardware and audio set across every model.

CapabilityMiralasGPT-RealtimeGemini 3.1 LiveGrok Voice
Realtime audio
Audio input/output
Voice cloning
Open-source baseline
Custom language training
Uzbek-first training track
WebRTC / realtime APIInternalLive APIVendor API
Miralas Baseline
500M
Chatterbox Multilingual V3
Chatterbox Languages
23+
Official multilingual baseline
GPT Realtime
Audio I/O
Realtime API
Gemini Live
A2A Audio
Gemini 3.1 Flash Live
OpenAI: GPT-Realtime supports realtime text/audio over WebRTC, WebSocket and SIP.Google: Gemini 3.1 Flash Live is documented as a low-latency audio-to-audio model.Chatterbox: Multilingual V3 is documented as 500M and 23+ languages.
Benchmark policy: Miralas performance numbers will only be shown here after you run the same evaluation suite on the same hardware and publish the methodology. This prevents the page from presenting invented MOS, WER, latency or leaderboard claims as independent research.

Real-Time System Metrics

Live telemetry from your Miralas endpoint. If the endpoint is unavailable, the page does not invent vendor metrics.

Audio Intelligence Accuracy

Miralas Chatterbox V30%
GPT Realtime0%
Gemini Voice0%
Grok Audio0%

Time to First Byte (Latency)

Miralas (Edge)0ms
GPT Realtime0ms
Gemini Voice0ms
Grok Audio0ms

Upcoming Languages & Expansion

Languages currently scaling in our regional training pipelines. Uzbek voice casting is now active.

🇬🇧Available

English

Baseline / evaluation

Progress100%
🇪🇸Available

Spanish

Baseline / evaluation

Progress100%
🇨🇳Available

Chinese

Baseline / evaluation

Progress100%
🇮🇳Available

Hindi

Baseline / evaluation

Progress100%
🇸🇦Available

Arabic

Baseline / evaluation

Progress100%
🇯🇵Available

Japanese

Baseline / evaluation

Progress100%
🇺🇿In training

Uzbek

Miralas native-language training

Progress22%

Uzbek Voice Pipeline — Active Development

Uzbek is the core language research direction for Miralas. The goal is to build a native-quality training pipeline around Uzbek phoneme coverage, natural prosody, regional variation and code-switching. Replace this status with your real dataset and training telemetry as the pipeline progresses.

Research TrackNative UzbekLive Dataset Metrics