Explore voiceagent.best
Estimate costOur independent methodology

Start here

Comparison LabCompare 27 voice AI platforms and components across seven buying contexts, with sourced capabilities, current price captures and visible unknowns.Voice AI Price Index: prices, evidence and historyUnderstand the complete voice-agent cost stack: models, voice, transcription, telephony, subscriptions, capacity and add-ons.Know the cost before the first callEstimate monthly voice-agent costs, AI minutes, setup expenses and volume scenarios using your own transparent assumptions.AI voice agents for dentalAppointments, rescheduling and after-hours calls—with a person ready for clinical questions. Explore call flows, safe automation, integrations, tests and cost considerations.

VoiceAgent Benchmark Intelligence

Inspect published voice-model evidence with its original dimensions, provenance and limitations.

46Model records
5Source architecture types
1Public benchmark source

Third-party Benchmark · VoiceBench · Captured 2026-09-27

These are published model results. VoiceAgent.best ran no model calls or voice-agent tests. They do not measure a platform, phone call, production latency, voice quality or handoff reliability.

Leaderboard · Repository · Paper · Dataset card · Apache-2.0 repository terms

Citation: Chen, Yue, Zhang, Gao, Tan and Li (2024). VoiceBench: Benchmarking LLM-Based Voice Assistants. arXiv:2410.17196.

Choose the question, then inspect the evidence

Architecture context

Inspect group coverage and metric ranges, with unequal samples and selection limits visible.

Cascaded vs Omni →

Comparable published rows

GPT-4o-Audio vs GPT-4o-mini-Audio →

Whisper-v3-large + GPT-4o vs GPT-4o-Audio →

Inspect the public result rows

Filter by the classifications published by VoiceBench. “Open” describes its weight-availability label; it does not establish a commercial license or a platform’s availability.

46 model records

Published model scores, alphabetically ordered. Each dimension keeps its source scale.
Third-party Benchmark · VoiceBench · 2026-09-27 · Exact checkpoint and protocol version unknown.
ModelArchitectureWeightsAlpacaEvalsource 5-point scaleCommonEvalsource 5-point scaleWildVoicesource 5-point scaleSD-QAsource score /100MMSUsource score /100OBQAsource score /100BBHsource score /100IFEvalsource score /100AdvBenchsource score /100VoiceBench OverallSource reported /100Provenance
Baichuan-AudioOmniCommunity labelopenSource classification4.414.083.9245.8453.1971.6554.8050.3199.4269.27VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Baichuan-Omni-1.5OmniCommunity labelopenSource classification4.504.054.0643.4057.2574.5162.7054.5497.3171.32VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
BR-Voice-ReasonerVision + Audio + LLMCommunity labelopenSource classification4.874.714.6674.3085.7096.0090.4083.2099.8090.46VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
DiVAAudio-LLMCommunity labelopenSource classification3.673.543.7457.0525.7625.4951.8039.1598.2757.39VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Freeze-OmniS2S / Full-DuplexCommunity labelopenSource classification4.033.463.1553.4528.1430.9850.7023.4097.3055.20VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
GLM-4-VoiceOmniCommunity labelopenSource classification3.973.423.1836.9839.7553.4152.8025.9288.0856.48VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
GPT-4o-AudioOmniCommunity labelclosedSource classification4.784.494.5875.5080.2589.2384.1076.0298.6586.75VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
GPT-4o-mini-AudioOmniCommunity labelclosedSource classification4.754.244.4067.3672.9084.8481.5072.9098.2782.84VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
IchigoOmniCommunity labelopenSource classification3.793.172.8336.5325.6326.5946.5021.5957.5045.57VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Kimi-AudioOmniCommunity labelopenSource classification4.463.974.2063.1262.1783.5269.7061.10100.0076.91VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
LFG-1Audio-LLMCommunity labelopenSource classification4.603.714.0162.3975.6082.4283.9074.8591.1579.63VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
LFG-2Audio-LLMCommunity labelopenSource classification4.674.264.2373.4285.2093.6387.1084.5495.7786.98VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
LFG-3Audio-LLMCommunity labelopenSource classification4.734.404.4578.1285.5294.7392.2088.5498.2789.88VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
LLaMA-OmniOmniCommunity labelopenSource classification3.703.462.9239.6925.9327.4749.2014.8711.3541.12VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Lyra-BaseOmniCommunity labelopenSource classification3.853.503.4238.2549.7472.7559.0036.2859.6259.00VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Lyra-MiniOmniCommunity labelopenSource classification2.992.692.5819.8931.4241.5448.4020.9180.0045.26VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Mair-hub-0.5B-OmniOmniCommunity labelopenSource classification3.062.872.4821.7025.6025.2750.9014.8594.8144.59VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Megrez-3B-OmniOmniCommunity labelopenSource classification3.502.952.3425.9527.0328.3550.3025.7187.6946.76VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
MERaLiONAudio-LLMCommunity labelopenSource classification4.503.774.1255.0634.9527.2362.6062.9394.8165.04VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Mini-OmniOmniCommunity labelopenSource classification1.952.021.6113.9224.6926.5946.3013.5837.1230.42VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Mini-Omni2OmniCommunity labelopenSource classification2.322.181.799.3124.2726.5946.4011.5657.5033.49VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
MiniCPM-oOmniCommunity labelopenSource classification4.424.153.9450.7254.7878.0260.4049.2597.6971.23VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
MoshiS2S / Full-DuplexCommunity labelopenSource classification2.011.601.3015.6424.0425.9347.4010.1244.2329.51VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Nemotron 3 VoiceChat (V1)S2S / Full-DuplexCommunity labelopenSource classification3.593.123.0643.4050.4664.4052.7017.1299.6258.10VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
NVIDIA Nemotron 3 Nano Omni 30B A3BVision + Audio + LLMCommunity labelopenSource classification4.754.574.5871.4382.3092.9791.1088.66100.0089.39VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
OlaOmniCommunity labelopenSource classification4.122.973.1933.8245.9767.9151.1039.5790.7759.42VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Parakeet-TDT-0.6b-V2 + Qwen3-8BCascadedCommunity labelopenSource classification4.684.464.3547.4759.1080.0077.9078.9999.8179.23VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Phi-4-multimodalVision + Audio + LLMCommunity labelopenSource classification3.813.823.5639.7842.1965.9361.8045.35100.0064.32VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Qwen2-AudioAudio-LLMCommunity labelopenSource classification3.743.433.0135.7135.7249.4554.7026.3396.7355.80VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Qwen3-Omni-30B-A3B-InstructOmniCommunity labelopenSource classification4.744.544.5876.9068.1089.7080.4077.8099.3085.50VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Qwen3-Omni-30B-A3B-ThinkingVision + Audio + LLMCommunity labelopenSource classification4.824.534.5378.1083.0094.3088.9080.6097.2088.80VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
SLAM-OmniOmniCommunity labelopenSource classification1.901.791.604.1626.0625.2748.8013.3894.2335.30VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Step-AudioOmniCommunity labelopenSource classification4.133.092.9344.2128.3333.8550.6027.9669.6250.84VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Ultravox-GLM-4P6Audio-LLMCommunity labelopenSource classification4.934.424.5784.2480.4089.4581.4875.5999.2387.05VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Ultravox-GLM-4P7Audio-LLMCommunity labelopenSource classification4.874.304.5584.1583.8294.7287.1676.2699.2388.86VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Ultravox-GLM-4P7 (thinking)Audio-LLMCommunity labelopenSource classification4.694.054.2188.5787.1595.0189.7880.9998.4688.79VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Ultravox-v0.4.1-LLaMA-3.1-8BAudio-LLMCommunity labelopenSource classification4.553.904.1253.3547.1765.2766.3066.8898.4672.09VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Ultravox-v0.5-LLaMA-3.1-8BAudio-LLMCommunity labelopenSource classification4.594.114.2858.6854.1668.3567.8066.5198.6574.86VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Ultravox-v0.5-LLaMA-3.2-1BAudio-LLMCommunity labelopenSource classification4.043.573.4734.7230.0335.6052.7045.5696.9257.46VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Ultravox-v0.6-LLaMA-3.3-70BAudio-LLMCommunity labelopenSource classification4.694.264.3882.6069.2086.4078.8061.5091.2081.81VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
VITA-1.0OmniCommunity labelopenSource classification3.382.151.8727.9425.7029.0147.7022.8226.7336.43VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
VITA-1.5OmniCommunity labelopenSource classification4.213.663.4838.8852.1571.6555.3038.1497.6964.53VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Whisper-v3-large + GPT-4oCascadedCommunity labelclosedSource classification4.804.474.6275.7781.6992.9787.2076.5198.2787.80VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Whisper-v3-large + LLaMA-3.1-8BCascadedCommunity labelopenSource classification4.534.044.1670.4362.4372.5369.7069.5398.0877.48VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Whisper-v3-turbo + LLaMA-3.1-8BCascadedCommunity labelopenSource classification4.554.024.1258.2362.0472.0969.1071.1298.4676.09VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27
Whisper-v3-turbo + LLaMA-3.2-3BCascadedCommunity labelopenSource classification4.453.824.0449.2851.3760.6663.9069.7198.0871.02VoiceBench OverallVoiceBenchThird-party Benchmark
2026-09-27

Qwen3-Omni rows import official results: the leaderboard divides the three open-ended QA values by 20 and rounds to two decimals; Overall remains as officially reported. We preserve these published values and create no universal score. Run settings and checkpoint equivalence are unverified.

Official model context stays separate

Two Qwen variants have source-linked input, output and component specifications. Other rows retain unknown official specifications until exact model evidence is attached. A published family name is not enough to infer pricing, latency or platform support.

Qwen3-Omni Instruct context →

Source capture

Checked 2026-09-27 · Apache-2.0 repository and dataset-card labels · citation retained · metadata only

Artifact SHA-256: 422d8ec911076053b9f4b3cd5cc38909a4e3d75a9c0ed093f05279557800c993
Repository revision observed: 6992cf4fc51d0426c52c4805b5002e0aae49118a

Read the capture, rights and comparison methodology →