Skip to content
The Humanness Index™
Built by VapiGitHub

The Humanness Index™

The open benchmark for how human voice AI sounds, so you can pick the model that passes. Built by Vapi.

MethodologyGitHubContactvapi.ai

Code is Apache-2.0. Standings data is CC BY 4.0. Audio clips and source voices are licensed recordings, all rights reserved. Provider logomarks belong to their respective owners and are used nominatively. “The Humanness Index™” name and logo are Vapi trademarks; see TRADEMARKS.md.

  1. Humanness Index™
  2. ElevenLabs
  3. Flash v2.5

Humanness Index™ · TTS model

ElevenLabs

ElevenLabs Flash v2.5

by ElevenLabs

Flash v2.5 is the multilingual member of ElevenLabs' fastest family, generating speech in about 75 ms across 32 languages.

Rank
#14
Humanness
68
Likely rank
#10–16
Blind votes
1,044

Standings as of Jul 28, 2026, 02:12 UTC

LowerHigher

A real arena clip: a cloned source voice reading a customer support prompt at phone quality.

Flash v2.5 key stats

Latency (measured)
197 ms1
Languages
322
Price / 1M chars
$503
Released
December 18, 20244
  1. Vapi streaming benchmark (50 trials per model) (checked 2026-06-10) Median of 50 sequential live streaming trials, June 2026; includes network RTT from the benchmark machine.
  2. elevenlabs.io/docs/overview/models (checked 2026-06-10) Low-latency tier default (Turbo/Flash v2.5); Turbo v2 and Flash v2 are English-only, Multilingual v2 lists 29, and Eleven v3 lists 70+, encoded per model.
  3. elevenlabs.io/pricing/api (checked 2026-06-10) ElevenAPI pay-as-you-go: Flash/Turbo $0.05 per 1k characters = $50 per 1M. Eleven v3 and Multilingual v2 bill $0.10 per 1k ($100 per 1M), encoded per model.
  4. elevenlabs.io/blog/meet-flash (checked 2026-06-10)

Background

Flash v2.5 is the multilingual member of ElevenLabs' fastest family, generating speech in about 75 ms across 32 languages. Announced in December 2024, it is the model ElevenLabs recommends for real time agents beyond English, and it bills at half the credit cost per character of the flagship v3. That mix makes it the default choice for latency sensitive, cost sensitive voice deployments on the ElevenLabs platform.

Sources: elevenlabs.io

At a glance

Flash v2.5 pairs the Flash latency tier with 32 languages. In our 50 trial streaming benchmark it returned first audio in a median of 197 ms, the fastest ElevenLabs result on the Index.

Sources: elevenlabs.io

Position in the rankings

Standings as of Jul 28, 2026, 02:12 UTC

RankProviderModelHumannessLatency
#12CartesiaCartesiaSonic 3.570128 ms
#13MiniMaxMiniMaxSpeech 2 Turbo70315 ms
#14ElevenLabsElevenLabsFlash v2.568197 ms
#15CartesiaCartesiaSonic 267159 ms
#16CartesiaCartesiaSonic 365166 ms

See the full Humanness Index™ rankings

Frequently asked questions

How is Flash v2.5 tested on the Humanness Index™?
Listeners hear Flash v2.5 against another model in a blind head to head round, both voices reading the same customer support prompt from the same cloned source voice, and they pick whichever sounds more human. Its Humanness score derives purely from those votes.
Which languages does Flash v2.5 support?
Flash v2.5 generates speech in 32 languages, making it the multilingual member of the fastest ElevenLabs family. The English only sibling is Flash v2.

Keep exploring

ElevenLabsElevenLabsAll ElevenLabs models on the IndexElevenLabsTurbo v2Rank #10 · Humanness 75ElevenLabsTurbo v2.5Latency 265 msElevenLabsFlash v2Rank #9 · Humanness 76ElevenLabsEleven v3Rank #1 · Humanness 96ElevenLabsMultilingual v2Latency 1006 ms

Back to the Humanness Index™

Find the most human-sounding voice for your agent.

Compare the models in blind tests, read the methodology, or get in touch.

Read the methodologyStar on GitHub

Build a TTS model? Add yours to the Index.