Skip to content
The Humanness Index™
Built by VapiGitHub

The Humanness Index™

The open benchmark for how human voice AI sounds, so you can pick the model that passes. Built by Vapi.

MethodologyGitHubContactvapi.ai

Code is Apache-2.0. Standings data is CC BY 4.0. Audio clips and source voices are licensed recordings, all rights reserved. Provider logomarks belong to their respective owners and are used nominatively. “The Humanness Index™” name and logo are Vapi trademarks; see TRADEMARKS.md.

  1. Humanness Index™
  2. Speechify
  3. Simba 3.2

Humanness Index™ · TTS model

Speechify

Speechify Simba 3.2

by Speechify

Simba 3.2 is Speechify's streaming native flagship, released to the SpeechifyAI API in July 2026 as the recommended model for new English integrations.

Rank
#7
Humanness
83
Likely rank
#1–16
Blind votes
90

Standings as of Jul 27, 2026, 22:55 UTC

LowerHigher

A real arena clip: a cloned source voice reading a customer support prompt at phone quality.

Simba 3.2 key stats

Latency (measured)
—Not measured: no publicly reachable API at benchmark time. The Index never shows vendor latency estimates.
Languages
English1
Price / 1M chars
$102
Voice cloning
zero-shot, manual approval3
Released
July 8, 20264
  1. docs.speechify.ai/build/guides/concepts/models (checked 2026-07-23) English only at launch; Speechify says multilingual support will land under the same model id.
  2. speechify.ai/text-to-speech-api (checked 2026-07-23) Flat per-character rate by plan: $10 per 1M characters on Starter, $8 on Pro, $6 on Scale; no credit conversion.
  3. docs.speechify.ai/build/guides/concepts/models (checked 2026-07-23) Cloned voices are supported on simba-3.2, but each voice key currently requires manual Speechify approval.
  4. docs.speechify.ai/build/changelog/2026/7/8 (checked 2026-07-23) API availability of simba-3.2 on the speech and stream endpoints.

Background

Simba 3.2 is Speechify's streaming native flagship, released to the SpeechifyAI API in July 2026 as the recommended model for new English integrations. Speechify positions it on expressivity and time to first byte, and it debuted statistically tied for first place on the Artificial Analysis TTS leaderboard. It serves a curated voice allow list, with cloned voices supported behind a manual approval step.

Sources: docs.speechify.ai, speechify.ai

At a glance

The arena clips for Simba 3.2 were rendered by Speechify with the four licensed source voices cloned on its platform, then verified and hosted by the Index team under the same frozen content hash scheme as every other model, the same vendor supplied path used where the pipeline has no API access. Latency shows a dash because the 50 trial streaming benchmark has not run against a Speechify key yet; we never substitute vendor figures.

Sources: docs.speechify.ai

Position in the rankings

Standings as of Jul 27, 2026, 22:55 UTC

RankProviderModelHumannessLatency
#5MiniMaxMiniMaxSpeech 2 HD89357 ms
#6xAIxAIGrok TTS (Streaming)86285 ms
#7SpeechifySpeechifySimba 3.283—
#8InworldInworldTTS-1.5-max78337 ms
#9ElevenLabsElevenLabsFlash v276226 ms

See the full Humanness Index™ rankings

Frequently asked questions

How is Simba 3.2 tested on the Humanness Index™?
Listeners hear Simba 3.2 against another model in a blind head to head round, both voices reading the same customer support prompt from the same cloned source voice, and they pick whichever sounds more human. Its Humanness score derives purely from those votes.
Where did the Simba 3.2 arena clips come from?
Speechify rendered the 80 arena clips (four cloned source voices reading the 20 frozen prompts) with simba-3.2 and supplied them to the Index team, who normalized, verified, and hosted them under the frozen content hash scheme. Blind battles and scoring work exactly as for every other model.

Keep exploring

SpeechifySpeechifyAll Speechify models on the Index

Back to the Humanness Index™

Find the most human-sounding voice for your agent.

Compare the models in blind tests, read the methodology, or get in touch.

Read the methodologyStar on GitHub

Build a TTS model? Add yours to the Index.