MiniMax models on the Humanness Index™
About MiniMax
MiniMax is the Shanghai AI lab behind the MiniMax speech line, served through its hosted t2a_v2 API with HD and Turbo tiers. It is widely regarded as the strongest text to speech provider for Chinese, and its recent generations brought English accuracy and rhythm up alongside that strength.
Sources: minimax.io
Speech generations
The speech line moved fast through 2025 and into 2026, with the Speech-02 series arriving in April 2025, Speech 2.5 following in August, and the Speech 2.8 generation current since early 2026. The current generation supports more than 40 languages and clones a voice from roughly six to ten seconds of reference audio using a learnable speaker encoder that needs no transcript. The clips on this Index were generated with the Speech 2.8 generation, turbo tier.
Sources: platform.minimax.io, platform.minimax.io
Find the most human-sounding voice for your agent.
Compare the models in blind tests, read the methodology, or get in touch.
Build a TTS model? Add yours to the Index.