
Source
IndexTTS2 official model/project documentation
Read the primary source.
Reference illustration, not a screenshot of this source.
Speech-generation engine · Voice
Controllable zero-shot speech-generation system.

Evaluate when precise control over the generated speech is needed.
Do not treat a model listing as installed and ready.
Text; Reference audio where supported
Speech audio
Model runtime or compatible application
Provider capabilities and historical Studio notes are distinguished below. Current account access, credits and runtime readiness were not tested.
Model compute; check the selected release and weights.
Self-hosted/runtime dependent.
IndexTTS 2.5 documents Chinese, English, Japanese, Spanish and Arabic; do not assume French support.
Editorial task fit based on the documented role and integration requirements. Check the released model and supported controls. A documented feature is not a verified local integration.
Decision: Evaluate when precise control over the generated speech is needed.
Do not transfer capabilities between versions: the 2.0 release notes say precise duration control was not enabled; 2.5 documents a speaking-speed duration factor. Check supported languages.
Official model/project documentation reviewed; no same-script listening test
Catalogue review: 2026-09-05. The evidence note identifies what was verified. This date does not imply a fresh tool test or verification of every linked source.