ModelRadar
/

ZONOS2

Zyphra🇺🇸 USMajor lab· Jun 11, 2026

Audioopen weights

ZONOS2 is our latest text-to-speech model trained on more than 6 million hours of varied multilingual speech, delivering expressiveness and quality on par with—or even surpassing—top TTS providers at low latency with MoE. ZONOS2 excels at high-fidelity and naturalistic voice cloning.

From the model card

License
apache-2.0
Modalities
text → audio

2 variants

GGUFHF
StandardHF

Sources

HuggingFace

Provenance

Detected
Sep 27, 2026
First source
huggingface
Description
From the model card
Editorially reviewed
—