OmniVoice
OmniVoice is a state-of-the-art massively multilingual zero-shot text-to-speech (TTS) model supporting over 600 languages. Built on a novel diffusion language model-style architecture, it generates high-quality speech with superior inference speed, supporting voice cloning and voice design. We recommend using a fresh virtual environment (e.g., conda, venv, etc.) to avoid conflicts. Intel Arc GPUs (Alchemist and Battlemage architectures) are supported via PyTorch's XPU backend.
View OmniVoice on GitHub
OmniVoice is a state-of-the-art massively multilingual zero-shot text-to-speech (TTS) model supporting over 600 languages. Built on a novel diffusion language model-style architecture, it generates high-quality speech with superior inference speed, supporting voice cloning and voice design.
We recommend using a fresh virtual environment (e.g., conda, venv, etc.) to avoid conflicts.
Intel Arc GPUs (Alchemist and Battlemage architectures) are supported via PyTorch's XPU backend.
OmniVoice at a glance
| Stars | 14k |
|---|---|
| Forks | 2.1k |
| Language | Python |
| License | Apache-2.0 |
| Last update | 2026-09-14 |
| Contributors | 21 |
How to install OmniVoice
bash pip install torch==2.8.0 torchaudio==2.8.0
Where OmniVoice is listed
Navid.me is reader-supported. When you buy through links on this site, I may earn an affiliate commission. Learn more.
GitHub repos like this
More repo topics
More free tools
Related MCP servers & CLIs
The most actionable AI newsletter for founders
Every week, get proven AI strategies, curated tools, and step-by-step systems to grow your audience, create better content, and build a profitable creator business.
No fluff, no filler, no BS. Just five minutes each week that might level up your online business and life.
P.S. Sign up now to get free access to my ultimate AI tools guide for creators.






































