OmniVoice: Towards Omnilingual Zero-Shot Text-to-Speech with Diffusion Language Models

Source

arxiv.orgfull article ↗

Publisher summary· verbatim

arXiv:2604.00688v3 Announce Type: replace Abstract: We present OmniVoice, a massively multilingual zero-shot text-to-speech (TTS) model that scales to over 600 languages. At its core is a novel diffusion language model-style discrete non-autoregressive (NAR) architecture. Unlike conventional discret

Stay posted· Newsletter

A 5-min weekly brief — top movers, price watch, story of the week.

Discussion

No replies yet. Be first.

OmniVoice: Towards Omnilingual Zero-Shot Text-to-Speech with Diffusion Language Models

Related coverage

OmniVoice: Towards Omnilingual Zero-Shot Text-to-Speech with Diffusion Language Models

Related coverage