STORY · PRODUKTER_
Audio8 launches TTS model with 0.6B parameters featuring zero-shot voice cloning
Audio8 has launched a multilingual text-to-speech model with 0.6 billion parameters supporting zero-shot voice cloning. The model uses a DualAR architecture and supports 11 languages including English, Mandarin, Japanese, and several European languages.
WHY IT MATTERS
The model represents an attempt to achieve high-quality text-to-speech in a compact size, making it accessible to more users and implementations.
SOURCES
MACHINE-GENERATED SUMMARY This summary is written by machine from the sources below. We sort and explain — but we are a way into the field, not the final word. Check the source when something matters to you.