2 articles
VoxCPM has launched an open-source text-to-speech system that ditches tokenization for more natural, expressive speech. The model handles context-aware generation and clones voices from just five seconds of audio.
OpenBMB has released VoxCPM, a tokenizer-free text-to-speech model that clones voices in real-time with streaming inference and LoRA fine-tuning capabilities.
We use cookies to improve your experience on our site and to show you relevant advertising. To find our more, read our privacy policy and cookie policy