Voice Cloning
- Multimedia
40,145 Stars and a 6-Point Hacker News Thread: Auditing VoiceStudio, the 'Fully-Local ElevenLabs Alternative' Whose Default Engine Is CC-BY-NC
VoiceStudio (AGPL-3.0, formerly OmniVoice Studio) reached 40,145 GitHub stars in under six months on the promise of 'fully-local voice cloning, dubbing and audiobooks in …
- Open Source
VoxCPM2: OpenBMB's Tokenizer-Free TTS for Multilingual Speech Generation
VoxCPM2 is a tokenizer-free text-to-speech (TTS) model developed by OpenBMB, an open-source AI research community affiliated with Tsinghua University...
- Multimedia
IndexTTS-vLLM: Accelerated Open-Source Text-to-Speech with vLLM Inference
Text-to-speech technology has advanced dramatically in the past three years. Zero-shot voice cloning, where a system can synthesize speech in a novel...
- Multimedia
Higgs Audio: Boson AI's Open-Source Text-Audio Foundation Model
Text-to-speech technology has advanced dramatically in recent years, transitioning from robotic, monotone synthesis to remarkably natural voice...
- Open Source
GPT-SoVITS: Few-Shot Voice Cloning with Just 1 Minute of Voice Data
GPT-SoVITS is an open-source voice cloning and text-to-speech system developed by RVC-Boss that has taken the AI audio community by storm. The...
- Multimedia
CosyVoice: Alibaba's Open-Source Multi-Lingual Voice Generation Model
Voice generation technology has seen remarkable progress, but most open-source text-to-speech (TTS) models still struggle with a fundamental...