Voice Cloning

  1. 40,145 Stars and a 6-Point Hacker News Thread: Auditing VoiceStudio, the 'Fully-Local ElevenLabs Alternative' Whose Default Engine Is CC-BY-NC

    VoiceStudio (AGPL-3.0, formerly OmniVoice Studio) reached 40,145 GitHub stars in under six months on the promise of 'fully-local voice cloning, dubbing and audiobooks in …

    Multimedia
  2. VoxCPM2: OpenBMB's Tokenizer-Free TTS for Multilingual Speech Generation

    VoxCPM2 is a tokenizer-free text-to-speech (TTS) model developed by OpenBMB, an open-source AI research community affiliated with Tsinghua University...

    Open Source
  3. IndexTTS-vLLM: Accelerated Open-Source Text-to-Speech with vLLM Inference

    Text-to-speech technology has advanced dramatically in the past three years. Zero-shot voice cloning, where a system can synthesize speech in a novel...

    Multimedia
  4. Higgs Audio: Boson AI's Open-Source Text-Audio Foundation Model

    Text-to-speech technology has advanced dramatically in recent years, transitioning from robotic, monotone synthesis to remarkably natural voice...

    Multimedia
  5. GPT-SoVITS: Few-Shot Voice Cloning with Just 1 Minute of Voice Data

    GPT-SoVITS is an open-source voice cloning and text-to-speech system developed by RVC-Boss that has taken the AI audio community by storm. The...

    Open Source
  6. CosyVoice: Alibaba's Open-Source Multi-Lingual Voice Generation Model

    Voice generation technology has seen remarkable progress, but most open-source text-to-speech (TTS) models still struggle with a fundamental...

    Multimedia