#Voice & Audio
10 items
RVC-Boss/GPT-SoVITS
Clone a voice and train a text-to-speech model from about one minute of voice data using a WebUI.
debpalash/VoiceStudio
Clone and design voices, dub videos, dictate, transcribe and create audiobooks locally in 646 languages as an ElevenLabs alternative.
coqui-ai/TTS
Generate speech from text with pretrained models, or train and fine-tune your own text-to-speech models, including XTTS voice models.
2noise/ChatTTS
Generate natural conversational speech in English and Chinese, suited to dialogue use cases such as LLM assistants.
OpenBMB/VoxCPM
Generate multilingual speech in 30 languages, design new voices from text descriptions, and clone voices from short reference clips.
myshell-ai/OpenVoice
Clone a voice from a reference clip and generate speech in multiple languages with control over emotion, accent, and rhythm.
babysor/MockingBird
Clone a voice from a few seconds of audio and generate arbitrary speech in real time, with Mandarin support.
index-tts/index-tts
Clone a voice from one reference clip and generate controllable speech in Chinese, English, Japanese, Spanish, and Arabic.
QwenAudio/CosyVoice
Run multilingual, zero-shot voice cloning text-to-speech with support for Chinese dialects, plus training and deployment tools.
nari-labs/dia
Turn a transcript into realistic English dialogue in one pass, including nonverbal sounds like laughter, with audio-based emotion control.