VoiceStudio
VoiceStudio is a local open-source voice AI suite for cloning, dubbing, dictation, transcription, and long-form audio on Windows,…
VoiceStudio is a local open-source voice AI suite for cloning, dubbing, dictation, transcription, and long-form audio on Windows,…
Speech-to-Speech is Hugging Face's modular open-source voice-agent pipeline, combining VAD, speech recognition, an LLM, and speech synthesis behind…
F5-TTS is an open-source open-source audio ai for developers, creators, and researchers working with speech, voice, transcription, or…
ChatTTS is an open-source open-source audio ai for developers, creators, and researchers working with speech, voice, transcription, or…
GPT-SoVITS is an open-source open-source audio ai for developers, creators, and researchers working with speech, voice, transcription, or…
CosyVoice is an open-source open-source audio ai for developers, creators, and researchers working with speech, voice, transcription, or…
Fish Speech is an open-source open-source audio ai for developers, creators, and researchers working with speech, voice, transcription,…
Faster Whisper is an open-source open-source audio ai for developers, creators, and researchers working with speech, voice, transcription,…
Whisper is an open-source open-source audio ai for developers, creators, and researchers working with speech, voice, transcription, or…
ComfyUI is an open-source, node-based AI creation tool designed for users who want precise control over image, video,…