# voice-cloning

このトピックのトレンドリポジトリ(4件)

OpenBMB/VoxCPM

OpenBMB/VoxCPMOtherPython
30.0k7回登場

VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning

audiodeeplearningminicpmmultilingualpythonpytorchspeechspeech-synthesistext-to-speechttstts-modelvoice-cloningvoice-designvoxcpm

debpalash/VoiceStudio

debpalash/VoiceStudioOtherPython
17.0k3回登場

VoiceStudio is the open-source, fully-local ElevenLabs alternative — voice cloning, voice design, video dubbing, dictation, transcription & audiobook creation in 646 languages.

aiaudiobookcudadubbingelevenlabs-alternativehuggingfacelocal-firstmlxomnivoice-studiospeech-to-texttauritext-to-speechtranscriptiontranslatettsvoice-aivoice-cloningvoice-generationvoicestudioworkflow

abus-aikorea/voice-pro

abus-aikorea/voice-proOtherPython
11.9k

Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.

audiobookfaster-whispergradiokaraokepodcastsspeech-recognitionspeech-synthesisspeech-to-textsubtitlestext-to-speechtranscriptiontranslatorttsvoice-cloningvoice-conversionwebuiwhisperwhisperxyt-dlp

OpenMOSS/MOSS-TTS

OpenMOSS/MOSS-TTSOtherPython
2.7k2回登場

MOSS‑TTS Family is an open‑source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high‑fidelity, high‑expressiveness, and complex real‑world scenarios, covering stable long‑form speech, multi‑speaker dialogue, voice/character design, environmental sound effects, and real‑time streaming TTS.

audioaudio-tokenizerllmmultimodaltext-to-speechvoice-cloning