# text-to-speech

このトピックのトレンドリポジトリ(7件)

harry0703/MoneyPrinterTurbo

harry0703/MoneyPrinterTurboOtherPython
114.3k13回登場

利用 AI 大模型和自动化工作流,根据主题或关键词一键生成高清短视频。Generate HD short videos from a topic or keyword with an automated AI workflow.

ai-video-generatorcontent-creationffmpeginstagram-reelsllmpythonshort-videosubtitlestext-to-speechtiktokvideo-automationvideo-workflowworkflow-automationyoutube-shorts

unslothai/unsloth

unslothai/unslothOtherPython
73.0k6回登場

Local UI to run and train LLMs and diffusion models, including Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, DeepSeek-V4, FLUX and more.

agentaichatgptdeepseekfine-tuninggemmaimage-generationllamallmllmsopenaipythonqwenreinforcement-learningself-hostedstable-diffusiontext-to-speechttsuiunsloth

calesthio/OpenMontage

calesthio/OpenMontageOtherPython
53.2k13回登場

World's first open-source, agentic video production system. 12 production pipelines, 100+ tools, 700+ agent skill and production-knowledge files. Turn your AI coding assistant into a full video production studio.

agentagentic-aiaiclaudecopilotcursorelevenlabsffmpegfluximage-generationopen-sourceopenaipythonremotionstable-diffusiontext-to-speechtext-to-videovideo-generationvideo-production

OpenBMB/VoxCPM

OpenBMB/VoxCPMOtherPython
30.0k7回登場

VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning

audiodeeplearningminicpmmultilingualpythonpytorchspeechspeech-synthesistext-to-speechttstts-modelvoice-cloningvoice-designvoxcpm

abus-aikorea/voice-pro

abus-aikorea/voice-proOtherPython
11.9k

Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.

audiobookfaster-whispergradiokaraokepodcastsspeech-recognitionspeech-synthesisspeech-to-textsubtitlestext-to-speechtranscriptiontranslatorttsvoice-cloningvoice-conversionwebuiwhisperwhisperxyt-dlp

supertone-inc/supertonic

supertone-inc/supertonicOtherSwift
8.0k6回登場

Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.

cppcsharpfluttergoiosjavalightweightmultilingualnodejson-deviceonnxonnxruntimepythonrustspeech-synthesisswifttext-to-speechttswebwebgpu

OpenMOSS/MOSS-TTS

OpenMOSS/MOSS-TTSOtherPython
2.7k2回登場

MOSS‑TTS Family is an open‑source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high‑fidelity, high‑expressiveness, and complex real‑world scenarios, covering stable long‑form speech, multi‑speaker dialogue, voice/character design, environmental sound effects, and real‑time streaming TTS.

audioaudio-tokenizerllmmultimodaltext-to-speechvoice-cloning