🎙️ Open source

Speech & Audio

Transcription, text-to-speech, and voice cloning you can self-host.

12repositories
443kstars combined
MO

MoneyPrinterTurbo

harry0703

利用 AI 大模型和自动化工作流,根据主题或关键词一键生成高清短视频。Generate HD short videos from a topic or keyword with an automated AI workflow.

★ 99k · Python · MITai-video-generatorcontent-creationffmpeg
updated 2 days ago View on GitHub →
GP

GPT-SoVITS

RVC-Boss

1 min voice data can also be used to train a good TTS model! (few shot voice cloning)

★ 60k · Python · MITtext-to-speechttsvits
updated today View on GitHub →
TT

TTS

coqui-ai

🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production

★ 46k · Python · MPL-2.0deep-learningglow-ttshifigan
updated 1 years ago View on GitHub →
CH

ChatTTS

2noise

A generative speech model for daily dialogue.

★ 40k · Python · AGPL-3.0agentchatchatgpt
updated 3 months ago View on GitHub →
OP

OpenVoice

myshell-ai

Instant voice cloning by MIT and MyShell. Audio foundation model.

★ 37k · Python · MITtext-to-speechttsvoice-clone
updated 1 years ago View on GitHub →
MO

MockingBird

babysor

🚀Clone a voice in 5 seconds to generate arbitrary speech in real-time

★ 37k · Pythonaideep-learningpytorch
updated 4 months ago View on GitHub →
VO

VoxCPM

OpenBMB

VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning

★ 34k · Python · Apache-2.0audiodeeplearningminicpm
updated 14 days ago View on GitHub →
IN

index-tts

index-tts

An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System

★ 22k · Pythonbigvgancross-lingualindextts
updated 8 days ago View on GitHub →
DI

dia

nari-labs

A TTS model capable of generating ultra-realistic dialogue in one pass.

★ 19k · Python · Apache-2.0aiopen-weighttext-to-speech
updated 8 months ago View on GitHub →
PY

pyvideotrans

jianchang512

Translate the video from one language to another and embed dubbing & subtitles.

★ 18k · Python · GPL-3.0speech-to-texttext-to-speechvideo-transition
updated today View on GitHub →
LE

leon

leon-ai

🧠 Leon is your open-source personal assistant.

★ 17k · TypeScript · MITaiai-agentai-assistant
updated today View on GitHub →
SH

sherpa-onnx

k2-fsa

Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Support embedded systems, Android, iOS, HarmonyOS, Raspberry Pi, RISC-V, RK NPU, Axera NPU, Ascend NPU, x86_64 servers, websocket server/client, support 12 programming languages

★ 14k · C++ · Apache-2.0aarch64androidarm32
updated today View on GitHub →

Where this comes from: GitHub search: topic:text-to-speech stars:>2000, sorted by stars. Last refreshed 22 July 2026. Nothing on this page is a paid placement.