Back to Topics Directory
Topic Hub

#text-to-speech (30 Repositories)

Ranked open-source repositories tagged with #text-to-speech, scored by pull request acceptance likelihood and maintainer engagement velocity.

Topic Avg Merge Rate

54.3%

Avg Review Latency

98.5h

Filter by language

30 repositories tagged #text-to-speech

B TierJavaScript 555 4 GFIs

MartinDelophy/ai-video-editor

Open-source, local-first video editor where creators and AI agents edit the same real timeline.

97.1%
Merge Rate
1d
First Review
100%
1st-Timers
1
Maintainers
A TierGo 113

heygen-com/heygen-cli

Create AI videos from the terminal. Official CLI for the HeyGen video generation API.

89.6%
Merge Rate
1d
First Review
100%
1st-Timers
4
Maintainers
B TierSwift 357 2 GFIs

PowerBeef/Vocello

Vocello: a local, private voice studio for Apple Silicon. Write a script, pick or describe a voice, and generate speech on-device, faster than realtime on an 8 GB M2 Mac mini. Native Swift + MLX, no Python. Mac app out now, iPhone beta on TestFlight. (Formerly QwenVoice.)

93.3%
Merge Rate
4d
First Review
100%
1st-Timers
0
Maintainers
A TierPython 2.4k

pnnbao97/VieNeu-TTS

Vietnamese TTS with instant voice cloning • On-device • Real-time CPU inference • 24kHz audio quality • Chuyển văn bản thành giọng nói tiếng Việt • Text to speech tiếng Việt • TTS tiếng Việt

87.3%
Merge Rate
2d
First Review
100%
1st-Timers
3
Maintainers
B TierPython 91

kadirnar/voicehub

VoiceHub: A Unified Inference Interface for TTS Models

98.4%
Merge Rate
7d
First Review
100%
1st-Timers
0
Maintainers
A TierC++ 1.7k

software-mansion/react-native-executorch

Declarative way to run AI models in React Native on device, powered by ExecuTorch.

88.5%
Merge Rate
4d
First Review
83%
1st-Timers
11
Maintainers
A TierPython 198

ayutaz/piper-plus

Multilingual neural TTS (6 languages: JA/EN/ZH/ES/FR/PT, code supports SV) — C++, C#, Rust, Go, Python, npm (WASM). VITS + Prosody, streaming, CUDA/CoreML/DirectML. pip install piper-plus | npm install piper-plus | cargo install piper-plus-cli

91.5%
Merge Rate
3d
First Review
100%
1st-Timers
2
Maintainers
A TierPython 75.2k 9 GFIs

unslothai/unsloth

Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, FLUX and more.

72.5%
Merge Rate
8h
First Review
46%
1st-Timers
134
Maintainers
A TierPython 5.4k

remsky/Kokoro-FastAPI

Dockerized OpenAI-compatible wrapper for Kokoro-82M text-to-speech w/multiplatform CPU, AMD, NVIDIA GPU PyTorch; multi-speaker, voice-mixing, auto-stitching, caption timestamps, SSML, readalong web UI

87.5%
Merge Rate
15h
First Review
100%
1st-Timers
3
Maintainers
B TierRust 381

izwi-ai/izwi

Voice AI runtime. Local first transcription, speaker diarization, TTS, and voice cloning with an OpenAI compatible API.

96.9%
Merge Rate
18h
First Review
0%
1st-Timers
0
Maintainers
A TierGo 719 1 GFIs

rapidaai/voice-ai

Rapida is an open-source, end-to-end voice AI orchestration platform for building real-time conversational voice agents with audio streaming, STT, TTS, VAD, multi-channel integration, agent state management, and observability.

84.2%
Merge Rate
5d
First Review
100%
1st-Timers
3
Maintainers
B TierDart 117

RWKV-APP/RWKV_APP

Cross-platform, local-first RWKV chat built with Flutter for Android, iOS, Windows, macOS, and Linux.

85.7%
Merge Rate
10d
First Review
100%
1st-Timers
0
Maintainers
B TierPython 118.6k

harry0703/MoneyPrinterTurbo

利用 AI 大模型和自动化工作流,根据主题或关键词一键生成高清短视频。Generate HD short videos from a topic or keyword with an automated AI workflow.

36.2%
Merge Rate
16h
First Review
47%
1st-Timers
40
Maintainers
B TierKotlin 127 1 GFIs

mitchib1440/SpeakThat

The world's most comprehensive notification reader.

100.0%
Merge Rate
18d
First Review
0%
1st-Timers
0
Maintainers
B TierTypeScript 17.5k

leon-ai/leon

🧠 Leon is your open-source personal assistant.

66.7%
Merge Rate
4d
First Review
100%
1st-Timers
1
Maintainers
B TierPython 4.0k

KoljaB/RealtimeTTS

Converts text to speech in realtime

100.0%
Merge Rate
10d
First Review
0%
1st-Timers
0
Maintainers
B TierJavaScript 117

asterics/Asterics-AAC

Free, easy-to-use AAC app with offline support, flexible input options, media & smart home access

61.8%
Merge Rate
3d
First Review
25%
1st-Timers
10
Maintainers
B TierPython 610

lukaszliniewicz/Pandrator

Turn PDFs and EPUBs into audiobooks; subtitles or videos into dubbed videos (including translation), and more. For free. Pandrator uses local models, including voice-cloning (instant, RVC-enhanced, XTTS fine-tuning) and LLM processing. It aspires to be a user-friendly app with a GUI, an installer and all-in-one packages.

30.0%
Merge Rate
7h
First Review
0%
1st-Timers
3
Maintainers
B TierPython 1.5k 1 GFIs

waybarrios/vllm-mlx

High-performance OpenAI and Anthropic compatible LLM inference server for Apple Silicon. Native MLX, continuous batching, multimodal models, MCP tool calling, and Claude Code support.

61.5%
Merge Rate
12d
First Review
34%
1st-Timers
20
Maintainers
C TierSwift 761

Blaizzy/mlx-audio-swift

A modular Swift SDK for audio processing with MLX on Apple Silicon

33.3%
Merge Rate
12d
First Review
0%
1st-Timers
2
Maintainers
C TierPython 61.3k 1 GFIs

RVC-Boss/GPT-SoVITS

1 min voice data can also be used to train a good TTS model! (few shot voice cloning)

28.9%
Merge Rate
4d
First Review
44%
1st-Timers
19
Maintainers
C TierKotlin 311

yuga-hashimoto/openclaw-assistant

OpenClaw voice assistant app for Android - Wake word activation & system assistant integration

37.6%
Merge Rate
12d
First Review
0%
1st-Timers
0
Maintainers
D TierTypeScript 6.3k

promptslab/Awesome-Prompt-Engineering

This repository contains a hand-curated resources for Prompt Engineering with a focus on Generative Pre-trained Transformer (GPT), ChatGPT, PaLM etc

0.0%
Merge Rate
9d
First Review
0%
1st-Timers
2
Maintainers
D TierPython 4.9k

MoonInTheRiver/DiffSinger

DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism (SVS & TTS); AAAI 2022; Official code

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierSwift 284

mlalma/kokoro-ios

Kokoro TTS for iOS and macOSX

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 339

MatteoFasulo/Whisper-TikTok

From AI tools to TikTok video creation using FFMPEG, Microsoft Edge read aloud and OpenAI Whisper model

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 349

keonlee9420/DiffGAN-TTS

PyTorch Implementation of DiffGAN-TTS: High-Fidelity and Efficient Text-to-Speech with Denoising Diffusion GANs

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 39.8k

2noise/ChatTTS

A generative speech model for daily dialogue.

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierJUJupyter Notebook 647

daniilrobnikov/vits2

VITS2: Improving Quality and Efficiency of Single-Stage Text-to-Speech with Adversarial Learning and Architecture Design

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 644

lucasnewman/f5-tts-mlx

Implementation of F5-TTS in MLX

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
Best Text-to-speech Open Source Repositories & C-Rank™ | GetMerged