Back to Topics Directory
Topic Hub

#speech-synthesis (25 Repositories)

Ranked open-source repositories tagged with #speech-synthesis, scored by pull request acceptance likelihood and maintainer engagement velocity.

Topic Avg Merge Rate

19.8%

Avg Review Latency

52.3h

Filter by language

25 repositories tagged #speech-synthesis

A TierPython 2.4k

pnnbao97/VieNeu-TTS

Vietnamese TTS with instant voice cloning • On-device • Real-time CPU inference • 24kHz audio quality • Chuyển văn bản thành giọng nói tiếng Việt • Text to speech tiếng Việt • TTS tiếng Việt

87.3%
Merge Rate
2d
First Review
100%
1st-Timers
3
Maintainers
A TierPython 198

ayutaz/piper-plus

Multilingual neural TTS (6 languages: JA/EN/ZH/ES/FR/PT, code supports SV) — C++, C#, Rust, Go, Python, npm (WASM). VITS + Prosody, streaming, CUDA/CoreML/DirectML. pip install piper-plus | npm install piper-plus | cargo install piper-plus-cli

91.5%
Merge Rate
3d
First Review
100%
1st-Timers
2
Maintainers
B TierPython 18.4k

NVIDIA-NeMo/Speech

A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)

69.4%
Merge Rate
3d
First Review
57%
1st-Timers
41
Maintainers
B TierTypeScript 17.5k

leon-ai/leon

🧠 Leon is your open-source personal assistant.

66.7%
Merge Rate
4d
First Review
100%
1st-Timers
1
Maintainers
B TierPython 4.0k

KoljaB/RealtimeTTS

Converts text to speech in realtime

100.0%
Merge Rate
10d
First Review
0%
1st-Timers
0
Maintainers
C TierPython 5.7k

denizsafak/abogen

Generate audiobooks from EPUBs, PDFs and text with synchronized captions.

80.0%
Merge Rate
33d
First Review
0%
1st-Timers
1
Maintainers
D TierPython 9.9k

espnet/espnet

End-to-End Speech Processing Toolkit

0.0%
Merge Rate
3h
First Review
0%
1st-Timers
0
Maintainers
D TierPython 36.3k

OpenBMB/VoxCPM

VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierC 1.5k

s-macke/SAM

Software Automatic Mouth - Tiny Speech Synthesizer

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 3.0k

keithito/tacotron

A TensorFlow implementation of Google's Tacotron speech synthesis with pre-trained model (unofficial)

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 342

keonlee9420/PortaSpeech

PyTorch Implementation of PortaSpeech: Portable and High-Quality Generative Text-to-Speech

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 4.9k

MoonInTheRiver/DiffSinger

DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism (SVS & TTS); AAAI 2022; Official code

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierRust 497

kxxt/aspeak

A simple text-to-speech client for Azure TTS API.

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 2.0k

r9y9/deepvoice3_pytorch

PyTorch implementation of convolutional neural networks-based text-to-speech synthesis models

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierShell 301

kan-bayashi/PytorchWaveNetVocoder

WaveNet-Vocoder implementation with pytorch.

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierJUJupyter Notebook 647

daniilrobnikov/vits2

VITS2: Improving Quality and Efficiency of Single-Stage Text-to-Speech with Adversarial Learning and Architecture Design

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 349

keonlee9420/DiffGAN-TTS

PyTorch Implementation of DiffGAN-TTS: High-Fidelity and Efficient Text-to-Speech with Denoising Diffusion GANs

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 348

liusongxiang/ppg-vc

PPG-Based Voice Conversion

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 339

bshall/ZeroSpeech

VQ-VAE for Acoustic Unit Discovery and Voice Conversion

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierTypeScript 336

Migushthe2nd/MsEdgeTTS

A simple Azure Speech Service module that uses the Microsoft Edge Read Aloud API. https://www.npmjs.com/package/msedge-tts

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierC++ 341

kigner/audio.cpp-webui

audio.cpp with a full-task WebUI - pure C++ audio-model inference engine powered by ggml. TTS, ASR/STT, VAD, voice conversion, speaker diarization, music generation. No Python dependency.

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierJUJupyter Notebook 4.6k

WhisperSpeech/WhisperSpeech

An Open Source text-to-speech system built by inverting Whisper.

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 980

Finrandojin/alexandria-audiobook

AI-powered multi-voice audiobook generator — LLM script annotation, voice cloning, voice design, LoRA training, per-line style control, and export to MP3, chaptered M4B, or Audacity multi-track. Built on Qwen3-TTS.

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierTypeScript 449

echogarden-project/echogarden

Cross-platform speech toolset, used from the command-line or as a Node.js library. Includes a variety of engines for speech synthesis, speech recognition, forced alignment, speech translation, voice isolation, language detection and more.

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierJavaScript 1.5k

AlekPet/ComfyUI_Custom_Nodes_AlekPet

Custom nodes that extend the capabilities of Comfyui

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
Best Speech-synthesis Open Source Repositories & C-Rank™ | GetMerged