#speech-synthesis (25 Repositories)
Ranked open-source repositories tagged with #speech-synthesis, scored by pull request acceptance likelihood and maintainer engagement velocity.
19.8%
52.3h
25 repositories tagged #speech-synthesis
pnnbao97/VieNeu-TTS
Vietnamese TTS with instant voice cloning • On-device • Real-time CPU inference • 24kHz audio quality • Chuyển văn bản thành giọng nói tiếng Việt • Text to speech tiếng Việt • TTS tiếng Việt
ayutaz/piper-plus
Multilingual neural TTS (6 languages: JA/EN/ZH/ES/FR/PT, code supports SV) — C++, C#, Rust, Go, Python, npm (WASM). VITS + Prosody, streaming, CUDA/CoreML/DirectML. pip install piper-plus | npm install piper-plus | cargo install piper-plus-cli
NVIDIA-NeMo/Speech
A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)
leon-ai/leon
🧠 Leon is your open-source personal assistant.
KoljaB/RealtimeTTS
Converts text to speech in realtime
denizsafak/abogen
Generate audiobooks from EPUBs, PDFs and text with synchronized captions.
espnet/espnet
End-to-End Speech Processing Toolkit
OpenBMB/VoxCPM
VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning
s-macke/SAM
Software Automatic Mouth - Tiny Speech Synthesizer
keithito/tacotron
A TensorFlow implementation of Google's Tacotron speech synthesis with pre-trained model (unofficial)
keonlee9420/PortaSpeech
PyTorch Implementation of PortaSpeech: Portable and High-Quality Generative Text-to-Speech
MoonInTheRiver/DiffSinger
DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism (SVS & TTS); AAAI 2022; Official code
kxxt/aspeak
A simple text-to-speech client for Azure TTS API.
r9y9/deepvoice3_pytorch
PyTorch implementation of convolutional neural networks-based text-to-speech synthesis models
kan-bayashi/PytorchWaveNetVocoder
WaveNet-Vocoder implementation with pytorch.
daniilrobnikov/vits2
VITS2: Improving Quality and Efficiency of Single-Stage Text-to-Speech with Adversarial Learning and Architecture Design
keonlee9420/DiffGAN-TTS
PyTorch Implementation of DiffGAN-TTS: High-Fidelity and Efficient Text-to-Speech with Denoising Diffusion GANs
liusongxiang/ppg-vc
PPG-Based Voice Conversion
bshall/ZeroSpeech
VQ-VAE for Acoustic Unit Discovery and Voice Conversion
Migushthe2nd/MsEdgeTTS
A simple Azure Speech Service module that uses the Microsoft Edge Read Aloud API. https://www.npmjs.com/package/msedge-tts
kigner/audio.cpp-webui
audio.cpp with a full-task WebUI - pure C++ audio-model inference engine powered by ggml. TTS, ASR/STT, VAD, voice conversion, speaker diarization, music generation. No Python dependency.
WhisperSpeech/WhisperSpeech
An Open Source text-to-speech system built by inverting Whisper.
Finrandojin/alexandria-audiobook
AI-powered multi-voice audiobook generator — LLM script annotation, voice cloning, voice design, LoRA training, per-line style control, and export to MP3, chaptered M4B, or Audacity multi-track. Built on Qwen3-TTS.
echogarden-project/echogarden
Cross-platform speech toolset, used from the command-line or as a Node.js library. Includes a variety of engines for speech synthesis, speech recognition, forced alignment, speech translation, voice isolation, language detection and more.
AlekPet/ComfyUI_Custom_Nodes_AlekPet
Custom nodes that extend the capabilities of Comfyui