#text-to-speech (30 Repositories)
Ranked open-source repositories tagged with #text-to-speech, scored by pull request acceptance likelihood and maintainer engagement velocity.
54.3%
98.5h
30 repositories tagged #text-to-speech
MartinDelophy/ai-video-editor
Open-source, local-first video editor where creators and AI agents edit the same real timeline.
heygen-com/heygen-cli
Create AI videos from the terminal. Official CLI for the HeyGen video generation API.
PowerBeef/Vocello
Vocello: a local, private voice studio for Apple Silicon. Write a script, pick or describe a voice, and generate speech on-device, faster than realtime on an 8 GB M2 Mac mini. Native Swift + MLX, no Python. Mac app out now, iPhone beta on TestFlight. (Formerly QwenVoice.)
pnnbao97/VieNeu-TTS
Vietnamese TTS with instant voice cloning • On-device • Real-time CPU inference • 24kHz audio quality • Chuyển văn bản thành giọng nói tiếng Việt • Text to speech tiếng Việt • TTS tiếng Việt
kadirnar/voicehub
VoiceHub: A Unified Inference Interface for TTS Models
software-mansion/react-native-executorch
Declarative way to run AI models in React Native on device, powered by ExecuTorch.
ayutaz/piper-plus
Multilingual neural TTS (6 languages: JA/EN/ZH/ES/FR/PT, code supports SV) — C++, C#, Rust, Go, Python, npm (WASM). VITS + Prosody, streaming, CUDA/CoreML/DirectML. pip install piper-plus | npm install piper-plus | cargo install piper-plus-cli
unslothai/unsloth
Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, FLUX and more.
remsky/Kokoro-FastAPI
Dockerized OpenAI-compatible wrapper for Kokoro-82M text-to-speech w/multiplatform CPU, AMD, NVIDIA GPU PyTorch; multi-speaker, voice-mixing, auto-stitching, caption timestamps, SSML, readalong web UI
izwi-ai/izwi
Voice AI runtime. Local first transcription, speaker diarization, TTS, and voice cloning with an OpenAI compatible API.
rapidaai/voice-ai
Rapida is an open-source, end-to-end voice AI orchestration platform for building real-time conversational voice agents with audio streaming, STT, TTS, VAD, multi-channel integration, agent state management, and observability.
RWKV-APP/RWKV_APP
Cross-platform, local-first RWKV chat built with Flutter for Android, iOS, Windows, macOS, and Linux.
harry0703/MoneyPrinterTurbo
利用 AI 大模型和自动化工作流,根据主题或关键词一键生成高清短视频。Generate HD short videos from a topic or keyword with an automated AI workflow.
mitchib1440/SpeakThat
The world's most comprehensive notification reader.
leon-ai/leon
🧠 Leon is your open-source personal assistant.
KoljaB/RealtimeTTS
Converts text to speech in realtime
asterics/Asterics-AAC
Free, easy-to-use AAC app with offline support, flexible input options, media & smart home access
lukaszliniewicz/Pandrator
Turn PDFs and EPUBs into audiobooks; subtitles or videos into dubbed videos (including translation), and more. For free. Pandrator uses local models, including voice-cloning (instant, RVC-enhanced, XTTS fine-tuning) and LLM processing. It aspires to be a user-friendly app with a GUI, an installer and all-in-one packages.
waybarrios/vllm-mlx
High-performance OpenAI and Anthropic compatible LLM inference server for Apple Silicon. Native MLX, continuous batching, multimodal models, MCP tool calling, and Claude Code support.
Blaizzy/mlx-audio-swift
A modular Swift SDK for audio processing with MLX on Apple Silicon
RVC-Boss/GPT-SoVITS
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
yuga-hashimoto/openclaw-assistant
OpenClaw voice assistant app for Android - Wake word activation & system assistant integration
promptslab/Awesome-Prompt-Engineering
This repository contains a hand-curated resources for Prompt Engineering with a focus on Generative Pre-trained Transformer (GPT), ChatGPT, PaLM etc
MoonInTheRiver/DiffSinger
DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism (SVS & TTS); AAAI 2022; Official code
mlalma/kokoro-ios
Kokoro TTS for iOS and macOSX
MatteoFasulo/Whisper-TikTok
From AI tools to TikTok video creation using FFMPEG, Microsoft Edge read aloud and OpenAI Whisper model
keonlee9420/DiffGAN-TTS
PyTorch Implementation of DiffGAN-TTS: High-Fidelity and Efficient Text-to-Speech with Denoising Diffusion GANs
2noise/ChatTTS
A generative speech model for daily dialogue.
daniilrobnikov/vits2
VITS2: Improving Quality and Efficiency of Single-Stage Text-to-Speech with Adversarial Learning and Architecture Design
lucasnewman/f5-tts-mlx
Implementation of F5-TTS in MLX