#audio-generation (9 Repositories)
Ranked open-source repositories tagged with #audio-generation, scored by pull request acceptance likelihood and maintainer engagement velocity.
21.5%
17.1h
9 repositories tagged #audio-generation
vllm-project/vllm-omni
A framework for efficient model inference with omni-modality models
sgl-project/sglang-omni
SGLang-Omni is a high-performance serving framework for audio models (TTS, ASR) and unified multimodal models.
mudler/LocalAI
LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.
light-and-ray/Minimalistic-Comfy-Wrapper-WebUI
MCWW: Additional non-node based UI for ComfyUI focused on inference. Stable UI states; presets; and advanced queue. Based on Gradio
BinWang28/audio-ai-hub
The hub for audio AI research: papers, open models, benchmarks & datasets across audio LLMs, speech recognition, TTS, music & audio generation.
OpenDCAI/GameFactory-3A
A comprehensive open-source 3A game-generation skill and asset framework.
xiaomi-research/controlfoley
[ACM MM 2026] ControlFoley: Unified and Controllable Video-to-Audio Generation with Cross-Modal Conflict Handling
haoheliu/AudioLDM
AudioLDM: Generate speech, sound effects, music and beyond, with text.
gokayfem/ComfyUI-fal-API
Production-ready ComfyUI custom nodes for 1,400+ fal.ai models, auto-updated image, video, audio, LLM and VLM APIs with native media, caching and cost controls.