Back to Topics Directory
Topic Hub

#sglang (15 Repositories)

Ranked open-source repositories tagged with #sglang, scored by pull request acceptance likelihood and maintainer engagement velocity.

Topic Avg Merge Rate

49.3%

Avg Review Latency

63.8h

Filter by language

15 repositories tagged #sglang

A TierRust 7.9k 6 GFIs

ai-dynamo/dynamo

A Datacenter Scale Distributed Inference Serving Framework

62.2%
Merge Rate
8h
First Review
48%
1st-Timers
161
Maintainers
A TierPython 1.6k 3 GFIs

intel/auto-round

A SOTA quantization toolkit for high-accuracy low-bit LLM inference, seamlessly optimized for CPU/XPU/CUDA, with multi-datatype support and full compatibility with vLLM, SGLang, and Transformers|简洁且高效的量化工具包

77.3%
Merge Rate
2d
First Review
63%
1st-Timers
19
Maintainers
A TierPython 1.2k

ModelCloud/GPTQModel

LLM model quantization (compression) toolkit with HW acceleration support for Nvidia, AMD, Intel GPU and Intel/AMD/Apple CPU via HF, vLLM, and SGLang.

94.6%
Merge Rate
14d
First Review
78%
1st-Timers
2
Maintainers
A TierC++ 6.3k

kvcache-ai/Mooncake

Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.

70.2%
Merge Rate
20h
First Review
73%
1st-Timers
82
Maintainers
A TierPython 903

Tencent-Hunyuan/UniRL

UniRL is a Framework for Unified Multimodal Model Reinforcement Learning

65.7%
Merge Rate
6h
First Review
67%
1st-Timers
23
Maintainers
A TierPython 1.5k

SemiAnalysisAI/InferenceX

Open Source Continuous Inference Benchmark Research Platform — Kimi K3 2.8T, MiniMax M3, DeepSeekv4, GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72 & soon™ TPUv6e/v7/Trainium2/3 | 开源持续推理基准研究平台 — Kimi K2.7-Code、MiniMax M3、DeepSeekv4、GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72,即将推出™ TPUv6e/v7/Trainium2/3

61.5%
Merge Rate
1d
First Review
47%
1st-Timers
34
Maintainers
A TierPython 946 4 GFIs

sgl-project/sglang-omni

SGLang-Omni is a high-performance serving framework for audio models (TTS, ASR) and unified multimodal models.

55.0%
Merge Rate
13h
First Review
41%
1st-Timers
90
Maintainers
B TierRust 489

smg-project/smg

Engine-agnostic LLM gateway in Rust. Full OpenAI & Anthropic API compatibility across vLLM, TRT-LLM, TokenSpeed, SGLang, OpenAI, Gemini & more. Industry-first gRPC pipeline, KV cache-aware routing, chat history, tokenization caching, Responses API, embeddings, WASM plugins, MCP, and multi-tenant auth.

73.6%
Merge Rate
5d
First Review
48%
1st-Timers
18
Maintainers
B TierGo 498

ome-projects/ome

Open Model Engine (OME) — Kubernetes operator for LLM serving, GPU scheduling, and model lifecycle management. Works with SGLang, vLLM, TensorRT-LLM, and Triton

70.5%
Merge Rate
9d
First Review
47%
1st-Timers
10
Maintainers
B TierTypeScript 1.7k

sybil-solutions/local-studio

Control panel for VLLM, Sglang, llama.cpp, exllamav3

52.6%
Merge Rate
3d
First Review
18%
1st-Timers
8
Maintainers
B TierPython 1.1k

sgl-project/SpecForge

Train speculative decoding models effortlessly and port them smoothly to SGLang serving.

56.6%
Merge Rate
2d
First Review
41%
1st-Timers
47
Maintainers
D TierGo 288

sgl-project/rbg

A workload for deploying LLM inference services on Kubernetes

0.0%
Merge Rate
20h
First Review
0%
1st-Timers
2
Maintainers
D TierPython 1.1k 2 GFIs

ovg-project/kvcached

Virtualized Elastic KV Cache for Dynamic GPU Sharing and Beyond

0.0%
Merge Rate
5h
First Review
0%
1st-Timers
4
Maintainers
D TierGo 309

InftyAI/llmaz

☸️ Easy, advanced inference platform for large language models on Kubernetes. 🌟 Star to support our work!

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 1.1k

OpenMOSS/MOVA

MOVA: Towards Scalable and Synchronized Video–Audio Generation

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers