Back to Topics Directory
Topic Hub

#llama (30 Repositories)

Ranked open-source repositories tagged with #llama, scored by pull request acceptance likelihood and maintainer engagement velocity.

Topic Avg Merge Rate

70.5%

Avg Review Latency

55.5h

Filter by language

30 repositories tagged #llama

B TierPython 188

harleyszhang/lite_llama

A light llama-like llm inference framework based on the triton kernel.

96.0%
Merge Rate
2d
First Review
100%
1st-Timers
0
Maintainers
S TierRust 191

timtoole02/Camelid

Camelid: a Rust-native local inference backend with evidence-gated model compatibility.

86.7%
Merge Rate
23h
First Review
100%
1st-Timers
3
Maintainers
A TierPython 863 5 GFIs

NVIDIA-NeMo/Automodel

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

78.5%
Merge Rate
1d
First Review
71%
1st-Timers
36
Maintainers
A TierPython 5.6k

gpustack/gpustack

A GPU cluster manager for high-performance AI model serving (vLLM, SGLang) and on-demand SSH-accessible GPU instances.

88.5%
Merge Rate
2d
First Review
65%
1st-Timers
25
Maintainers
A TierGo 561

hybridgroup/yzma

Go with your own intelligence - Go applications that directly integrate llama.cpp for local inference using hardware acceleration.

90.9%
Merge Rate
9d
First Review
100%
1st-Timers
5
Maintainers
A TierC++ 1.6k 18 GFIs

tenstorrent/tt-metal

:metal: TT-NN operator library, and TT-Metalium low level kernel programming model.

65.8%
Merge Rate
7h
First Review
58%
1st-Timers
254
Maintainers
B TierC++ 809

invergent-ai/surogate

Training/Fine-tuning at the speed of light

96.7%
Merge Rate
10d
First Review
100%
1st-Timers
0
Maintainers
A TierPython 75.2k 9 GFIs

unslothai/unsloth

Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, FLUX and more.

72.5%
Merge Rate
8h
First Review
46%
1st-Timers
134
Maintainers
A TierC++ 1.8k

ROCm/FastFlowLM

Run LLMs on AMD Ryzen™ AI NPUs in minutes. Just like Ollama - but purpose-built and deeply optimized for the AMD NPUs.

81.6%
Merge Rate
4d
First Review
44%
1st-Timers
4
Maintainers
A TierPython 11.9k

dataelement/bisheng

BISHENG is an open LLM devops platform for next generation Enterprise AI applications. Powerful and comprehensive features include: GenAI workflow, RAG, Agent, Unified model management, Evaluation, SFT, Dataset Management, Enterprise-level System Management, Observability and more.

87.4%
Merge Rate
9d
First Review
73%
1st-Timers
3
Maintainers
B TierJavaScript 2.6k

papersgpt/papersgpt-for-zotero

A powerful Zotero AI and MCP plugin with ChatGPT, Gemini 3.6, Claude Fable 5, Claude Sonnet 5, DeepSeek V4, Grok, OpenRouter, Kimi k3, GLM 5.2, SiliconFlow, GPT-oss, Gemma 4, Qwen 3.7

100.0%
Merge Rate
22h
First Review
100%
1st-Timers
1
Maintainers
A TierKotlin 7.1k 5 GFIs

AAswordman/Operit

The most powerful AI agent and AI chat software on Android/Operit是一款Android上能力最为强大、发展最久的AI Agent

64.0%
Merge Rate
1d
First Review
72%
1st-Timers
23
Maintainers
A TierTypeScript 452

tetherto/qvac

Open-source local AI SDK - run AI on-device with no cloud, no API keys. Supports GGUF, RAG, image, music, and video generation, speech-to-text, P2P inference, and more. Cross-platform: Linux, macOS, Windows, Android, iOS.

71.4%
Merge Rate
1d
First Review
66%
1st-Timers
41
Maintainers
A TierPython 32.9k 12 GFIs

sgl-project/sglang

SGLang is a high-performance serving framework for large language models and multimodal models.

49.1%
Merge Rate
5h
First Review
29%
1st-Timers
593
Maintainers
B TierKotlin 178

ferranpons/Llamatik

True on-device AI for Kotlin Multiplatform (Android, iOS, Desktop, JVM, WASM). LLM, Speech-to-Text and Image Generation — powered by llama.cpp, whisper.cpp and stable-diffusion.cpp.

100.0%
Merge Rate
17h
First Review
100%
1st-Timers
1
Maintainers
A TierC++ 5.4k

lemonade-sdk/lemonade

Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs. Join our discord: https://discord.gg/5xXzkMu8Zk

71.1%
Merge Rate
20h
First Review
67%
1st-Timers
63
Maintainers
A TierPython 15.4k

modelscope/ms-swift

Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, Phi4, ...) (AAAI 2025).

77.9%
Merge Rate
3d
First Review
51%
1st-Timers
61
Maintainers
A TierPython 8.0k

InternLM/lmdeploy

LMDeploy is a toolkit for compressing, deploying, and serving LLMs.

70.3%
Merge Rate
9h
First Review
50%
1st-Timers
34
Maintainers
B TierGo 48.8k

mudler/LocalAI

LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.

82.4%
Merge Rate
5d
First Review
72%
1st-Timers
34
Maintainers
B TierPython 90.5k 25 GFIs

vllm-project/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

40.9%
Merge Rate
10h
First Review
30%
1st-Timers
939
Maintainers
B TierC++ 1.6k

UbiquitousLearning/mllm

Fast Multimodal LLM on Mobile Devices

75.0%
Merge Rate
17h
First Review
50%
1st-Timers
2
Maintainers
B TierPython 15.3k

GaiZhenbiao/ChuanhuChatGPT

GUI for ChatGPT API and many LLMs. Supports agents, file-based QA, GPT finetuning and query with web search. All with a neat UI.

50.0%
Merge Rate
-
First Review
100%
1st-Timers
0
Maintainers
B TierPython 6.6k

linkedin/Liger-Kernel

Efficient Triton Kernels for LLM Training

44.3%
Merge Rate
4h
First Review
47%
1st-Timers
29
Maintainers
B TierCUCuda 1.3k

alibaba/rtp-llm

RTP-LLM: Alibaba's high-performance LLM inference engine for diverse applications.

29.1%
Merge Rate
4h
First Review
36%
1st-Timers
11
Maintainers
B TierGo 179.8k

ollama/ollama

Get up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.

27.7%
Merge Rate
9h
First Review
7%
1st-Timers
275
Maintainers
B TierJava 13.0k

langchain4j/langchain4j

LangChain4j is an idiomatic, open-source Java library for building LLM-powered applications on the JVM. It offers a unified API over popular LLM providers and vector stores, and makes implementing tool calling (including MCP support), agents and RAG easy. It integrates seamlessly with enterprise Java frameworks like Quarkus and Spring Boot.

60.8%
Merge Rate
3d
First Review
56%
1st-Timers
85
Maintainers
B TierGo 5.5k

mostlygeek/llama-swap

Reliable model swapping for any local OpenAI/Anthropic compatible server - llama.cpp, vllm, etc

60.0%
Merge Rate
2d
First Review
50%
1st-Timers
59
Maintainers
B TierPython 4.2k

ModelTC/LightLLM

LightLLM is a Python-based LLM (Large Language Model) inference and serving framework, notable for its lightweight design, easy scalability, and high-speed performance.

62.3%
Merge Rate
6d
First Review
50%
1st-Timers
5
Maintainers
B TierRust 2.2k

floneum/kalosm

Instant, controllable, local pre-trained AI models in Rust

33.3%
Merge Rate
3h
First Review
50%
1st-Timers
1
Maintainers
B TierTypeScript 1.6k

kwaroran/Risuai

Make your own story. User-friendly software for LLM roleplaying

100.0%
Merge Rate
4d
First Review
0%
1st-Timers
1
Maintainers
Best Llama Open Source Repositories & C-Rank™ | GetMerged