Back to Topics Directory
Topic Hub

#cuda (30 Repositories)

Ranked open-source repositories tagged with #cuda, scored by pull request acceptance likelihood and maintainer engagement velocity.

Topic Avg Merge Rate

88.4%

Avg Review Latency

44.1h

Filter by language

30 repositories tagged #cuda

B TierC++ 128

glotzerlab/fresnel

Publication quality path tracing in real time.

100.0%
Merge Rate
<1h
First Review
100%
1st-Timers
1
Maintainers
B TierNINix 153

utensils/comfyui-nix

A slightly opinionated Nix flake for ComfyUI with curated custom nodes. Supports macOS (Apple Silicon) and Linux with CUDA.

100.0%
Merge Rate
6h
First Review
100%
1st-Timers
1
Maintainers
S TierCUCuda 827

brucefan1983/GPUMD

Graphics Processing Units Molecular Dynamics

92.9%
Merge Rate
19h
First Review
100%
1st-Timers
6
Maintainers
B TierOCOCaml 146

mathiasbourgoin/Sarek

SIMT Abstractions for Runtime Extensible Kernels (GPGPU programing with OCaml)

96.5%
Merge Rate
4h
First Review
100%
1st-Timers
1
Maintainers
B TierC 1.7k

Zaneham/Booth

Open-source CUDA, Triton and HIP compiler targeting multiple GPU and CPU architectures.

94.7%
Merge Rate
1d
First Review
100%
1st-Timers
0
Maintainers
S TierPython 3.4k

roflcoopter/viseron

Self-hosted, local only NVR and AI Computer Vision software. With features such as object detection, motion detection, face recognition and more, it gives you the power to keep an eye on your home, office or any other place you want to monitor.

84.4%
Merge Rate
2h
First Review
50%
1st-Timers
4
Maintainers
S TierPython 100

Blackwellboy/model-serving-minefield

Community registry of LLM serving-path traps that produce confidently wrong measurements: templates, tool parsers, reasoning fields, quant kernel paths, CUDA toolchains, KV allocation, eval harnesses, versioning. Symptom-first, with the check that catches each.

90.0%
Merge Rate
5h
First Review
100%
1st-Timers
2
Maintainers
S TierPython 887

unilabsim/UniLab

UniLab: A Heterogeneous Architecture for Robot RL Beyond GPU-Dominant Paradigms

88.2%
Merge Rate
22h
First Review
77%
1st-Timers
4
Maintainers
A TierC++ 398 2 GFIs

QMCPACK/qmcpack

Main repository for QMCPACK, an open-source production level many-body ab initio Quantum Monte Carlo code for computing the electronic structure of atoms, molecules, and solids with full performance portable GPU support

89.9%
Merge Rate
1d
First Review
71%
1st-Timers
12
Maintainers
A TierRust 160

ultralytics/inference

High-performance Ultralytics YOLO inference in Rust with ONNX Runtime, GPU backends, CLI, and WebGPU/WASM.

94.7%
Merge Rate
2d
First Review
75%
1st-Timers
2
Maintainers
A TierPython 3.1k

NVIDIA/skills

Agent Skills for NVIDIA products — install into Claude Code, Codex, and other coding agents to run Physical AI, robotics, simulation, CUDA, and RAG workflows end to end.

86.8%
Merge Rate
7h
First Review
83%
1st-Timers
26
Maintainers
A TierC++ 3.6k 3 GFIs

MrNeRF/LichtFeld-Studio

Train, inspect, edit, automate, and export 3D Gaussian Splatting scenes from a single native application.

80.5%
Merge Rate
17h
First Review
79%
1st-Timers
30
Maintainers
A TierC++ 1.0k

LuisaGroup/LuisaCompute

High-Performance Rendering Framework on Stream Architectures

80.6%
Merge Rate
12h
First Review
88%
1st-Timers
2
Maintainers
A TierC++ 2.5k 8 GFIs

NVIDIA/cccl

CUDA Core Compute Libraries

78.9%
Merge Rate
19h
First Review
60%
1st-Timers
54
Maintainers
A TierPython 5.6k

gpustack/gpustack

A GPU cluster manager for high-performance AI model serving (vLLM, SGLang) and on-demand SSH-accessible GPU instances.

88.5%
Merge Rate
2d
First Review
65%
1st-Timers
25
Maintainers
A TierC++ 780

luigifcruz/CyberEther

High-performance GPU-accelerated signal processing and visualization framework that runs anywhere.

95.2%
Merge Rate
4d
First Review
100%
1st-Timers
6
Maintainers
A TierPython 169

inclusionAI/Awex

A high-performance RL training-inference weight synchronization framework, designed to enable second-level parameter updates from training to inference in RL workflows

100.0%
Merge Rate
1h
First Review
100%
1st-Timers
2
Maintainers
B TierRust 328

avifenesh/memra

Rust + CUDA inference engine for NVIDIA RTX PRO 6000 Blackwell and RTX 5090. Serves safetensors and GGUF over an OpenAI-compatible API, with per-device tuned defaults and speculative decode gated byte-identical to plain decode. Hosted instance: inference.tiyuvta.ai

73.1%
Merge Rate
2h
First Review
100%
1st-Timers
0
Maintainers
A TierGo 561

hybridgroup/yzma

Go with your own intelligence - Go applications that directly integrate llama.cpp for local inference using hardware acceleration.

90.9%
Merge Rate
9d
First Review
100%
1st-Timers
5
Maintainers
A TierRust 112

tracel-ai/cubek

CubeK: high-performance multi-platform kernels in CubeCL

87.7%
Merge Rate
5d
First Review
60%
1st-Timers
7
Maintainers
B TierC++ 809

invergent-ai/surogate

Training/Fine-tuning at the speed of light

96.7%
Merge Rate
10d
First Review
100%
1st-Timers
0
Maintainers
A TierPython 11.5k

debpalash/VoiceStudio

VoiceStudio is the open-source, fully-local ElevenLabs alternative — voice cloning, voice design, video dubbing, dictation, transcription & audiobook creation in 646 languages.

91.1%
Merge Rate
2d
First Review
100%
1st-Timers
4
Maintainers
A TierC++ 1.5k

uccl-project/uccl

UCCL is an efficient communication library for GPUs, covering collectives, P2P (e.g., KV cache transfer, RL weight transfer), and EP (e.g., GPU-driven)

91.0%
Merge Rate
7d
First Review
86%
1st-Timers
12
Maintainers
A TierC++ 296

llnl/blt

A streamlined CMake build system foundation for developing HPC software

100.0%
Merge Rate
20h
First Review
100%
1st-Timers
2
Maintainers
A TierC++ 104

celeritas-project/celeritas

Celeritas is a new Monte Carlo transport code designed to accelerate scientific discovery in high energy physics by improving detector simulation throughput and energy efficiency using GPUs.

85.6%
Merge Rate
1d
First Review
83%
1st-Timers
15
Maintainers
B TierRust 156

mlx-node/mlx-node

96.7%
Merge Rate
4d
First Review
100%
1st-Timers
0
Maintainers
B TierC++ 281

crazyguitar/cppcheatsheet

C/C++ Cheat Sheet

100.0%
Merge Rate
-
First Review
100%
1st-Timers
0
Maintainers
A TierCUCuda 2.2k

rapidsai/cugraph

cuGraph - RAPIDS Graph Analytics Library

70.0%
Merge Rate
2h
First Review
40%
1st-Timers
9
Maintainers
A TierC++ 5.6k

shader-slang/slang

Making it easier to work with shaders

77.3%
Merge Rate
15h
First Review
55%
1st-Timers
29
Maintainers
A TierPython 32.9k 12 GFIs

sgl-project/sglang

SGLang is a high-performance serving framework for large language models and multimodal models.

49.1%
Merge Rate
5h
First Review
29%
1st-Timers
593
Maintainers
Best Cuda Open Source Repositories & C-Rank™ | GetMerged