Back to Topics Directory
Topic Hub

#rocm (17 Repositories)

Ranked open-source repositories tagged with #rocm, scored by pull request acceptance likelihood and maintainer engagement velocity.

Topic Avg Merge Rate

52.9%

Avg Review Latency

49.0h

Filter by language

17 repositories tagged #rocm

S TierPython 887

unilabsim/UniLab

UniLab: A Heterogeneous Architecture for Robot RL Beyond GPU-Dominant Paradigms

88.2%
Merge Rate
22h
First Review
77%
1st-Timers
4
Maintainers
A TierC++ 398 2 GFIs

QMCPACK/qmcpack

Main repository for QMCPACK, an open-source production level many-body ab initio Quantum Monte Carlo code for computing the electronic structure of atoms, molecules, and solids with full performance portable GPU support

89.9%
Merge Rate
1d
First Review
71%
1st-Timers
12
Maintainers
A TierC++ 217

ROCm/MIVisionX

AMD MIVisionX is a computer vision toolkit built around a highly optimized, conformant open-source implementation of the Khronos OpenVX™ 1.3.2 specification. As of the 4.0.0 release, MIVisionX ships three components: the AMD OpenVX™ engine, the AMD RPP OpenVX extension, and the RunVX graph executor — across CPU, HIP, and OpenCL backends.

93.2%
Merge Rate
21h
First Review
100%
1st-Timers
5
Maintainers
A TierGo 561

hybridgroup/yzma

Go with your own intelligence - Go applications that directly integrate llama.cpp for local inference using hardware acceleration.

90.9%
Merge Rate
9d
First Review
100%
1st-Timers
5
Maintainers
A TierRust 112

tracel-ai/cubek

CubeK: high-performance multi-platform kernels in CubeCL

87.7%
Merge Rate
5d
First Review
60%
1st-Timers
7
Maintainers
A TierJUJulia 342

JuliaGPU/AMDGPU.jl

AMD GPU (ROCm) programming in Julia

85.7%
Merge Rate
3d
First Review
75%
1st-Timers
10
Maintainers
A TierRust 554 4 GFIs

warpfront/hipfire

RDNA-native LLM inference engine in Rust.

70.2%
Merge Rate
6d
First Review
61%
1st-Timers
12
Maintainers
A TierPython 13.7k 1 GFIs

apache/tvm

Open Machine Learning Compiler Framework

80.3%
Merge Rate
3d
First Review
75%
1st-Timers
25
Maintainers
A TierPython 1.5k

SemiAnalysisAI/InferenceX

Open Source Continuous Inference Benchmark Research Platform — Kimi K3 2.8T, MiniMax M3, DeepSeekv4, GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72 & soon™ TPUv6e/v7/Trainium2/3 | 开源持续推理基准研究平台 — Kimi K2.7-Code、MiniMax M3、DeepSeekv4、GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72,即将推出™ TPUv6e/v7/Trainium2/3

61.4%
Merge Rate
1d
First Review
52%
1st-Timers
35
Maintainers
A TierFOFortran 246

ROCm/aomp

AOMP is an open source Clang/LLVM based compiler with added support for the OpenMP® API on Radeon™ GPUs. Use this repository for releases, issues, documentation, packaging, and examples.

29.1%
Merge Rate
13h
First Review
80%
1st-Timers
11
Maintainers
B TierPython 11.4k 2 GFIs

LMCache/LMCache

LMCache: Supercharge Your LLM with the Fastest KV Cache Layer

53.9%
Merge Rate
20h
First Review
43%
1st-Timers
121
Maintainers
B TierC++ 328

ROCm/AMDMIGraphX

AMD's graph optimization engine.

68.8%
Merge Rate
3d
First Review
60%
1st-Timers
22
Maintainers
D TierPython 257

wjluoxiao/XB_ToolBox

XB_ToolBox: An easy-to-use ComfyUI custom node suite that streamlines workflow generation. Featuring exclusive data-flow logic and pioneering experimental kernel optimizations dedicated to the native AMD GPU ecosystem.

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierZig 160

AuleTechnologies/Aule-Attention

High-performance FlashAttention-2 for AMD, Intel, and Apple GPUs. Drop-in replacement for PyTorch SDPA. Triton backend for ROCm (MI300X, RDNA3), Vulkan backend for consumer GPUs. No CUDA required.

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 136

julianmb/q38rocm

Qwen 3.8 27B ROCmFP4 on AMD Strix Halo (Ryzen AI Max+ 395). Up to 36 tok/s via MTP Speculation, TurboQuant & Mesa RADV Wave64.

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 291

hogeheer499-commits/strix-halo-guide

Strix Halo guide for AMD Ryzen AI MAX+ 395 / Radeon 8060S local LLM setup and benchmarks: Ollama, llama.cpp, Vulkan/RADV, ROCm, GGUF, and raw evidence.

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
D TierPython 104

Nathanw1014/strix-halo-llamacpp

Performance-tuned llama.cpp for AMD Strix Halo (gfx1151): FA + MoE-prefill fixes with a bundled current Mesa driver. Vulkan and HIP; portable dir, Docker, and distrobox.

0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers
Best Rocm Open Source Repositories & C-Rank™ | GetMerged