#amd (25 Repositories)
Ranked open-source repositories tagged with #amd, scored by pull request acceptance likelihood and maintainer engagement velocity.
36.5%
41.5h
25 repositories tagged #amd
uccl-project/uccl
UCCL is an efficient communication library for GPUs, covering collectives, P2P (e.g., KV cache transfer, RL weight transfer), and EP (e.g., GPU-driven)
ROCm/FastFlowLM
Run LLMs on AMD Ryzen™ AI NPUs in minutes. Just like Ollama - but purpose-built and deeply optimized for the AMD NPUs.
lemonade-sdk/lemonade
Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs. Join our discord: https://discord.gg/5xXzkMu8Zk
SemiAnalysisAI/InferenceX
Open Source Continuous Inference Benchmark Research Platform — Kimi K3 2.8T, MiniMax M3, DeepSeekv4, GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72 & soon™ TPUv6e/v7/Trainium2/3 | 开源持续推理基准研究平台 — Kimi K2.7-Code、MiniMax M3、DeepSeekv4、GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72,即将推出™ TPUv6e/v7/Trainium2/3
ROCm/aomp
AOMP is an open source Clang/LLVM based compiler with added support for the OpenMP® API on Radeon™ GPUs. Use this repository for releases, issues, documentation, packaging, and examples.
vllm-project/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
LMCache/LMCache
LMCache: Supercharge Your LLM with the Fastest KV Cache Layer
ChefKissInc/NootedRed
The macOS kext for GPU bug fixes, QoL improvements, and extended hardware support.
ROCm/AMDMIGraphX
AMD's graph optimization engine.
JuliaGPU/AcceleratedKernels.jl
Cross-architecture parallel algorithms for Julia's CPU and GPU backends. Targets multithreaded CPUs, and GPUs via Intel oneAPI, AMD ROCm, Apple Metal, Nvidia CUDA.
zolotukhin/zinc
Zig INferenCe Engine — Local LLM inference on AMD GPUs and Apple Silicon
Genesis-Embodied-AI/quadrants
High-performance multi-platform compiler for physics simulation
HorizonUnix/ZenTune
Universal x86 Tuning Utility for AMD Ryzen CPUs on Linux and macOS based on ZenMaster
dependents/node-dependency-tree
Get the dependency tree of a module
optiscaler/OptiScaler
OptiScaler bridges upscaling/frame gen across GPUs. Supports DLSS2+/XeSS/FSR2+ inputs, replaces native upscalers, enables FSR-FG/XeFG on non-FG titles. Supports Nukem mod for DLSSG-to-FSR3 FG.
wjluoxiao/XB_ToolBox
XB_ToolBox: An easy-to-use ComfyUI custom node suite that streamlines workflow generation. Featuring exclusive data-flow logic and pioneering experimental kernel optimizations dedicated to the native AMD GPU ecosystem.
mayankk2308/set-egpu
Display-agnostic acceleration of macOS applications using external GPUs.
Nathanw1014/strix-halo-llamacpp
Performance-tuned llama.cpp for AMD Strix Halo (gfx1151): FA + MoE-prefill fixes with a bundled current Mesa driver. Vulkan and HIP; portable dir, Docker, and distrobox.
hogeheer499-commits/strix-halo-guide
Strix Halo guide for AMD Ryzen AI MAX+ 395 / Radeon 8060S local LLM setup and benchmarks: Ollama, llama.cpp, Vulkan/RADV, ROCm, GGUF, and raw evidence.
mikeroyal/GPU-Guide
Graphics Processing Unit (GPU) Architecture Guide
Nukem9/dlssg-to-fsr3
Adds AMD FSR 3 Frame Generation to games by replacing Nvidia DLSS Frame Generation (nvngx_dlssg).
ocerman/zenpower
Zenpower is Linux kernel driver for reading temperature, voltage(SVI2), current(SVI2) and power(SVI2) for AMD Zen family CPUs.
dumbie/RadeonTuner
RadeonTuner is an easy to use alternative for the AMD Adrenalin Software for users that just want the basics and use the driver only software install type.
irusanov/SMUDebugTool
A dedicated tool to help write/read various parameters of Ryzen-based systems, such as manual overclock, SMU, PCI, CPUID, MSR and Power Table.
tandasat/barevisor
A bare minimum hypervisor on AMD and Intel processors for learners.