#benchmark (30 Repositories)
Ranked open-source repositories tagged with #benchmark, scored by pull request acceptance likelihood and maintainer engagement velocity.
82.5%
95.0h
30 repositories tagged #benchmark
martinus/nanobench
Simple, fast, accurate single-header microbenchmarking functionality for C++11/14/17/20
minghinmatthewlam/openbench
Same model, different wrapper: a from-scratch benchmark comparing coding-agent harnesses (codex, pi, opencode, cursor, devin) and open models on correctness, speed, and token cost
jsdelivr/globalping
A global network of probes to run network tests like ping, traceroute and DNS resolve
Ammaar-Alam/minebench
Minecraft-style voxel benchmark for comparing AI models (Arena + Sandbox)
JuliaCI/BenchmarkTools.jl
A benchmarking framework for the Julia language
ClickHouse/ClickBench
ClickBench: a Benchmark For Analytical Databases
pawurb/hotpath-rs
Quickly find bottlenecks in Rust - one profiler for CPU, memory, SQL, HTTP, I/O and async code.
1a1a11a/libCacheSim
a high performance library for building cache simulators
StanfordVL/BEHAVIOR-1K
BEHAVIOR-1K: a platform for accelerating Embodied AI research. Join our Discord for support: https://discord.gg/bccR5vGFEx
Tantalor93/dnspyre
CLI tool for a high QPS DNS benchmark
pnpm/benchmarks
Benchmarks of JavaScript Package Managers
SemiAnalysisAI/InferenceX
Open Source Continuous Inference Benchmark Research Platform — Kimi K3 2.8T, MiniMax M3, DeepSeekv4, GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72 & soon™ TPUv6e/v7/Trainium2/3 | 开源持续推理基准研究平台 — Kimi K2.7-Code、MiniMax M3、DeepSeekv4、GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72,即将推出™ TPUv6e/v7/Trainium2/3
gungraun/gungraun
High-precision, one-shot and consistent benchmarking framework/harness for Rust. All Valgrind tools at your fingertips.
benchflow-ai/benchflow
Research infra for creating RL environments, post-training, and evals.
weavebench/WeaveBench
[EMNLP 2026] WeaveBench: A Long-Horizon, Real-World Benchmark for Computer-Use Agents with Hybrid Interfaces
NVIDIA/SkillEvaluator
Multi-tier framework for evaluating AI agent skills with quality gates, semantic overlap detection, synthetic evaluation dataset generation, and live agent evaluation that measures how skills affect agent behavior.
TIGER-AI-Lab/ClawBench
Open-source benchmark for browser AI agents on daily tasks.
the-benchmarker/web-frameworks
Which is the fastest web framework?
NeuraLegion/brokencrystals
A Broken Application - Very Vulnerable!
sbt/sbt-jmh
"Trust no one, bench everything." - sbt plugin for JMH (Java Microbenchmark Harness)
CodSpeedHQ/codspeed
CodSpeed is the all-in-one performance testing toolkit. Optimize code performance and catch regressions early.
EvolvingLMMs-Lab/lmms-eval
One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks
hogeheer499-commits/strix-halo-guide
Strix Halo guide for AMD Ryzen AI MAX+ 395 / Radeon 8060S local LLM setup and benchmarks: Ollama, llama.cpp, Vulkan/RADV, ROCm, GGUF, and raw evidence.
OpenBMB/UltraEval-Audio
Your faithful, impartial partner for audio evaluation — know yourself, know your rivals. 真实评测,知己知彼。A unified benchmark framework for ASR/TTS/Audio Codec/audio LLM evaluation
lissy93/framework-benchmarks
🌈 The same app built in 10 different frontend frameworks. For automated performance benchmarking
Kotlin/kotlinx-benchmark
Kotlin multiplatform benchmarking toolkit
sosy-lab/benchexec
BenchExec: A Framework for Reliable Benchmarking and Resource Measurement
RafaelGSS/bench-node
A powerful Node.js benchmark library
oneclickvirt/ecs
VPS Fusion Monster Server Test GO Version Aiming to be the most comprehensive server testing project, implemented in Go with zero environment dependencies. VPS融合怪服务器测评项目 GO版本 尽量成为最全能的服务器测评项目,使用 Go 实现,无需任何环境依赖。
rstackjs/build-tools-performance
Benchmarks for bundlers and build tools, including Rspack, Rsbuild, webpack, Vite, Rolldown, esbuild, Parcel, Farm and Utoo.