Topic Hub
#llm-eval (3 Repositories)
Ranked open-source repositories tagged with #llm-eval, scored by pull request acceptance likelihood and maintainer engagement velocity.
Topic Avg Merge Rate
71.8%
Avg Review Latency
59.9h
Filter by language
3 repositories tagged #llm-eval
B TierTypeScript 24.4k 1 GFIs
promptfoo/promptfoo
Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI and Anthropic.
72.8%
Merge Rate
3d
First Review
53%
1st-Timers
39
Maintainers
B TierPython 11.2k
Arize-ai/phoenix
AI Observability & Evaluation
80.6%
Merge Rate
3d
First Review
51%
1st-Timers
40
Maintainers
B TierPython 5.8k 1 GFIs
Giskard-AI/giskard-oss
🐢 Open-Source Evaluation & Testing library for LLM Agents
62.0%
Merge Rate
2d
First Review
27%
1st-Timers
28
Maintainers