Back to Topics Directory
Topic Hub

#llm-eval (3 Repositories)

Ranked open-source repositories tagged with #llm-eval, scored by pull request acceptance likelihood and maintainer engagement velocity.

Topic Avg Merge Rate

71.8%

Avg Review Latency

59.9h

Filter by language

3 repositories tagged #llm-eval

B TierTypeScript 24.4k 1 GFIs

promptfoo/promptfoo

Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI and Anthropic.

72.8%
Merge Rate
3d
First Review
53%
1st-Timers
39
Maintainers
B TierPython 11.2k

Arize-ai/phoenix

AI Observability & Evaluation

80.6%
Merge Rate
3d
First Review
51%
1st-Timers
40
Maintainers
B TierPython 5.8k 1 GFIs

Giskard-AI/giskard-oss

🐢 Open-Source Evaluation & Testing library for LLM Agents

62.0%
Merge Rate
2d
First Review
27%
1st-Timers
28
Maintainers
Best Llm-eval Open Source Repositories & C-Rank™ | GetMerged