Topic Hub
#ai-evaluation-tools (2 Repositories)
Ranked open-source repositories tagged with #ai-evaluation-tools, scored by pull request acceptance likelihood and maintainer engagement velocity.
Topic Avg Merge Rate
0.0%
Avg Review Latency
0.0h
Filter by language
Rankings/rankings/python
2 repositories tagged #ai-evaluation-tools
C TierPython 114
ianarawjo/evalstats
Statistical analysis for LLM evaluations, from model and prompt comparisons to inference resilient to LLM judge bias, including at small sample sizes. All defaults battle-tested in Monte Carlo simulations.
0.0%
Merge Rate
-
First Review
0%
1st-Timers
1
Maintainers
D TierPython 16.2k
raga-ai-hub/RagaAI-Catalyst
Python SDK for Agent AI Observability, Monitoring and Evaluation Framework. Includes features like agent, llm and tools tracing, debugging multi-agentic system, self-hosted dashboard and advanced analytics with timeline and execution graph view
0.0%
Merge Rate
-
First Review
0%
1st-Timers
0
Maintainers