AI models and infrastructure

Model evaluation and monitoring

Tools for evaluating models and tracking their usage, quality, or cost

Where it stands

Agent skills13 products1.04× its peers

Fastest products

1google-agents-cli-eval · Agent skills10×
2mo-qa · Agent skills2.1×
3langsmith-online-eval-engineering · Agent skills1.7×
4eval-engineering · Agent skills1.5×
5signal-postmortem · Agent skills1.5×
6ponytail-gain · Agent skills1.3×
7Inspect AI · Code editor extensions1.1×
8Inspect AI · Open-source editor extensions1.0×
9Antigravity Usage Monitor · Open-source editor extensions1.0×
10gb-voicebench-sftjudge · AI apps1.0×
11llm-evaluation · Agent skills1.0×
12Tolokaforge Evaluation · AI apps1.0×
13evaluation-methodology · Agent skills0.9×
14expo-skill-eval · Agent skills0.9×
15langfuse · Agent skills0.9×
16Qwen/Qwen-Image-Bench · Open models0.8×
17agent-platform-eval-flywheel · Agent skills0.8×
18vectara/hallucination_evaluation_model · Open models0.7×
19cargo-analytics · Agent skills0.5×
20skill-doctor · Agent skills0.4×

45 products carry this theme. “× its peers” is the theme's median growth against the place's median.

Measured, not estimated · snapshot 2026-10-03 · Growth is shown only where the starting size is big enough to mean something.