ModelsAgree

AI ranking change · 2026-07-12

Inspect AI overtakes Promptfoo as ChatGPT's #1 pick

for open-source llm eval framework

PromptfooInspect AI

On 2026-07-12, ChatGPT changed its #1 recommendation for best open-source llm eval framework — dropping Promptfoo from the top spot in favor of Inspect AI. The previous #1 had held since 2026-07-12.

Best overall architecture for modern evals: composable tasks, agents, tools, sandboxes, scorers, 200-plus reusable evaluations, strong model-provider support, and excellent logs, web visualization, and VS Code toolingChatGPT

Is your product in this race?

open-source LLM eval framework rankings re-poll every week. Check where the AI models place your product — and get an email the moment it moves.

Get your AI Visibility Grade →

Source: modelsagree.com · CC BY 4.0 · Every poll is public and re-checked continuously.