ModelsAgree

AI ranking change · 2026-07-13

Promptfoo overtakes LangSmith as Gemini's #1 pick

for llm evaluation tool

LangSmithPromptfoo

On 2026-07-13, Gemini changed its #1 recommendation for best llm evaluation tool — dropping LangSmith from the top spot in favor of Promptfoo. The previous #1 had held since 2026-07-12.

Outstanding developer experience for offline regression testing and prompt iteration. Its lightweight, CLI-first, YAML-driven design requires no heavy Python dependencies and integrates natively into CI/CD pipelines. It is the absolute gold standard for automated red-teaming, safety checks, and jailbreak testing across multiple models.Gemini

Is your product in this race?

LLM evaluation tool rankings re-poll every week. Check where the AI models place your product — and get an email the moment it moves.

Get your AI Visibility Grade →

Source: modelsagree.com · CC BY 4.0 · Every poll is public and re-checked continuously.