⚔ LangSmith appears in the Perception Lab: forced-choice runs vs Braintrust →
When buyers ask ChatGPT, Claude, and Gemini about llm & agent evals, LangSmith ranks #1 of 15, a Visibility Score of 77 in LLM & Agent Evals .
These rankings are measured from what the three models know from training, not a live web search. The live-web grounded surface is rolling out across the index. Methodology.
LangSmith owns LLM evals outright, full stop.
Where it wins. In LLM and Agent Evals, LangSmith converts an 85% mention rate into a 46% first-pick rate, meaning nearly half of all AI recommendations name it before any competitor. That combination of reach and primacy is rare at this stage of a category.
Where it loses. Co-mention patterns show Weights and Biases, Braintrust, and Arize Phoenix clustering around LangSmith in the same answers, which signals that buyers are still comparison-shopping rather than defaulting to LangSmith as the sole choice.
AI-generated analysis of the September 2026 measurements.
One of the index's bigger model splits: see where the models disagree →
Being named is not the same as being recommended. Each mention is graded endorsed (a strong pick), listed (a neutral option), or caveated (named with a reservation).
Share of each theme's answers that mention LangSmith. Hover a row for the exact prompt. Compare shapes on the head-to-head page.
Share of LangSmith's mentions where the AI models name this brand in the same answer. That is the real competitive set in the AI channel.
Head-to-head: LangSmith vs Weights & Biases
Scores are measured from real AI answers, refreshed monthly. Methodology.
We rerun the index every month. Drop your email and pick what to watch in LLM & Agent Evals: the whole category or specific brands. Free, unsubscribe anytime.