Thunderdome B2B SaaS AI Perception Index

← LLM & Agent Evals ranking

Brand report · September 2026

Braintrust Braintrust: AI Visibility Report

braintrust.dev ↗

The Perception Lab measured Sep 25, 2026
92%of aided recommendations are qualified, not a flat yes
70%of answers still describe Braintrust as prompt-testing and evals tool
83%forced-choice win rate vs 10 named competitors
50%of answer citations are competitor-authored pages
The full lab: aided verdicts, the forced-choice matrix, and whose content grounds the answers →

When buyers ask ChatGPT, Claude, and Gemini about llm & agent evals, Braintrust ranks #3 of 15, a Visibility Score of 36 in LLM & Agent Evals .

These rankings are measured from what the three models know from training, not a live web search. The live-web grounded surface is rolling out across the index. Methodology.

How AI sees Braintrust

Solid eval-tool contender, but first-pick rate tells the real story.

Where it wins. Braintrust lands rank 3 of 15 in LLM & Agent Evals with a 44% mention rate, meaning AI models surface it reliably when buyers ask about evaluation tooling. Its strongest foothold is startup and small-team prompts, where it competes visibly against the category leaders.

Where it loses. A 15% first-pick rate means buyers hear the name but typically hear LangSmith or Weights & Biases first, and Braintrust's entire AI presence rests on this single category with no adjacent territory to fall back on.

AI-generated analysis of the September 2026 measurements.

LLM & Agent Evals · #3 of 15

36 / 100 first snapshot (September 2026)
44%
mention rate (72 answers)
3
avg. position
15%
first pick
8%
share of voice
First snapshot: September 2026. The trend line appears with the second monthly measurement; no modeled history.
ChatGPT
32
Claude
66
Gemini
9

One of the index's bigger model splits: see where the models disagree →

How the models portray Braintrust

Being named is not the same as being recommended. Each mention is graded endorsed (a strong pick), listed (a neutral option), or caveated (named with a reservation).

56% 31% 13%
endorsed · 95% band 39–72% listed caveated graded across 32 mentions in LLM & Agent Evals answers
Prompt themes Braintrust would want to own

Share of each theme's answers that mention Braintrust. Hover a row for the exact prompt. Compare shapes on the head-to-head page.

Startup & small team
8/9 · #1×3
Use-case fit
5/9 · #1×3
If you could pick one
4/9 · #1×2
Best overall
3/9 · #1×2
Budget & alternatives
1/9 · #1×1
Most co-mentioned competitors

Share of Braintrust's mentions where the AI models name this brand in the same answer. That is the real competitive set in the AI channel.

LangSmith
88%
Weights & Biases
66%
Promptfoo
41%
Ragas
41%
Langfuse
34%
Arize AI
31%

Scores are measured from real AI answers, refreshed monthly. Methodology.

Get alerted on big movements

We rerun the index every month. Drop your email and pick what to watch in LLM & Agent Evals: the whole category or specific brands. Free, unsubscribe anytime.

Or watch specific brands:

Work at Braintrust? Put this on your site: a live "AI-recommended" badge, grounded in this data, that updates monthly and links back here. Free. Get the badge →