Thunderdome B2B SaaS AI Perception Index

← LLM & Agent Evals ranking

Brand report · September 2026

Weights & Biases Weights & Biases: AI Visibility Report

wandb.ai ↗

⚔ Weights & Biases appears in the Perception Lab: forced-choice runs vs Braintrust →

When buyers ask ChatGPT, Claude, and Gemini about llm & agent evals, Weights & Biases ranks #2 of 15, a Visibility Score of 51 in LLM & Agent Evals .

These rankings are measured from what the three models know from training, not a live web search. The live-web grounded surface is rolling out across the index. Methodology.

How AI sees Weights & Biases

Strong #2 in evals, but first-pick rate barely registers.

Where it wins. Weights & Biases holds the #2 rank in LLM & Agent Evals across 15 competitors, with a 64% mention rate that puts it near the top of how often AI models surface it to buyers in that category.

Where it loses. A 2.8% first-pick rate means models almost never open with Weights & Biases as the go-to evals tool, and the brand has no presence outside this single category to fall back on.

AI-generated analysis of the September 2026 measurements.

LLM & Agent Evals · #2 of 15

51 / 100 first snapshot (September 2026)
64%
mention rate (72 answers)
3
avg. position
3%
first pick
11%
share of voice
First snapshot: September 2026. The trend line appears with the second monthly measurement; no modeled history.
ChatGPT
64
Claude
56
Gemini
32

One of the index's bigger model splits: see where the models disagree →

How the models portray Weights & Biases

Being named is not the same as being recommended. Each mention is graded endorsed (a strong pick), listed (a neutral option), or caveated (named with a reservation).

33% 33% 35%
endorsed · 95% band 21–47% listed caveated graded across 46 mentions in LLM & Agent Evals answers
Prompt themes Weights & Biases would want to own

Share of each theme's answers that mention Weights & Biases. Hover a row for the exact prompt. Compare shapes on the head-to-head page.

If you could pick one
7/9 · #1×1
Use-case fit
2/9 · #1×1
Most co-mentioned competitors

Share of Weights & Biases's mentions where the AI models name this brand in the same answer. That is the real competitive set in the AI channel.

LangSmith
83%
Arize AI
52%
Langfuse
48%
Braintrust
46%
Arize Phoenix
33%
Ragas
26%

Head-to-head: Weights & Biases vs LangSmith

Scores are measured from real AI answers, refreshed monthly. Methodology.

Get alerted on big movements

We rerun the index every month. Drop your email and pick what to watch in LLM & Agent Evals: the whole category or specific brands. Free, unsubscribe anytime.

Or watch specific brands:

Work at Weights & Biases? Put this on your site: a live "AI-recommended" badge, grounded in this data, that updates monthly and links back here. Free. Get the badge →