Thunderdome B2B SaaS AI Perception Index

← LLM & Agent Evals ranking

Brand report · September 2026

Confident AI Confident AI: AI Visibility Report

confident-ai.com ↗

When buyers ask ChatGPT, Claude, and Gemini about llm & agent evals, Confident AI ranks #15 of 15, a Visibility Score of 4 in LLM & Agent Evals .

These rankings are measured from what the three models know from training, not a live web search. The live-web grounded surface is rolling out across the index. Methodology.

How AI sees Confident AI

Dead last in its only category, invisible as a first pick.

Where it wins. Confident AI holds a foothold in LLM & Agent Evals through co-mention with category anchors like Weights & Biases and LangSmith, surfacing in 6.9% of relevant queries.

Where it loses. Rank 15 of 15 with zero first-pick selections tells the real story: buyers hear the name alongside stronger players but never land on it as the answer, and it fails to register across five prompt themes including budget comparisons and use-case fit questions where purchase decisions actually form.

AI-generated analysis of the September 2026 measurements.

LLM & Agent Evals · #15 of 15

4 / 100 first snapshot (September 2026)
7%
mention rate (72 answers)
5
avg. position
0%
first pick
1%
share of voice
First snapshot: September 2026. The trend line appears with the second monthly measurement; no modeled history.
ChatGPT
3
Claude
10
Gemini
0

How the models portray Confident AI

Being named is not the same as being recommended. Each mention is graded endorsed (a strong pick), listed (a neutral option), or caveated (named with a reservation).

60% 40%
endorsed · 95% band 0–43% listed caveated graded across 5 mentions in LLM & Agent Evals answers
Prompt themes Confident AI would want to own

Share of each theme's answers that mention Confident AI. Hover a row for the exact prompt. Compare shapes on the head-to-head page.

Gaps Use-case fit · Top tools in 2026 · Startup & small team · Feature-led ask · Budget & alternatives

Most co-mentioned competitors

Share of Confident AI's mentions where the AI models name this brand in the same answer. That is the real competitive set in the AI channel.

Weights & Biases
100%
LangSmith
100%
Langfuse
60%
Arize AI
60%
Arize Phoenix
60%
Braintrust
60%

Scores are measured from real AI answers, refreshed monthly. Methodology.

Get alerted on big movements

We rerun the index every month. Drop your email and pick what to watch in LLM & Agent Evals: the whole category or specific brands. Free, unsubscribe anytime.

Or watch specific brands: