Thunderdome B2B SaaS AI Perception Index

← LLM & Agent Evals ranking

Brand report · September 2026

Humanloop Humanloop: AI Visibility Report

humanloop.com ↗

⚔ Humanloop appears in the Perception Lab: forced-choice runs vs Braintrust →

When buyers ask ChatGPT, Claude, and Gemini about llm & agent evals, Humanloop ranks #11 of 15, a Visibility Score of 9 in LLM & Agent Evals .

These rankings are measured from what the three models know from training, not a live web search. The live-web grounded surface is rolling out across the index. Methodology.

How AI sees Humanloop

Humanloop exists in one race and is losing it.

Where it wins. It does register in LLM & Agent Evals, earning a 13.9% mention rate inside a competitive 15-tool field, which at least confirms models know the product exists.

Where it loses. Ranked 11th with zero first-pick selections, it gets passed over whenever a buyer wants a concrete recommendation, and the models have no answer for it on use-case fit or small team questions, which is where purchasing decisions actually get made.

AI-generated analysis of the September 2026 measurements.

LLM & Agent Evals · #11 of 15

9 / 100 first snapshot (September 2026)
14%
mention rate (72 answers)
4
avg. position
0%
first pick
2%
share of voice
First snapshot: September 2026. The trend line appears with the second monthly measurement; no modeled history.
ChatGPT
28
Claude
0
Gemini
0

One of the index's bigger model splits: see where the models disagree →

How the models portray Humanloop

Being named is not the same as being recommended. Each mention is graded endorsed (a strong pick), listed (a neutral option), or caveated (named with a reservation).

40% 40% 20%
endorsed · 95% band 17–69% listed caveated graded across 10 mentions in LLM & Agent Evals answers
Prompt themes Humanloop would want to own

Share of each theme's answers that mention Humanloop. Hover a row for the exact prompt. Compare shapes on the head-to-head page.

Gaps Use-case fit · Startup & small team

Most co-mentioned competitors

Share of Humanloop's mentions where the AI models name this brand in the same answer. That is the real competitive set in the AI channel.

Weights & Biases
90%
Langfuse
80%
LangSmith
70%
Arize AI
70%
Patronus AI
50%
Braintrust
50%

Scores are measured from real AI answers, refreshed monthly. Methodology.

Get alerted on big movements

We rerun the index every month. Drop your email and pick what to watch in LLM & Agent Evals: the whole category or specific brands. Free, unsubscribe anytime.

Or watch specific brands: