Thunderdome B2B SaaS AI Perception Index

← LLM & Agent Evals ranking

Brand report · September 2026

Honeyhive Honeyhive: AI Visibility Report

honeyhive.ai ↗

When buyers ask ChatGPT, Claude, and Gemini about llm & agent evals, Honeyhive ranks #14 of 15, a Visibility Score of 5 in LLM & Agent Evals .

These rankings are measured from what the three models know from training, not a live web search. The live-web grounded surface is rolling out across the index. Methodology.

How AI sees Honeyhive

One category, near the bottom, no first picks.

Where it wins. Honeyhive holds a foothold in LLM & Agent Evals, appearing in 9.7% of relevant prompts and landing alongside recognized names like Weights & Biases and Arize AI.

Where it loses. At rank 14 of 15 with zero first-pick selections, it trails the field badly, and the models omit it entirely from broad buying questions, budget comparisons, and use-case-specific asks, leaving it visible only at the margins of the one race it runs.

AI-generated analysis of the September 2026 measurements.

LLM & Agent Evals · #14 of 15

5 / 100 first snapshot (September 2026)
10%
mention rate (72 answers)
6
avg. position
0%
first pick
2%
share of voice
First snapshot: September 2026. The trend line appears with the second monthly measurement; no modeled history.
ChatGPT
0
Claude
6
Gemini
8

How the models portray Honeyhive

Being named is not the same as being recommended. Each mention is graded endorsed (a strong pick), listed (a neutral option), or caveated (named with a reservation).

71% 29%
endorsed · 95% band 0–35% listed caveated graded across 7 mentions in LLM & Agent Evals answers
Prompt themes Honeyhive would want to own

Share of each theme's answers that mention Honeyhive. Hover a row for the exact prompt. Compare shapes on the head-to-head page.

Gaps Best overall · Use-case fit · Startup & small team · If you could pick one · Budget & alternatives

Most co-mentioned competitors

Share of Honeyhive's mentions where the AI models name this brand in the same answer. That is the real competitive set in the AI channel.

Weights & Biases
86%
LangSmith
86%
Arize AI
43%
Langfuse
43%
Datadog
43%
Arize Phoenix
43%

Scores are measured from real AI answers, refreshed monthly. Methodology.

Get alerted on big movements

We rerun the index every month. Drop your email and pick what to watch in LLM & Agent Evals: the whole category or specific brands. Free, unsubscribe anytime.

Or watch specific brands: