⚔ Helicone appears in the Perception Lab: forced-choice runs vs Braintrust →
When buyers ask ChatGPT, Claude, and Gemini about llm & agent evals, Helicone ranks #10 of 15, a Visibility Score of 13 in LLM & Agent Evals .
These rankings are measured from what the three models know from training, not a live web search. The live-web grounded surface is rolling out across the index. Methodology.
One category, one shot, and it's landing outside the top half.
Where it wins. Helicone earns a 23.6% mention rate in LLM & Agent Evals, which means nearly one in four relevant queries surfaces it at all. That reach, inside a 15-tool field, is its entire footprint.
Where it loses. A rank of 10 and zero first picks tells you the models treat Helicone as list filler rather than a go-to in evals. It also drops out entirely when buyers frame the question around startups or ask for a single recommendation, exactly the prompts that convert intent into a vendor decision.
AI-generated analysis of the September 2026 measurements.
Being named is not the same as being recommended. Each mention is graded endorsed (a strong pick), listed (a neutral option), or caveated (named with a reservation).
Share of each theme's answers that mention Helicone. Hover a row for the exact prompt. Compare shapes on the head-to-head page.
Gaps Startup & small team · If you could pick one
Share of Helicone's mentions where the AI models name this brand in the same answer. That is the real competitive set in the AI channel.
Scores are measured from real AI answers, refreshed monthly. Methodology.
We rerun the index every month. Drop your email and pick what to watch in LLM & Agent Evals: the whole category or specific brands. Free, unsubscribe anytime.