⚔ Langfuse appears in the Perception Lab: forced-choice runs vs Braintrust →
When buyers ask ChatGPT, Claude, and Gemini about llm & agent evals, Langfuse ranks #4 of 15, a Visibility Score of 35 in LLM & Agent Evals .
These rankings are measured from what the three models know from training, not a live web search. The live-web grounded surface is rolling out across the index. Methodology.
Fourth in a 15-horse eval race, nowhere else.
Where it wins. Langfuse gets named in 43% of LLM and agent eval queries, a real foothold in a crowded 15-tool field. The 'Enterprise pick' prompt theme is doing work, signaling that at least one buyer segment hears a credible story.
Where it loses. A first-pick rate of under 10% means the three tools ranked above it are absorbing most of the purchase intent. The complete absence from startup and small-team themes cuts off a natural adoption funnel that rivals like LangSmith are almost certainly capturing.
AI-generated analysis of the September 2026 measurements.
One of the index's bigger model splits: see where the models disagree →
Being named is not the same as being recommended. Each mention is graded endorsed (a strong pick), listed (a neutral option), or caveated (named with a reservation).
Share of each theme's answers that mention Langfuse. Hover a row for the exact prompt. Compare shapes on the head-to-head page.
Gaps Startup & small team
Share of Langfuse's mentions where the AI models name this brand in the same answer. That is the real competitive set in the AI channel.
Scores are measured from real AI answers, refreshed monthly. Methodology.
We rerun the index every month. Drop your email and pick what to watch in LLM & Agent Evals: the whole category or specific brands. Free, unsubscribe anytime.