Thunderdome B2B SaaS AI Perception Index

← LLM & Agent Evals ranking

Brand report · September 2026

Arize AI Arize AI: AI Visibility Report

arize.com ↗

When buyers ask ChatGPT, Claude, and Gemini about llm & agent evals, Arize AI ranks #5 of 15, a Visibility Score of 30 in LLM & Agent Evals .

These rankings are measured from what the three models know from training, not a live web search. The live-web grounded surface is rolling out across the index. Methodology.

How AI sees Arize AI

One field, one race, sitting fifth with room to climb.

Where it wins. In LLM & Agent Evals, Arize AI earns a 36% mention rate, meaning models surface it in more than a third of relevant buying conversations. The 'Enterprise pick' prompt theme confirms it registers as a credible option when buyers frame the question around scale.

Where it loses. A 12.5% first-pick rate against 14 rivals means it trails whatever sits above it in the same category, and the complete absence from startup and small-team conversations hands those buyers to Langfuse and LangSmith before Arize enters the frame.

AI-generated analysis of the September 2026 measurements.

LLM & Agent Evals · #5 of 15

30 / 100 first snapshot (September 2026)
36%
mention rate (72 answers)
3
avg. position
13%
first pick
6%
share of voice
First snapshot: September 2026. The trend line appears with the second monthly measurement; no modeled history.
ChatGPT
38
Claude
25
Gemini
27

How the models portray Arize AI

Being named is not the same as being recommended. Each mention is graded endorsed (a strong pick), listed (a neutral option), or caveated (named with a reservation).

58% 27% 15%
endorsed · 95% band 39–74% listed caveated graded across 26 mentions in LLM & Agent Evals answers
Prompt themes Arize AI would want to own

Share of each theme's answers that mention Arize AI. Hover a row for the exact prompt. Compare shapes on the head-to-head page.

Enterprise pick
9/9 · #1×9

Gaps Startup & small team

Most co-mentioned competitors

Share of Arize AI's mentions where the AI models name this brand in the same answer. That is the real competitive set in the AI channel.

Weights & Biases
92%
LangSmith
81%
Langfuse
62%
Helicone
42%
Braintrust
38%
Humanloop
27%

Scores are measured from real AI answers, refreshed monthly. Methodology.

Get alerted on big movements

We rerun the index every month. Drop your email and pick what to watch in LLM & Agent Evals: the whole category or specific brands. Free, unsubscribe anytime.

Or watch specific brands:

Work at Arize AI? Put this on your site: a live "AI-recommended" badge, grounded in this data, that updates monthly and links back here. Free. Get the badge →