Buyer prompt · published evidence

Create a reproducible AI customer agent benchmark covering resolution, accuracy, policy compliance, actions, escalation, latency, and customer effort.

An exact prompt-level view of which reviewed brands surfaced and which public URLs the configured AI models cited. No answer text or tenant data is published.

Intent
commercial
Model providers
3
Categories
1
Last observed
Sep 1, 2026, 12:00 AM UTC
Competitive outcome

Brands surfaced

Ordered by provider breadth, then total mentions and owned-domain citations. A mention is not an endorsement or a position claim.

#BrandCategoryProvidersMentionsOwned citationsProvider evidence
01Zendesk AI Agentszendesk.comService agents101Openai
Model comparison

Answer matrix

One row per configured provider and category snapshot. Hashes prove answer identity without publishing stored answer text.

ProviderModel versionCategoryBrands surfacedCitationsUnique domainsAnswer hash
Openaiopenai/gpt-4o-miniService agentsZendesk AI Agents775c5f795e…edc85
Anthropicanthropic/claude-haiku-4-5Service agentsNone observed002f3ea886…1794f
Geminigoogle/gemini-2.5-flashService agentsNone observed22fb155c3d…a216c
URL-derived citations

Sources cited

Hostnames are derived from the returned public URLs. Best position is the smallest citation position observed.

Snapshot IDs327bfab8-d298-4c5f-8adc-58f406c1f295
Corpus versionsai-customer-agent-platforms-2026-07-27
Benchmark versions4-ai-customer-agent-platforms-2026-07-27-models-64232fd32583

Coverage: Service agents. Full answers stored: no · Full answers published: no · Tenant data included: no.