Stunt Double Index
The independent ranking of how real AI agents experience the world's websites across discovery, information retrieval, task completion, delegated access and more.
- index.stuntdouble.io
- AudienceDeveloper
- MarketUS
- Localeen-US
- Rank
- Top 51%Prev: #267 · top 51% of 518 tracked
- #267
- SaaS rank
- Sector average 71.3
- #94 of 216
- Benchmark runs
- 3 providersAcross 3 providers
- 7
- SaaS average 71.3
- +6.0
- Market average 72.5
- +4.8
Category breakdown
Strongest is Accuracy; weakest is Market ranking.
Category breakdown
Strongest is Accuracy; weakest is Market ranking.
Score over time
Every run against the SaaS and market medians.
▲ 1.8
Score over time
Every run against the SaaS and market medians.
2 runs since 29 Sept 2026, ▲ 1.8 overall.
- This site
- SaaS median
- Market median
Each category over time
- Brand awareness96▲ 8.2
- Discovery67▼ 1.4
- Information retrieval75▼ 1.2
- Market ranking530.0
- Accuracy97▲ 8.4
- Task completion63▼ 2.0
- Delegated access89▼ 0.2
- Contact & communication67▼ 3.9
Every run, as a table
| Run | This site | SaaS median | Market median |
|---|---|---|---|
| 30 Sept 2026 | 77.3 | 75.8 | 76.3 |
| 29 Sept 2026 | 75.5 | 75.8 | 76.3 |
Medians are by month: each domain counts once, at its newest run that month. Dates are UTC. Latest run 30 Sept 2026.
Agents
ChatGPT Agent does best. Each provider's model and every session behind its score.
70.5
Agents
ChatGPT Agent does best. Each provider's model and every session behind its score.
How these are scored
The score beside each provider is its session quality over a rolling 30 days, from at least 10 sessions per provider.
Scores run 0 to 100. Each is one agent on one task, graded against that task’s fixed rubric, so no agent marks its own work. The median under a score is the sector median: the middle score for the same category, or the same provider, across the other SaaS sites the Index has sent agents to.
- Excellent85+
- Good72-84
- OK58-71
- Poor40-57
- Critical<40
Latest run 30 Sept 2026 (UTC). Anyone signed in can re-run it from 1 Oct 2026, 00:20 UTC. The owner can re-run it any time from their dashboard.
Claudeclaude-sonnet-507 tasks69.5
- Market rankingmarket-ranking-ecom-v2 · 11s
- Delegated accessdelegated-access-ecom-v1 · 1m 34s
- Task completionjourney-ecom-v1 · 1m 57s
- Discoveryjourney-ecom-v1 · 1m 57s
- Information retrievalresearch-ecom-v1 · 2m 37s
- Brand awarenessresearch-ecom-v1 · 2m 42s
- Accuracyresearch-ecom-v1 · 2m 37s
ChatGPT Agentgpt-5-mini08 tasks70.5
- Market rankingmarket-ranking-ecom-v2 · 18s
- Contact & communicationsupport-ecom-v1 · 1m 59s
- Delegated accessdelegated-access-ecom-v1 · 2m 39s
- Task completionjourney-ecom-v1 · 4m 32s
- Discoveryjourney-ecom-v1 · 4m 30s
- Information retrievalresearch-ecom-v1 · 10m 55s
- Accuracyresearch-ecom-v1 · 10m 56s
- Brand awarenessresearch-ecom-v1 · 10m 56s
Geminigemini-3.8-flash08 tasks69.4
- Market rankingmarket-ranking-ecom-v2 · 12s
- Contact & communicationsupport-ecom-v1 · 1m 21s
- Delegated accessdelegated-access-ecom-v1 · 3m 23s
- Task completionjourney-ecom-v1 · 5m 24s
- Discoveryjourney-ecom-v1 · 5m 24s
- Accuracyresearch-ecom-v1 · 7m 11s
- Brand awarenessresearch-ecom-v1 · 7m 12s
- Information retrievalresearch-ecom-v1 · 7m 13s
Where agents get stuck
5 issues, worst first.
Where agents get stuck
5 issues, worst first.
Show 1 more issueShow fewer
Want to find these in your own flows before an agent does? Read how to test your product for AI agent users.
Technical readiness
The protocol baseline: robots, sitemaps, llms.txt, structured data and more.
63/100
Technical readiness
The protocol baseline: robots, sitemaps, llms.txt, structured data and more.
Protocol baseline by Is Agentic
- Score
- 63/ 100
- Important blockers remain
- Essential
- 44.8/ 80
- 5 of 11 passing
- Recommended
- 14.4/ 20
- 10 of 18 passing
- Bonus
- +4.2
- 19 positive signals
Show 11 more issuesShow fewer
Related domains
index.stuntdouble.io is a subdomain of stuntdouble.io. Agent experience can differ sharply across subdomains.
Embed
Put this score on your site
A live badge for your footer, docs or README. It updates after every run and links back to this report.