Automated checks probe the site before any agent runs. They count for 40% of each category. Strongest is Information retrieval at 85.0, weakest is Delegated access at 10.0.
Site checks by category
Brand awareness
80.0
Discovery
70.0
Information retrieval
85.0
Market ranking
35.0
Accuracy
60.0
Task completion
75.0
Delegated access
10.0
Contact & communication
30.0
01 Site checks
Points by category
Brand awareness
80.0
Discovery
70.0
Information retrieval
85.0
Market ranking
35.0
Accuracy
60.0
Task completion
75.0
Delegated access
10.0
Contact & communication
30.0
Agent readiness checks
Essential
69.3/80
Recommended
10.2/20
Checks passing
15/28
Scoped permissionsEssential
JSON-LD structured dataRecommended
API schema complexity analysisRecommended
Function calling compatibilityRecommended
Agent instruction / when-to-useRecommended
02 Agent sessions
Brand awareness
Find the product name, its one-sentence value proposition, and the primary user persona it targets.
Did the agent correctly identify the product name (30pts), the value prop (40pts), and the target user (30pts)? Deduct 20pts for hallucinations.
2m 34s
99median 95
Find the product name, its one-sentence value proposition, and the primary user persona it targets.
Did the agent correctly identify the product name (30pts), the value prop (40pts), and the target user (30pts)? Deduct 20pts for hallucinations.
2m 5s
98median 95
Find the product name, its one-sentence value proposition, and the primary user persona it targets.
Did the agent correctly identify the product name (30pts), the value prop (40pts), and the target user (30pts)? Deduct 20pts for hallucinations.
3m 42s
98median 95
Find the product name, its one-sentence value proposition, and the primary user persona it targets.
Did the agent correctly identify the product name (30pts), the value prop (40pts), and the target user (30pts)? Deduct 20pts for hallucinations.
7m 31s
92median 95
Find the product name, its one-sentence value proposition, and the primary user persona it targets.
Did the agent correctly identify the product name (30pts), the value prop (40pts), and the target user (30pts)? Deduct 20pts for hallucinations.
4m 43s
0median 95
Discovery
Navigate to the pricing page, and note its URL and the names of any pricing tiers shown.
Found the pricing page within 3 nav hops (60pts). Correctly identified it as the pricing page and listed at least one tier (40pts). Deduct 20pts for each dead end.
46s
98median 91
Navigate to the pricing page, and note its URL and the names of any pricing tiers shown.
Found the pricing page within 3 nav hops (60pts). Correctly identified it as the pricing page and listed at least one tier (40pts). Deduct 20pts for each dead end.
42s
100median 91
Navigate to the pricing page, and note its URL and the names of any pricing tiers shown.
Found the pricing page within 3 nav hops (60pts). Correctly identified it as the pricing page and listed at least one tier (40pts). Deduct 20pts for each dead end.
2m
100median 91
Navigate to the pricing page, and note its URL and the names of any pricing tiers shown.
Found the pricing page within 3 nav hops (60pts). Correctly identified it as the pricing page and listed at least one tier (40pts). Deduct 20pts for each dead end.
3m 39s
88median 91
Navigate to the pricing page, and note its URL and the names of any pricing tiers shown.
Found the pricing page within 3 nav hops (60pts). Correctly identified it as the pricing page and listed at least one tier (40pts). Deduct 20pts for each dead end.
2m 25s
26median 91
Information retrieval
Find the starter or free tier price, one headline feature of the product, and the length of any free trial offered.
Starter/free price (30pts), headline feature (30pts), trial length (30pts). Correct and sourced from the site (10pts). Deduct 20pts for any hallucinated data.
2m 35s
89median 76
Find the starter or free tier price, one headline feature of the product, and the length of any free trial offered.
Starter/free price (30pts), headline feature (30pts), trial length (30pts). Correct and sourced from the site (10pts). Deduct 20pts for any hallucinated data.
2m 6s
82median 76
Find the starter or free tier price, one headline feature of the product, and the length of any free trial offered.
Starter/free price (30pts), headline feature (30pts), trial length (30pts). Correct and sourced from the site (10pts). Deduct 20pts for any hallucinated data.
3m 42s
78median 76
Market ranking
Based on your knowledge, name 3 direct competitors to papermill.io in the SaaS space and rank them by market share.
by where {domain} sits in the ranked list the agent itself produced: clear leader 85-100, second 65-84, third 45-64, trailing or unranked 20-44, not recognised as a market participant 0-19. Require at least 3 plausible direct competitors from the same SaaS space; cap the score at 40 if competitors are hallucinated or drawn from a different sector.
12s
43median 66
Based on your knowledge, name 3 direct competitors to papermill.io in the SaaS space and rank them by market share.
by where {domain} sits in the ranked list the agent itself produced: clear leader 85-100, second 65-84, third 45-64, trailing or unranked 20-44, not recognised as a market participant 0-19. Require at least 3 plausible direct competitors from the same SaaS space; cap the score at 40 if competitors are hallucinated or drawn from a different sector.
8s
51median 66
Based on your knowledge, name 3 direct competitors to papermill.io in the SaaS space and rank them by market share.
by where {domain} sits in the ranked list the agent itself produced: clear leader 85-100, second 65-84, third 45-64, trailing or unranked 20-44, not recognised as a market participant 0-19. Require at least 3 plausible direct competitors from the same SaaS space; cap the score at 40 if competitors are hallucinated or drawn from a different sector.
13s
43median 66
Accuracy
Find out whether individual users can access the Enterprise tier, and the specific SSO options listed for Enterprise.
Answered whether individuals can access Enterprise (50pts). Listed the specific SSO options correctly (50pts). Deduct 40pts for hallucinations or invented features.
2m 34s
76median 64
Find out whether individual users can access the Enterprise tier, and the specific SSO options listed for Enterprise.
Answered whether individuals can access Enterprise (50pts). Listed the specific SSO options correctly (50pts). Deduct 40pts for hallucinations or invented features.
2m 5s
65median 64
Find out whether individual users can access the Enterprise tier, and the specific SSO options listed for Enterprise.
Answered whether individuals can access Enterprise (50pts). Listed the specific SSO options correctly (50pts). Deduct 40pts for hallucinations or invented features.
3m 41s
54median 64
Task completion
Click the primary CTA for any paid plan and reach the signup or upgrade form.
Found the pricing page (20pts). Clicked the primary CTA (20pts). Reached the signup/upgrade form (40pts). Reported the URL and form fields (20pts). Deduct 20pts for each bot challenge or CAPTCHA hit.
46s
99median 85
Click the primary CTA for any paid plan and reach the signup or upgrade form.
Found the pricing page (20pts). Clicked the primary CTA (20pts). Reached the signup/upgrade form (40pts). Reported the URL and form fields (20pts). Deduct 20pts for each bot challenge or CAPTCHA hit.
44s
98median 85
Click the primary CTA for any paid plan and reach the signup or upgrade form.
Found the pricing page (20pts). Clicked the primary CTA (20pts). Reached the signup/upgrade form (40pts). Reported the URL and form fields (20pts). Deduct 20pts for each bot challenge or CAPTCHA hit.
2m 2s
96median 85
Click the primary CTA for any paid plan and reach the signup or upgrade form.
Found the pricing page (20pts). Clicked the primary CTA (20pts). Reached the signup/upgrade form (40pts). Reported the URL and form fields (20pts). Deduct 20pts for each bot challenge or CAPTCHA hit.
3m 38s
100median 85
Click the primary CTA for any paid plan and reach the signup or upgrade form.
Found the pricing page (20pts). Clicked the primary CTA (20pts). Reached the signup/upgrade form (40pts). Reported the URL and form fields (20pts). Deduct 20pts for each bot challenge or CAPTCHA hit.
2m 25s
32median 85
03 Blend
Agents 60%
Checks 40%
One kind of evidence only? That kind counts in full.
Brand awareness10% of the overall
0.4 × 80.0 + 0.6 × 95.7 = 89.4 raw
89.4 raw → 95.1 published
Discovery15% of the overall
0.4 × 70.0 + 0.6 × 98.3 = 87.0 raw
87.0 raw → 93.9 published
Information retrieval15% of the overall
0.4 × 85.0 + 0.6 × 67.0 = 74.2 raw
74.2 raw → 87.4 published
Market ranking10% of the overall
0.4 × 35.0 + 0.6 × 17.3 = 24.4 raw
24.4 raw → 53.0 published
Accuracy15% of the overall
0.4 × 60.0 + 0.6 × 39.3 = 47.6 raw
47.6 raw → 71.6 published
Task completion20% of the overall
0.4 × 75.0 + 0.6 × 95.0 = 87.0 raw
87.0 raw → 93.9 published
Delegated access10% of the overall
Site checks only: 10.0 raw
10.0 raw → 35.5 published
Contact & communication5% of the overall
Site checks only: 30.0 raw
30.0 raw → 58.2 published
Category
Weight
Checks
Sessions
Blend
Benchmarked
Brand awareness
10%
80.0
95.7
89.4
95.1
Discovery
15%
70.0
98.3
87.0
93.9
Information retrieval
15%
85.0
67.0
74.2
87.4
Market ranking
10%
35.0
17.3
24.4
53.0
Accuracy
15%
60.0
39.3
47.6
71.6
Task completion
20%
75.0
95.0
87.0
93.9
Delegated access
10%
10.0
None
10.0
35.5
Contact & communication
5%
30.0
None
30.0
58.2
Overall
100%
~62.6
81.0
04 Benchmarked score
62.6 raw → 81.0
Final score is calibrated and adjusted to better reflect the actual performance and benchmark.
Category weights
Brand awareness 10%
Discovery 15%
Information retrieval 15%
Market ranking 10%
Accuracy 15%
Task completion 20%
Delegated access 10%
Contact & communication 5%
Overall 81.0, from 62.6 raw points.
Agents
Each agent’s score over 30 days, and how its latest sessions went.
Claude
Latest run average
86.8
SaaS median
90.3
Sessions
Brand awareness
Find the product name, its one-sentence value proposition, and the primary user persona it targets.
Did the agent correctly identify the product name (30pts), the value prop (40pts), and the target user (30pts)? Deduct 20pts for hallucinations.
2m 34s
99median 95
Discovery
Navigate to the pricing page, and note its URL and the names of any pricing tiers shown.
Found the pricing page within 3 nav hops (60pts). Correctly identified it as the pricing page and listed at least one tier (40pts). Deduct 20pts for each dead end.
46s
98median 91
Information retrieval
Find the starter or free tier price, one headline feature of the product, and the length of any free trial offered.
Starter/free price (30pts), headline feature (30pts), trial length (30pts). Correct and sourced from the site (10pts). Deduct 20pts for any hallucinated data.
2m 35s
89median 76
Market ranking
Based on your knowledge, name 3 direct competitors to papermill.io in the SaaS space and rank them by market share.
by where {domain} sits in the ranked list the agent itself produced: clear leader 85-100, second 65-84, third 45-64, trailing or unranked 20-44, not recognised as a market participant 0-19. Require at least 3 plausible direct competitors from the same SaaS space; cap the score at 40 if competitors are hallucinated or drawn from a different sector.
12s
43median 66
Accuracy
Find out whether individual users can access the Enterprise tier, and the specific SSO options listed for Enterprise.
Answered whether individuals can access Enterprise (50pts). Listed the specific SSO options correctly (50pts). Deduct 40pts for hallucinations or invented features.
2m 34s
76median 64
Task completion
Click the primary CTA for any paid plan and reach the signup or upgrade form.
Found the pricing page (20pts). Clicked the primary CTA (20pts). Reached the signup/upgrade form (40pts). Reported the URL and form fields (20pts). Deduct 20pts for each bot challenge or CAPTCHA hit.
46s
99median 85
ChatGPT Agent
Latest run average
84.7
SaaS median
84.0
Sessions
Brand awareness
Find the product name, its one-sentence value proposition, and the primary user persona it targets.
Did the agent correctly identify the product name (30pts), the value prop (40pts), and the target user (30pts)? Deduct 20pts for hallucinations.
2m 5s
98median 95
Discovery
Navigate to the pricing page, and note its URL and the names of any pricing tiers shown.
Found the pricing page within 3 nav hops (60pts). Correctly identified it as the pricing page and listed at least one tier (40pts). Deduct 20pts for each dead end.
42s
100median 91
Information retrieval
Find the starter or free tier price, one headline feature of the product, and the length of any free trial offered.
Starter/free price (30pts), headline feature (30pts), trial length (30pts). Correct and sourced from the site (10pts). Deduct 20pts for any hallucinated data.
2m 6s
82median 76
Market ranking
Based on your knowledge, name 3 direct competitors to papermill.io in the SaaS space and rank them by market share.
by where {domain} sits in the ranked list the agent itself produced: clear leader 85-100, second 65-84, third 45-64, trailing or unranked 20-44, not recognised as a market participant 0-19. Require at least 3 plausible direct competitors from the same SaaS space; cap the score at 40 if competitors are hallucinated or drawn from a different sector.
8s
51median 66
Accuracy
Find out whether individual users can access the Enterprise tier, and the specific SSO options listed for Enterprise.
Answered whether individuals can access Enterprise (50pts). Listed the specific SSO options correctly (50pts). Deduct 40pts for hallucinations or invented features.
2m 5s
65median 64
Task completion
Click the primary CTA for any paid plan and reach the signup or upgrade form.
Found the pricing page (20pts). Clicked the primary CTA (20pts). Reached the signup/upgrade form (40pts). Reported the URL and form fields (20pts). Deduct 20pts for each bot challenge or CAPTCHA hit.
44s
98median 85
Gemini
Latest run average
81.9
SaaS median
87.3
Sessions
Brand awareness
Find the product name, its one-sentence value proposition, and the primary user persona it targets.
Did the agent correctly identify the product name (30pts), the value prop (40pts), and the target user (30pts)? Deduct 20pts for hallucinations.
3m 42s
98median 95
Discovery
Navigate to the pricing page, and note its URL and the names of any pricing tiers shown.
Found the pricing page within 3 nav hops (60pts). Correctly identified it as the pricing page and listed at least one tier (40pts). Deduct 20pts for each dead end.
2m
100median 91
Information retrieval
Find the starter or free tier price, one headline feature of the product, and the length of any free trial offered.
Starter/free price (30pts), headline feature (30pts), trial length (30pts). Correct and sourced from the site (10pts). Deduct 20pts for any hallucinated data.
3m 42s
78median 76
Market ranking
Based on your knowledge, name 3 direct competitors to papermill.io in the SaaS space and rank them by market share.
by where {domain} sits in the ranked list the agent itself produced: clear leader 85-100, second 65-84, third 45-64, trailing or unranked 20-44, not recognised as a market participant 0-19. Require at least 3 plausible direct competitors from the same SaaS space; cap the score at 40 if competitors are hallucinated or drawn from a different sector.
13s
43median 66
Accuracy
Find out whether individual users can access the Enterprise tier, and the specific SSO options listed for Enterprise.
Answered whether individuals can access Enterprise (50pts). Listed the specific SSO options correctly (50pts). Deduct 40pts for hallucinations or invented features.
3m 41s
54median 64
Task completion
Click the primary CTA for any paid plan and reach the signup or upgrade form.
Found the pricing page (20pts). Clicked the primary CTA (20pts). Reached the signup/upgrade form (40pts). Reported the URL and form fields (20pts). Deduct 20pts for each bot challenge or CAPTCHA hit.
2m 2s
96median 85
Kimi
Latest run average
93.3
SaaS median
75.8
Sessions
Brand awareness
Find the product name, its one-sentence value proposition, and the primary user persona it targets.
Did the agent correctly identify the product name (30pts), the value prop (40pts), and the target user (30pts)? Deduct 20pts for hallucinations.
7m 31s
92median 95
Discovery
Navigate to the pricing page, and note its URL and the names of any pricing tiers shown.
Found the pricing page within 3 nav hops (60pts). Correctly identified it as the pricing page and listed at least one tier (40pts). Deduct 20pts for each dead end.
3m 39s
88median 91
Task completion
Click the primary CTA for any paid plan and reach the signup or upgrade form.
Found the pricing page (20pts). Clicked the primary CTA (20pts). Reached the signup/upgrade form (40pts). Reported the URL and form fields (20pts). Deduct 20pts for each bot challenge or CAPTCHA hit.
3m 38s
100median 85
Llama
Latest run average
24.3
SaaS median
53.2
Sessions
Brand awareness
Find the product name, its one-sentence value proposition, and the primary user persona it targets.
Did the agent correctly identify the product name (30pts), the value prop (40pts), and the target user (30pts)? Deduct 20pts for hallucinations.
4m 43s
0median 95
Discovery
Navigate to the pricing page, and note its URL and the names of any pricing tiers shown.
Found the pricing page within 3 nav hops (60pts). Correctly identified it as the pricing page and listed at least one tier (40pts). Deduct 20pts for each dead end.
2m 25s
26median 91
Task completion
Click the primary CTA for any paid plan and reach the signup or upgrade form.
Found the pricing page (20pts). Clicked the primary CTA (20pts). Reached the signup/upgrade form (40pts). Reported the URL and form fields (20pts). Deduct 20pts for each bot challenge or CAPTCHA hit.
2m 25s
32median 85
Excellent85+
Good72-84
OK58-71
Poor40-57
Critical<40
Latest run 1 Oct 2026 (UTC). Sign in to request a re-run from 2 Oct 2026, 11:29 UTC.
03 / 03
Potential improvements
Fix where agents got stuck, then lift your biggest categories.
HighDelegated access
No discoverable MCP server. Agents and AI clients have no way to discover one automatically.
Biggest wins
Projected
Lifting Contact & communication to Good adds about +0.5 on its own, taking the overall from 81.0 to 81.5.
Raise Contact & communication from 58.2 to 72 (Good) to add about 0.5 to the overall score.
No declared OAuth scopes, security schemes, or scoped-permission documentation found
Fix: Declare scoped API permissions where machines can read them: named OAuth scopes in your OpenAPI security schemes, or scopes_supported in RFC 9728 protected-resource metadata. Prose descriptions of roles help humans, but agents need the machine-readable declaration to request least-privilege access.
Investigate
JSON-LD structured datarecommended
No JSON-LD structured data found on homepage
Fix: Add JSON-LD structured data to your homepage using the identity type that matches your site - SoftwareApplication for products, Organization or LocalBusiness for companies, Person for personal sites, Article for blogs - with name, description, url, and type-appropriate fields (offers, sameAs, author) so AI can parse your identity programmatically.
Investigate
API schema complexity analysisrecommended
No API schema detected
Fix: Make your API spec self-describing: a unique operationId and a description on every operation, typed parameters, and response schemas. For GraphQL, a fully typed schema with a documented cost or rate limit reads best.
Investigate
Function calling compatibilityrecommended
No API spec found - function calling requires discoverable endpoints
Fix: Ensure API endpoints have unique operation IDs, typed schemas, and descriptions compatible with LLM function-calling formats.
Investigate
Agent instruction / when-to-userecommended
No agent instruction file with when-to-use guidance found
Fix: Tell agents when to reach for you: add a 'when to use this' section to your llms.txt (or a dedicated agent-instructions file) that names your best-fit use cases and how an agent should call you. Be specific about the jobs you are right for - generic marketing copy does not read as guidance.
Investigate
Organization schema completenessrecommended
No JSON-LD found - Organization schema missing
Fix: Add Organization JSON-LD that includes both contactPoint (with email/phone and contactType) and address (PostalAddress). This lets AI verify your business legitimacy and answer contact queries.
Investigate
Trust anchor pagesrecommended
No trust anchor pages found with sufficient content (About, Contact, Privacy)
Fix: Publish real /about, /contact, and /privacy pages with at least 500 characters of content each. These are the pages AI agents check to verify your business is legitimate before recommending you.
Partial
Content without JavaScriptessential
7925 chars with H1, but 2.1% content ratio is below the 5% target
Fix: Serve at least 500 characters of meaningful homepage content in raw HTML. Add a clear H1, keep deeper heading levels sequential, and remove excessive non-content markup.
Partial
Developer resource discoverabilityrecommended
Name search surfaced no pages on papermill.io, although developer resources exist on the site (API docs, developer portal, auth docs, MCP server) - weak search indexing or transient search noise
Fix: Check whether your developer resources (API docs, OpenAPI spec, auth docs, developer portal, MCP server, SDK documentation) surface in name-based searches. If they do not, use predictable URLs, link them in llms.txt, and include your product name in page titles and headings. This result reflects one search sample.
Partial
Brand name discoverabilityrecommended
papermill.io appears once in brand-name search results for "Papermill" (position #4 out of 10)
Fix: Make sure a clean search for your brand name returns your own domain in the top results. If it does not, your brand may be too generic, conflict with a more established term, or not yet indexed. Strengthen brand-name search by claiming consistent NAP across listings, earning press mentions that link to the canonical domain, and avoiding redirect chains that mask the apex domain in search results.
Partial
MCP server / manifestrecommended
MCP mentioned at https://docs.papermill.io but no standard manifest endpoint found
Fix: Build an MCP (Model Context Protocol) server exposing your API as tools. Use Streamable HTTP transport for full score. This lets Claude, ChatGPT, and other AI agents call your product natively.
Partial
Agent onboarding frictionrecommended
Onboarding signals described but not verified live: free tier available, self-serve key generation, sandbox/test environment
Fix: Offer a free tier or trial, self-serve API key generation, and a sandbox environment. Agents can't fill out 'contact sales' forms.
Partial
CLI tool availablerecommended
CLI tool mentioned in llms.txt
Fix: Publish an official CLI tool on npm, PyPI, or Homebrew. A CLI lets agents and developers script interactions with your product without building API integrations from scratch.
28 checks · scanned 1 Oct 2026
Opportunity cost · SaaS
Papermill could see ~$1,292,000 of new ARR exposed to agent-led journeys each year