Cloudflare's AI-crawler default change is 15 Sep 2026 loading…Run a free scanGet lifetime access
The quiet visibility problem

Your site can rank fine on Google and still be invisible to ChatGPT. Most owners find out months later.

See which answer-engine crawlers can reach your pages, before you spend another dollar on content.

Raw HTML + robots precedence + ten live User-Agent probes.

CrawlProof is a deterministic audit. It does not guess what an LLM thinks, promise citations, or render a fake browser. It checks the request path your origin actually receives and shows the evidence in plain language.

Scan one URL now

No signup. No model tokens. No browser rendering. Paste a public page and get the first evidence in seconds.

Ready. Your URL stays in your browser until you submit.
n/abaseline status
n/acrawler probes
n/aH1 headings
n/aJS dependency

This is the output, not a design concept.

The report below is a real CrawlProof scan of a real public page. Status codes, crawler names, robots-only tokens, and the limitation note are all from the product.

source: example.com / captured 2026-08-08

Genuine CrawlProof scan output showing HTML signals and crawler verdicts
Genuine live scan output. No invented dashboard data.

Four checks. One useful answer.

The audit follows the evidence path instead of hiding it behind a score.

01

Fetch the page

One baseline request reads the raw HTML your origin sends without waiting for a browser.

02

Read the rules

The matching robots group, longest path rule, and equal-length Allow precedence are evaluated.

03

Probe the agents

Ten documented User-Agents receive header-only requests. Google-Extended and Applebot-Extended stay robots-only.

04

Show the signals

Headings, canonicals, JSON-LD, visible text, noscript, and app-shell clues become findings you can act on.

HONEST LIMITATION: a probe tests how the origin/CDN treats that supplied request. It does not replicate the real crawler's source IP, WAF reputation, TLS fingerprint or challenge path, so a 403 is strong evidence about the probe, not proof every real crawler is blocked.

What you actually get.

Small checks, assembled into a report an owner or agency can explain to a client.

01

Ten live crawler probes

GPTBot, OAI-SearchBot, ChatGPT-User, Claude, Perplexity, Google, Meta, Apple, and Amazon.

  • status, Location, challenge, server, Retry-After
  • no probe response bodies are read
02

Robots precedence, not a text search

Specific groups beat wildcards. Longest matching rules win. Equal-length Allow wins. The two robots-only tokens are evaluated without pretending they have HTTP User-Agents.

03

Raw HTML visibility

See whether the title, description, canonical, headings, visible body text, noscript, and structured data arrived without rendering.

04

JS dependency flag

A low-text app shell is called out. It is a signal to investigate, not a claim that every JavaScript site is broken.

05

JSON-LD validity

Structured data is parsed safely and checked for a context or type. We report the signal, not a magic ranking promise.

06

Sitemap fan-out

Solo starts with one page. Agency can fan out across up to 250 discovered URLs without a queue or a browser bill.

07

Reusable agency evidence

Run client sites again, keep the findings in D1, and compose a compact report in the R2-backed audit record.

08

White-label-ready output

Agency findings are structured for a branded report workflow. The underlying evidence stays visible and citable.

Same anxiety. Different bill.

These are the settled comparison figures from our research. Third-party captures remain labelled as such.

ToolPrice observedWhat it doesWhat CrawlProof does instead
Profound$99–$399/moLLM visibility and citation trackingOne-time technical evidence, no model inference
Peec AI~$95+/mo, third-party captureAI search monitoringPer-agent reachability and raw HTML checks
AthenaHQ~$295/mo, third-party captureAnswer-engine analyticsRobots and HTTP evidence without a subscription
Semrush AI Visibility Toolkit$99/mo per domainAI visibility and an AI-readiness site auditFocused deterministic audit, $69 once
Ahrefs Brand Radarfrom $199/moBrand visibility in AI answersNo citation tracking, no recurring model cost
ZeroRank AI$89 LTD observedAI-search citation monitoringTechnical half of GEO, not citation monitoring
CrawlProof Solo$69 onceTen crawler probes, robots precedence, HTML signalsWhat they do that we don't: LLM citation tracking.
A note from the founder

Cloudflare's next default change is a real date, not a marketing countdown. I built CrawlProof around the part of GEO that can be honest as a lifetime product: fetching public pages and reporting deterministic evidence. No metered LLM calls means your lifetime access does not quietly become a feature cap.

Start with the evidence.

$99/mo elsewhere for a visibility subscription
$69
Solo, one-time payment

Founder price: $48.30, 30% off while the real first-200 allocation remains.

  • 50-page sitemap cap
  • 10 crawler probes per page
  • robots, HTML, JSON-LD, and JS signals
  • 30-day money-back guarantee

Questions worth asking before you buy.

No inflated promise, no fake customer wall, no mystery score.

Isn't this free elsewhere?

Parts of it are. Free tools can parse robots.txt, look for schema, or check an llms.txt file. The gap is the combined evidence: a baseline fetch, ten live crawler User-Agent requests per page, correct robots precedence, and raw HTML signals across a sitemap. If you only need a robots text parser, use the free one.

Will this get me cited by ChatGPT?

No. CrawlProof cannot promise citations. Schema's causal effect is unsettled, and answer engines change. We show whether a supplied request was allowed, blocked, challenged, or served a thin app shell. That is evidence for remediation, not a citation guarantee.

Does llms.txt make me visible?

We can report whether it exists, but adoption by major providers is unconfirmed. It is not the headline feature and we do not claim it is a proven ranking lever.

Does a 403 prove the real crawler is blocked?

No. The probe does not reproduce the real crawler's IP reputation, WAF path, TLS fingerprint, or challenge flow. Treat it as strong evidence about the supplied request.

What does lifetime mean?

Lifetime access to the current CrawlProof product, with no recurring subscription and no usage-metered model bill. It does not promise every future feature forever or third-party crawler behavior that changes outside our control.

What is the refund policy?

30 days, stated plainly. If the evidence is not useful for your site, request a refund within 30 days through the merchant-of-record checkout.

30-day money-back guaranteeMerchant of Record: Dodo PaymentsOne-time checkout, not a subscriptionNo LLM inference in the scan

Find the block before the buyer finds it.

Run the free scan first. If the evidence is useful, choose Solo or Agency access for repeat audits.

Run a free scan