Fetch the page
One baseline request reads the raw HTML your origin sends without waiting for a browser.
See which answer-engine crawlers can reach your pages, before you spend another dollar on content.
Raw HTML + robots precedence + ten live User-Agent probes.
CrawlProof is a deterministic audit. It does not guess what an LLM thinks, promise citations, or render a fake browser. It checks the request path your origin actually receives and shows the evidence in plain language.
No signup. No model tokens. No browser rendering. Paste a public page and get the first evidence in seconds.
The report below is a real CrawlProof scan of a real public page. Status codes, crawler names, robots-only tokens, and the limitation note are all from the product.
source: example.com / captured 2026-08-08

The audit follows the evidence path instead of hiding it behind a score.
One baseline request reads the raw HTML your origin sends without waiting for a browser.
The matching robots group, longest path rule, and equal-length Allow precedence are evaluated.
Ten documented User-Agents receive header-only requests. Google-Extended and Applebot-Extended stay robots-only.
Headings, canonicals, JSON-LD, visible text, noscript, and app-shell clues become findings you can act on.
Small checks, assembled into a report an owner or agency can explain to a client.
GPTBot, OAI-SearchBot, ChatGPT-User, Claude, Perplexity, Google, Meta, Apple, and Amazon.
Specific groups beat wildcards. Longest matching rules win. Equal-length Allow wins. The two robots-only tokens are evaluated without pretending they have HTTP User-Agents.
See whether the title, description, canonical, headings, visible body text, noscript, and structured data arrived without rendering.
A low-text app shell is called out. It is a signal to investigate, not a claim that every JavaScript site is broken.
Structured data is parsed safely and checked for a context or type. We report the signal, not a magic ranking promise.
Solo starts with one page. Agency can fan out across up to 250 discovered URLs without a queue or a browser bill.
Run client sites again, keep the findings in D1, and compose a compact report in the R2-backed audit record.
Agency findings are structured for a branded report workflow. The underlying evidence stays visible and citable.
These are the settled comparison figures from our research. Third-party captures remain labelled as such.
| Tool | Price observed | What it does | What CrawlProof does instead |
|---|---|---|---|
| Profound | $99–$399/mo | LLM visibility and citation tracking | One-time technical evidence, no model inference |
| Peec AI | ~$95+/mo, third-party capture | AI search monitoring | Per-agent reachability and raw HTML checks |
| AthenaHQ | ~$295/mo, third-party capture | Answer-engine analytics | Robots and HTTP evidence without a subscription |
| Semrush AI Visibility Toolkit | $99/mo per domain | AI visibility and an AI-readiness site audit | Focused deterministic audit, $69 once |
| Ahrefs Brand Radar | from $199/mo | Brand visibility in AI answers | No citation tracking, no recurring model cost |
| ZeroRank AI | $89 LTD observed | AI-search citation monitoring | Technical half of GEO, not citation monitoring |
| CrawlProof Solo | $69 once | Ten crawler probes, robots precedence, HTML signals | What they do that we don't: LLM citation tracking. |
Founder price: $48.30, 30% off while the real first-200 allocation remains.
No inflated promise, no fake customer wall, no mystery score.
Parts of it are. Free tools can parse robots.txt, look for schema, or check an llms.txt file. The gap is the combined evidence: a baseline fetch, ten live crawler User-Agent requests per page, correct robots precedence, and raw HTML signals across a sitemap. If you only need a robots text parser, use the free one.
No. CrawlProof cannot promise citations. Schema's causal effect is unsettled, and answer engines change. We show whether a supplied request was allowed, blocked, challenged, or served a thin app shell. That is evidence for remediation, not a citation guarantee.
We can report whether it exists, but adoption by major providers is unconfirmed. It is not the headline feature and we do not claim it is a proven ranking lever.
No. The probe does not reproduce the real crawler's IP reputation, WAF path, TLS fingerprint, or challenge flow. Treat it as strong evidence about the supplied request.
Lifetime access to the current CrawlProof product, with no recurring subscription and no usage-metered model bill. It does not promise every future feature forever or third-party crawler behavior that changes outside our control.
30 days, stated plainly. If the evidence is not useful for your site, request a refund within 30 days through the merchant-of-record checkout.
Run the free scan first. If the evidence is useful, choose Solo or Agency access for repeat audits.
The founder allocation is still live: 200 licences remain. The per-visitor $48.30 price ends when the timer hits 0, then Solo is $69.
Keep the founder price