Before a model can cite your site, it has to be allowed to read it. Access is decided in one file most teams last touched when they set up the project — and never checked again.
Each is a separate user agent with its own rule in robots.txt. Allowing one says nothing about the others.
| User agent | Operated by | Used for |
|---|---|---|
| GPTBot | OpenAI | Training and answers in ChatGPT |
| ClaudeBot | Anthropic | Answers and browsing in Claude |
| PerplexityBot | Perplexity | Search results in Perplexity |
| Google-Extended | Grounding for Google AI answers | |
| CCBot | Common Crawl | Open dataset used by many models |
The list changes. New agents appear, existing ones split into separate names — which is why a rule set that is never updated silently drifts out of date.
Open your robots.txt and look for the agents above. No entry means no decision was made.
Request /llms.txt. A 404 means a model has to infer your site from your navigation.
View source on your pricing page and find the title. If it matches your home page, the two are indistinguishable.
Disable JavaScript and reload. Whatever disappears is content some crawlers will never see.
The tests above tell you where you stand today. The Web Ready report runs them with every scan, keeps the history and mails you when something flips — because the change that costs you visibility is almost never the one you made on purpose.
One tier, one price. There is no separate Web Ready subscription to work out.
Price excl. VAT — added at checkout where applicable. B2B only.