CrawlPact

Content Signals checker

Enter a domain to detect a recognised Content-Signal response header, plus its meta robots tag, X-Robots-Tag header, and canonical URL — and flag any disagreement against its RSL declaration or robots.txt rules.

What this checks

  • Detects a Content-Signal response header and its declared search/ai-train/ai-input values.
  • Also checks meta robots tags, the X-Robots-Tag header, and the canonical URL for the same page.
  • Flags disagreement between Content Signals and the domain's RSL declaration or robots.txt rules.

What this does not check

Content Signals is a proposal (published by Cloudflare, September 2025) — not a ratified standard, and no major crawler is currently known to honour it in practice. A separate, more formal IETF effort (AI Preferences) is still in draft. See limitations.

Related

Want the complete picture?

This checker is a view into a full CrawlPact audit, which also covers robots.txt, llms.txt, RSL, and your Policy Health Score.

Run the full audit

Analytics preferences

CrawlPact uses optional Google Analytics and Microsoft Clarity on public marketing pages to understand which content is useful and how visitors actually use it. Neither is used in the authenticated app or admin areas. You can accept or decline analytics without affecting the service.