Platform guides
How AI crawler policy signals — robots.txt, meta directives, and response headers — actually work on specific hosting platforms and CDNs. Platform behaviour can change; every guide below shows when its official sources were last verified, and links directly to them.
Managed CDN
Cloudflare
How Cloudflare's managed robots.txt and AI Crawl Control interact with your site's public AI crawler policy — verified against official Cloudflare documentation.
Hosted application platforms
Shopify
How Shopify's robots.txt.liquid customization mechanism works, its default rules, and its documented limits — verified against official Shopify documentation.
WordPress
How WordPress's virtual robots.txt, the 'Discourage search engines' setting, and theme/plugin interactions affect what your site tells AI crawlers — verified against official WordPress documentation.
Deployment platforms
Netlify
How Netlify's automatic deploy-preview noindex header, _headers file, and deploy contexts affect what your site tells AI crawlers — verified against official Netlify documentation.
Vercel
How Vercel's automatic preview-deployment noindex header, static and generated robots.txt files, and custom headers affect what your site tells AI crawlers — verified against official Vercel and Next.js documentation.
Check your own deployed policy
Run a free audit to see exactly what your domain currently tells AI crawlers, regardless of which platform serves it.
Audit a domain