GPTBot
Crawls public web pages to collect training data for OpenAI's GPT models. Allowing it means your content can inform future models; there is no traffic in return.
| robots.txt token | GPTBot |
| Respects robots.txt | Yes — documented and generally observed |
| Verification | OpenAI publishes GPTBot IP ranges; Cloudflare verified-bot flag covers it |
| Our recommendation | Your call — Legitimate and polite, but the exchange is one-sided — decide whether model exposure is worth donating your content. |
Block GPTBot with robots.txt
User-agent: GPTBot
Disallow: /robots.txt is a request, not a wall — requests claiming to be GPTBot from unverified networks should be treated as bad bots. How to verify crawlers by network →
See every GPTBot request hitting your site — live, with per-bot allow / block / tarpit controls. Try TrafficDATA free (100K page views)