Websites blocking OAI-SearchBot
OpenAI's search crawler. It builds the index behind ChatGPT search features and is separate from GPTBot, so sites can allow search visibility while blocking model training.
Operated by OpenAI · Search and AI assistant · operator documentation
* block
0.40%
restrict it by name
99,650 domains name it and close some of its paths
1.18%
name it in robots.txt
293K domains give it rules of their own
Data for Aug 2026. All four percentages
divide by the same 24.91 million domains. Blocked counts both
explicit User-agent: OAI-SearchBot blocks and blanket
* blocks the crawler inherits.
Historical mentions, HTTP Archive
High-authority sites blocking OAI-SearchBot by name
| Domain | OPROpen Page Rank | Blocking since |
|---|---|---|
| nytimes.com | 9.28/10 | Aug 2026 |
| bbc.com | 9.15/10 | Aug 2026 |
| congress.gov | 9.05/10 | Aug 2026 |
| msn.com | 9.04/10 | Aug 2026 |
| lefigaro.fr | 8.95/10 | Aug 2026 |
| quora.com | 8.92/10 | Aug 2026 |
| huffpost.com | 8.91/10 | Aug 2026 |
| claude.ai | 8.87/10 | Aug 2026 |
| mashable.com | 8.87/10 | Aug 2026 |
| francetvinfo.fr | 8.87/10 | Aug 2026 |
| smh.com.au | 8.86/10 | Aug 2026 |
| wikihow.com | 8.81/10 | Aug 2026 |
| pcmag.com | 8.80/10 | Aug 2026 |
| technologyreview.com | 8.78/10 | Aug 2026 |
| amazon.it | 8.78/10 | Aug 2026 |
| mirror.co.uk | 8.73/10 | Aug 2026 |
| rfi.fr | 8.73/10 | Aug 2026 |
| metro.co.uk | 8.72/10 | Aug 2026 |
| ctvnews.ca | 8.72/10 | Aug 2026 |
| mainichi.jp | 8.69/10 | Aug 2026 |
| france24.com | 8.68/10 | Aug 2026 |
| nikkeibp.co.jp | 8.68/10 | Aug 2026 |
| express.co.uk | 8.65/10 | Aug 2026 |
| nhk.jp | 8.64/10 | Aug 2026 |
| launchpad.net | 8.63/10 | Aug 2026 |
| vermont.gov | 8.63/10 | Aug 2026 |
| fedoraproject.org | 8.62/10 | Aug 2026 |
| adage.com | 8.62/10 | Aug 2026 |
| uic.edu | 8.61/10 | Aug 2026 |
| francebleu.fr | 8.61/10 | Aug 2026 |
Sites whose robots.txt names OAI-SearchBot (or a legacy
alias) with a full block, ranked by Open Page Rank. Sites blocking every crawler with a blanket
* rule are counted above but not listed here. Since dates start at our first
observation of the rule.
Restricting some paths
- linkedin.com
- flickr.com
- hbr.org
- jotform.com
- usatoday.com
- tripadvisor.com
- lnkd.in
- iubenda.com
- nbcnews.com
- investopedia.com
- netflix.com
- academia.edu
Crawler first seen in the wild around Jul 2024. How to block it and what each state means is on the methodology page.