It asks a fetcher not to request a path. Documented, well-behaved crawlers honour that request, and it has never been an access control: the file is public and enforcement happens entirely on the other side. The infrastructure provider behind the most widely deployed managed version of that file says as much about its own newer signalling scheme: content signals “express preferences; they are not technical countermeasures against scraping. Some companies might simply ignore them.”4
The split that decides most outcomes is not training against search, which the retrieval stage covers. It is scheduled crawling against a fetch that happened because a person asked for something, and on that boundary the vendors contradict each other in their own documentation. Every quotation below was checked against the live vendor page on 30 August 2026.