Hi, I’m building a personal website and I don’t want it to be used to train AI. In my robots.txt
file I blocked:
- ChatGPT-User
- GPTBot
- Google-Extended
- FacebookBot
What bots should I also add? Are there any other ways to block AI bots?
IMPORTANT: I don’t want to block search engine crawlers, only bots that are used to train AI.
I don’t really understand the reasoning behind doing any of this, they didn’t give a fuck about stealing clearly copyrighted content in the first place, why would they care about you (not OP specifically) begging them not to steal your stuff. (As long as theres no laws about this which afaik there aren’t).
So that leaves two options then. Leave the front door wide open, don’t bother with any locks. Or shut down the web site. I’m for at least closing the door with the right robots.txt