Hi, I’m building a personal website and I don’t want it to be used to train AI. In my robots.txt file I blocked:

  • ChatGPT-User
  • GPTBot
  • Google-Extended
  • FacebookBot

What bots should I also add? Are there any other ways to block AI bots?

IMPORTANT: I don’t want to block search engine crawlers, only bots that are used to train AI.

  • wagoner@infosec.pub
    link
    fedilink
    arrow-up
    1
    ·
    8 months ago

    So that leaves two options then. Leave the front door wide open, don’t bother with any locks. Or shut down the web site. I’m for at least closing the door with the right robots.txt

    • jlow (he/him)@beehaw.org
      link
      fedilink
      arrow-up
      1
      ·
      8 months ago

      The analogy should be either having the door open or having the door open but putting a note on the door saying to please not steal anything. I’m not saying you shouldn’t do it, I just don’t think it’s gonna do anything, so I’m not going to bother.