NYT looks like it's updated it's robots.txt file to disallow the Open AI bot from scraping it's data. Pretty interested to see if they just update their user agent string or if they'll respect it

you are viewing a single comment's thread
view the rest of the comments
[–] 36 points 3 years ago (1 child)

Updating user agent doesn't natter unless NYT is actively blocking that, too. Updating robots.txt is purely a "gentleman's agreement" that OpenAI will respect it. OpenAI would be dumb to ignore it, hat all said, because it'd trigger the lawyer shenanigans to ensue.

  • source
  • hideshow 2 child comments
  • [–] 12 points 3 years ago (1 child)

    NYT is already considering a lawsuit against OpenAI. So, not just dumb but arrogantly stupid when the lawyers are already in the room.

  • source
  • parent
  • hideshow 2 child comments
  • [–] 9 points 3 years ago

    The burden of proof will fall upon the NYT and it will be extremely difficult to prove OpenAI is culpable for any infringement that it's end users perform.

    It's new territory and will be expensive, but NYT is old money and has the liquidity to burn cash all day.

  • source
  • parent