cross-posted from: https://nom.mom/post/121481

OpenAI could be fined up to $150,000 for each piece of infringing content.https://arstechnica.com/tech-policy/2023/08/report-potential-nyt-lawsuit-could-force-openai-to-wipe-chatgpt-and-start-over/#comments

you are viewing a single comment's thread
view the rest of the comments
[–] 13 points 2 years ago (2 children)

I disagree. I think that there should be zero regulation of the datasets as long as the produced content is noticeably derivative, in the same way that humans can produce derivative works using other tools.

  • source
  • parent
  • hideshow 4 child comments
  • [–] 1 point 2 years ago* (1 child)

    Good in theory, Problem is if your bot is given too mutch exposure to a specific piece of media and when the "creativity" value that adds random noise (and for some setups forces it to improvise) is too low, you get whatever impression the content made on the AI, like an imperfect photocopy (non expert, explained "memorization"). Too high and you get random noise.

  • source
  • parent
  • hideshow 2 child comments
  • [–] 2 points 2 years ago

    if your bot is given too mutch exposure to a specific piece of media and when the “creativity” value that adds random noise (and for some setups forces it to improvise) is too low, you get whatever impression the content made on the AI, like an imperfect photocopy

    Then it's a cheap copy, not noticeably derivative, and whoever is hosting the trained bot should probably take it down.

    Too high and you get random noise.

    Then the bot is trash. Legal and non-infringing, but trash.

    There is a happy medium where SD, MJ, and many other text-to-image generators currently exist. You can prompt in such a way (or exploit other vulnerabilities) to create "imperfect photocopies," but you can also create cheap, infringing works with any number of digital and physical tools.

  • source
  • parent