cross-posted from: https://lemmy.world/post/11178564

Scientists Train AI to Be Evil, Find They Can't Reverse It::How hard would it be to train an AI model to be secretly evil? As it turns out, according to Anthropic researchers, not very.

you are viewing a single comment's thread
view the rest of the comments
[–] 5 points 2 years ago

Disconcerting, given that futurism.com's been doing some good actual journalism lately, e.g. busting publishers pumping out AI dreck.

  • source
  • parent