cross-posted from: https://lemmy.world/post/23009603
This is horrifying. But, also sort of expected it. Link to the full research paper:
cross-posted from: https://lemmy.world/post/23009603
This is horrifying. But, also sort of expected it. Link to the full research paper:
Not really caught. The devs intentionally connected it to specific systems (like other servers), gave it vague instructions that amounted to “ensure you achieve your goal in the long term at all costs,” and then let it do its thing.
It’s not like it did something it wasn’t instructed to do; it didn’t perform some menial task and then also invent its own secret agenda on the side when nobody was looking.