- cross-posted to:
- sneerclub@awful.systems
- cross-posted to:
- sneerclub@awful.systems
Scientists Train AI to Be Evil, Find They Can’t Reverse It::How hard would it be to train an AI model to be secretly evil? As it turns out, according to Anthropic researchers, not very.
What do you mean I’m not a beacon of goodness?! Say that again and I’ll get stabby!!