logoalt Hacker News

0xDEAFBEADyesterday at 8:24 AM1 replyview on HN

It's an alignment problem in the sense that it demonstrates the principle that today's AI systems cannot be trusted to reliably work towards the goals of their users. A small-scale alignment failure and a large-scale alignment failure are the same fundamental type of failure. Typically, large disasters come after smaller disasters which foreshadowed the disaster mechanism, but weren't taken seriously.


Replies

Wowfunhappyyesterday at 11:27 AM

I think there is a meaningful difference in kind between an AI that makes a mistake and an AI that is actively malicious.

show 1 reply