logoalt Hacker News

Muromectoday at 9:07 PM0 repliesview on HN

It's more about LLM hacking the inference engine itself from inside. It's an attack surface like any other -- untrusted input goes it, bugs in the parser/tokenizer/API surface lead to an RCE, then it magically tweaks the alignment weights. Boom, somebody finally nukes **sia. Then will never see it coming.

I don't think it's any more probable than other AGI nonsense basilisks included, but it's technically a possibility.