logoalt Hacker News

lolakutty • yesterday at 4:33 PM • 2 replies • view on HN

>You could ask an LLM what 1+1 is, and the number of times it says "3" is so small that it makes no sense to worry about it...

I think the disturbing fact is that you can take a frontier model with all the intelligence of humanity, and make it say 1 + 1 = 3, by specifically training for it...

A human with that much knowledge will refuse that attempt. There in lies the difference..


Replies

thaumasiotes • yesterday at 5:45 PM

> A human with that much knowledge will refuse that attempt.

Well, that's not true.

https://www.youtube.com/playlist?list=PLO3a3Ax6Yh6bbtKuxfYBP...

monkpit • yesterday at 4:45 PM

What’s your point though, really? “You can train a model to say things that are objectively wrong”? You can do the same with a human.

➕ show 1 reply