logoalt Hacker News

mitxelatoday at 2:13 AM1 replyview on HN

Some concrete facts about LLMs are explained by their next token predictor nature. Every time it says "wait, that's wrong." instead of generating the correct thing the first time.


Replies

howunfortunatetoday at 2:17 AM

I think that's relatively emergent too though! BERT never really did that (at least to my recollection), presumably because its training was never sufficient for it to develop corrective reasoning in a chain of thought.

show 1 reply