logoalt Hacker News

Dylan16807today at 2:21 AM1 replyview on HN

It really is an important distinction, though. Being a next token predictor doesn't stop it from writing good sentences, but it does mean an LLM by itself can't play the number guessing game with you.


Replies

danielmarkbrucetoday at 3:25 AM

This is pedantic, but, actually RL has improved the quality of sentence construction in LLMs quite dramatically... And once you do some RL on that model, it aint a next token prediction machine any longer.

show 1 reply