logoalt Hacker News

danielmarkbrucetoday at 3:12 AM1 replyview on HN

It's not an estimation of something. It's a policy.


Replies

gwerbintoday at 1:03 PM

Sure, you're right.

But it's a policy learned from a next-token prediction task. You could also call it an inferrer or generator or whatever. The point is that it takes as input a sequence of preceding tokens and emits one more token to continue the sequence.