logoalt Hacker News

croemer • yesterday at 7:31 PM • 1 reply • view on HN

Pretty crazy that the model doesn't know that it needs to stop before it hits 128k output tokens. I guess it has no sense of how many tokens in it is? Wouldn't this be possible to work into the architecture?


Replies

simonw • yesterday at 7:40 PM

I think this is a bug. I've not seen this problem from any of the other frontier models.

➕ show 2 replies