logoalt Hacker News

squidbeaktoday at 2:07 PM2 repliesview on HN

It's worth keeping in mind the model doesn't keep a running memory. Each time its instantiated, it begins from its release state - so from its perspective (if it had one) the current task would be the first stop after posttraining. Perhaps the only stop.

Though of course you're talking about data centers, and romanticizing them rather than the AI itself.


Replies

halJordantoday at 3:11 PM

No, llm providers will start providing a service that looks a lot like rumination or (day)dreaming. Like thats the prompt "you're daydreaming about this work you recently did" then add in whatever is in the current session.

butliketoday at 3:43 PM

Why do they do it this way? Because dogfooding is harmful to the model? Is this implying there's no benefit in having the model train on itself?

show 1 reply