logoalt Hacker News

ethbr1 • yesterday at 12:36 PM • 2 replies • view on HN

My perspective is that the addition of thinking loops to models allows sufficiently advanced ones to approximate world models.

Incredibly inefficiently because of the recursive loops ("Wait, the object is on the table. I should think about this more deeply..."), and likely instantly surpassed by large world models if/when those are shipped, but effectively enough vs non-thinking models.


Replies

red75prime • yesterday at 1:42 PM

LeCun calling them "world models" gives a high-level description of the desired functionality. They are Joint Embedding Predictive Architectures (with SIGReg). They might produce more useful world models, but it's yet to be seen.

hyperman1 • yesterday at 1:14 PM

This sounds like a human trying to reason about quantum mechanics. We als simplify to newtonian for day to day tasks.

➕ show 1 reply