logoalt Hacker News

demibabs • yesterday at 5:51 PM • 1 reply • view on HN

Is this an example of the bitter lesson? Could diffusion models not be this good with equal scaling/compute as LLMs?


Replies

moojacob • yesterday at 5:57 PM

Well the LLMs are much bigger than diffusion models so I think that's the bitter lesson. You could scale compute for diffusion models though.

➕ show 1 reply