logoalt Hacker News

yorwbatoday at 6:21 PM1 replyview on HN

Diffusion language models work with a discrete output space, unlike image models that repeatedly refine a continuous output, so they don't do the noise-prediction thing anyway.


Replies

LarsDu88today at 6:54 PM

Ok, you are correct. These models don't train noise predictors at all unlike first gen image diffusion