logoalt Hacker News

gdiamostoday at 5:21 AM0 repliesview on HN

I’d like to see more of these models.

I’ve been using diffusion Gemma and it is very fast on GPUs in output token/sec.

In the diffusion Gemma whitepaper, they say they could have done better with more time and compute.

Even with those caveats, it is very uses-able as a local model.