logoalt Hacker News

faangguyindiatoday at 4:29 AM0 repliesview on HN

Diffusion is already being used in Drafter in many LLMs.

many people are running Qwen 3.8 27b on TPU at 130tk/s for free on Kaggle TPUs:

https://www.reddit.com/r/Qwen_AI/comments/1w6gv32/qwen3827b_...

I wonder if we are going to see boxes appear soon, which can run these models for dirt cheap.