Comment by faangguyindia
7 hours ago
Diffusion is already being used in Drafter in many LLMs.
many people are running Qwen 3.8 27b on TPU at 130tk/s for free on Kaggle TPUs:
https://www.reddit.com/r/Qwen_AI/comments/1w6gv32/qwen3827b_...
I wonder if we are going to see boxes appear soon, which can run these models for dirt cheap.
No comments yet
Contribute on Hacker News ↗