HN
Hacker News
Top
New
Best
Ask
Show
Jobs
Timeline of Diffusion Language Models | Hacker News Reader
Timeline of Diffusion Language Models
github.com
1 point
·
tilt
·
·
1 comment
Open article
Save
View on HN
storystarling
·
I'm curious what the actual inference unit economics look like compared to standard autoregressive models. Parallel decoding helps with latency, but does the total compute cost per token make it viable for production workloads yet?
Reply on news.ycombinator.com