HN
Hacker News
Top
New
Best
Ask
Show
Jobs
Comment by dgritsko | Hacker News Reader
Full thread
dgritsko
·
This isn't quantized, right? Just a smaller context?
View on HN
DSingularity
·
Its 256k context window. Quantization is orthogonal. We cant really tell directly so it could be quantized.
gpugreg
·
The model is already natively MXFP4-quantized during training, so there is no quality loss.
Reply on news.ycombinator.com