7 ms·
This isn't quantized, right? Just a smaller context?
by dgritsko 2mo ago
This isn't quantized, right? Just a smaller context?
- DSingularity 2mo agoIts 256k context window. Quantization is orthogonal. We cant really tell directly so it could be quantized.
- gpugreg 2mo agoThe model is already natively MXFP4-quantized during training, so there is no quality loss.