7 ms·
Note: you are probably running a distilled version of R1, which is actually LLama or Qwen further trained on the input/output of R1. The full R1 is huge (~700G
by zbendefy 2y ago
Note: you are probably running a distilled version of R1, which is actually LLama or Qwen further trained on the input/output of R1.
The full R1 is huge (~700GB), altough there are still quantized versions, the smallest one is around 150gb (1.58bit)
- 9cb14c1ec0 2y agoOh, that's interesting. I didn't know that the ollama version wasn't the whole thing.
- postalrat 2y agoollama deepseek-r1:671b is