8 ms·
Thanks! I don't have any knowledge of running models locally. I assume it would not be able to handle an unquantized Qwen3.6-35B or is it irrelevant as you alm
by expedited123 1mo ago
Thanks! I don't have any knowledge of running models locally.
I assume it would not be able to handle an unquantized Qwen3.6-35B or is it irrelevant as you almost always would want to run a quantized version of the model on consumer hardware?
- seanmcdirmid 1mo agonot parent, but 4-bit quantization is generally consider a good trade off for speed/performance, so you might use it even when you aren't on consumer hardware, but definitely when you are on consumer hardware.