5 ms·
I run a 2x 3090 rig, but a single 3090 already provides a great experience at a reasonable quant and context size. On a single card I used to run Qwen_Qwen3.6-2
by nsbk 2mo ago
I run a 2x 3090 rig, but a single 3090 already provides a great experience at a reasonable quant and context size. On a single card I used to run Qwen_Qwen3.6-27B-Q4_K_M or similarly quantized 35B MoE at 65536 context size
- xiconfjs 2mo agoyou should look into https://github.com/noonghunna/club-3090 https://github.com/noonghunna/club-3090