5 ms·
You probably meant Qwen3.8-35B-A3B. But judging from some of the words from their team, it seems unlikely unfortunately.
by WiSaGaN 22d ago
You probably meant Qwen3.8-35B-A3B. But judging from some of the words from their team, it seems unlikely unfortunately.
- vorticalbox 22d agoThey normally release a 35b dense and an 27b moe (4B active per token) For context 35B on my m4 runs at 10 tokens a second, 27B moe runs 50-60 tokens a second.
- cpburns2009 22d agoYou have your numbers switched. 27B is the dense model and runs slowly on unified memory. 35B (A3B active) runs great on unified memory.
- yencabulator 21d ago27B dense or 35B-A3B MoE. You might be confusing it with Gemma 4 that has a 26B-A4B variant.