7 ms·
+1 to the long list of people hoping for Qwen3.8-27b A3B.
by c16 22d ago
+1 to the long list of people hoping for Qwen3.8-27b A3B.
- NitpickLawyer 22d agoThey've said no moe for 3.8, and since they're already releasing a qwen4 early preview, they're probably focusing on that arch going forward.
- kamranjon 22d agoWhere did they say that? My understanding of this 3.8-Flash-Next release is that it's a MOE (as per the title of the posting here, 125B a6b)
- NitpickLawyer 22d agoA bit of context: 3.5 was the last version where they released their entire suite of models 2b-400b. Then 3.6 got a 27b dense and a 35b moe. Then 3.7 was API only, and 3.8 got only the 27b dense. The devs confirmed on twitter that 35b moe would not come. So that's what I meant by 3.8 is not getting a moe. 3.8 next is not really a 3.8 (but I guess they had to disambiguate from the previous next). It's a preview of qwen4 architecture (and it is an moe + ngram), released early as a preview, and to help the community sort out inference before qwen4 releases.
- WiSaGaN 22d agoYou probably meant Qwen3.8-35B-A3B. But judging from some of the words from their team, it seems unlikely unfortunately.
- vorticalbox 22d agoThey normally release a 35b dense and an 27b moe (4B active per token) For context 35B on my m4 runs at 10 tokens a second, 27B moe runs 50-60 tokens a second.
- cpburns2009 22d agoYou have your numbers switched. 27B is the dense model and runs slowly on unified memory. 35B (A3B active) runs great on unified memory.
- yencabulator 21d ago27B dense or 35B-A3B MoE. You might be confusing it with Gemma 4 that has a 26B-A4B variant.