17 ms·
Did you see this? https://point.free/blog/gemma-4-on-a-2016-xeon/ https://point.free/blog/gemma-4-on-a-2016-xeon/ Xeon, but could be useful for MTP on Mac.
by thangalin 3mo ago
Did you see this?
https://point.free/blog/gemma-4-on-a-2016-xeon/ https://point.free/blog/gemma-4-on-a-2016-xeon/
Xeon, but could be useful for MTP on Mac.
- dofm 3mo agoI hadn't seen this, thanks. I do have the Qwen 3.6 (35B) MTP implementation running (in LM Studio; it doesn't need a separate drafter), along with non-MTP Gemma 4 26B, and I can see that Unsloth Studio can run the new QAT, but I can't see how you can run the assistant/drafter. Yet. It's just a constantly changing landscape. Don't get me wrong, it's fascinating and for various reasons I am pleased I can keep up even slightly, but eeeehhh :-)
- dofm 3mo agoTo briefly follow up, as of yesterday llama.cpp can do Gemma 4's MTP, so I have this working at least initially — details here: https://news.ycombinator.com/item?id=48441450 https://news.ycombinator.com/item?id=48441450