5 ms·
It actually has vision + tool_use even the 27B param model. The community tries to produce even a MoE version of Qwen3.8 now, because that'd allow to run the fu
by ALLTaken 27d ago
It actually has vision + tool_use even the 27B param model. The community tries to produce even a MoE version of Qwen3.8 now, because that'd allow to run the full model with some experts being pruned like with Ornith 1.5 35B A3B.
You can get this to run on 24GB Ram: https://huggingface.co/baa-ai/Qwen3.8-27B-RAM-24GB-MLX https://huggingface.co/baa-ai/Qwen3.8-27B-RAM-24GB-MLX
I found the FULL Qwen3.8-2.4T-A95B-MLX-reap50-3bit, but it's 540GB. So, I can't run it on my tiny laptop, but hope that the community finds ways to bring the size down and memory requirements too =)))
- ALLTaken 27d agoOMG I found it!! Research preview: Whittle MoE 27B (A18B): a mixture of experts rescued by its routers https://huggingface.co/logic65/Qwen3.8-Whittle-MoE-27B-A17.8B https://huggingface.co/logic65/Qwen3.8-Whittle-MoE-27B-A17.8...