4 ms·
prismML provides a llama.cpp fork which is compatible with the 1 bit models: https://github.com/PrismML-Eng/llama.cpp https://github.com/PrismML-Eng/llama.cpp
by m0do1 6mo ago
prismML provides a llama.cpp fork which is compatible with the 1 bit models:
https://github.com/PrismML-Eng/llama.cpp https://github.com/PrismML-Eng/llama.cpp
After fails with Ollama and main llama.cpp the fork worked on my M5 MBA.
Edit: Typos