6 ms·Is this similar to fastllm? https://github.com/ztxz16/fastllm https://github.com/ztxz16/fastllmby nogajun 2mo agoIs this similar to fastllm? https://github.com/ztxz16/fastllm https://github.com/ztxz16/fastllmvikmals 2mo agofastllm targets the GPU, while colibri uses CPU inference onlyaliljet 2mo agoI'd be curious about an.option that would allow glm use with a low end GPU like a 2080 ti...