5 ms·
+1 using llama.cpp Vulkan releases with the Qwen models - runs much better than the ROCm releases. I'll have to give the preserve_thinking a shot.
by ndom91 3mo ago
+1 using llama.cpp Vulkan releases with the Qwen models - runs much better than the ROCm releases.
I'll have to give the preserve_thinking a shot.
- jderekw 3mo agoThanks for sharing have been running ROCm primarily with Qwen 3.6 and Qwen Coder, on the runs much better statement is that a stability, performance or other capability your experiencing?