Hacker Timesnew | past | comments | ask | show | jobs | submitlogin

+1 using llama.cpp Vulkan releases with the Qwen models - runs much better than the ROCm releases.

I'll have to give the preserve_thinking a shot.



Thanks for sharing have been running ROCm primarily with Qwen 3.6 and Qwen Coder, on the runs much better statement is that a stability, performance or other capability your experiencing?




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: