Hacker Timesnew | past | comments | ask | show | jobs | submitlogin

Qwen 3.5-9b-Q4_K_M.

I have a 5080 too! For me, the key has been dropping Ollama for Llama.cpp, which is not particularly scary to configure anymore and just skyrocketed performance. I download the models with LM Studio, then run them with llama.cpp.



Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: