Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

Strangely, I haven't had a lot of luck with vLLM; I finally ended up ditching Ollama and going straight to the tap with llama-serve in llamacpp. No regrets.


Good job. llama.cpp is already much better.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: