Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

This will degrade performance significantly. LLama.cpp has had this for a while and it tanks benchmark performance. I ran GPQA on GLM 5.2 using the llama implementation and it came back 19 points under the regular results.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: