Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

Better than Qwen3.6-35B-A3B-8bit ?

When I tried glm found it way way slower (omlx as runtime)



Yes way better. We host both and while qwen3.6 is over 100tps we usually can do glm around that too.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: