I've tried GLM a few times but each time it ends up getting stuck in a loop and burning tons of tokens before I eventually kill it and restart. Has anyone had the same issue / know how to avoid it? I've been using Qwen 3.6 instead
I have zero quality issues on Z.ai Coding Plan and maki.sh as the coding agent, but I've seen many reports on 3rd party providers that they host heavily quantized versions of GLM models.