Tried the recent Tencent Hy3 and it produced a working cube, but plain visuals.
GLM still wins by a landslide
* I ran all agents in their web version with their default settings, I don't remember now if they were set to deep thinking or what but that's the result they produced by default.
GLM-5.2 just keeps blowing me away. I pretty much use it in the mode I used to use Opus for where I have it drive a lot of the reasoning with DeepSeek or other models for the smaller subtasks/subagents.
GLM still wins by a landslide
* I ran all agents in their web version with their default settings, I don't remember now if they were set to deep thinking or what but that's the result they produced by default.