Hacker News new | ask | show | jobs
by Revanche1367 12 days ago
In my experience, Gemini 3.x wasn’t just getting a bad rap, it was significantly worse in practice. It could analyze codebases and report back from a one-shot prompt as good as Claude or Codex but any slightly complex task that carried on for more than a few minutes led to hanging, seemingly infinite loops, and bizarre and nonsensical hallucinations, to the point of being unusable for serious work. The Claude and Codex counterpart models at the time rarely had such issues for the same type and duration of complex work, if at all. To be fair, later Claude especially started having hanging issues as many people noticed but that’s been better recently.