Hacker News new | ask | show | jobs
by argee 23 days ago
This repros nearly 100% of the time on most LLMs, even the most advanced ones: https://share.gemini.google/u9NwYu7lbgxe
1 comments

n=1 but I gave this to Sonnet 5 medium effort (free model) and it had no trouble with it
Try it without "reasoning". As you can see in my example (and GP), it meanders to correctness eventually after emphatically being wrong, and most reasoning modes hide that from you.

If LLMs worked the way people want to believe they do, there’d be no reason to start in the wrong place — a computer should have the facts!