Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

This repros nearly 100% of the time on most LLMs, even the most advanced ones: https://share.gemini.google/u9NwYu7lbgxe


n=1 but I gave this to Sonnet 5 medium effort (free model) and it had no trouble with it


Try it without "reasoning". As you can see in my example (and GP), it meanders to correctness eventually after emphatically being wrong, and most reasoning modes hide that from you.

If LLMs worked the way people want to believe they do, there’d be no reason to start in the wrong place — a computer should have the facts!




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: