3 ms·
This repros nearly 100% of the time on most LLMs, even the most advanced ones: https://share.gemini.google/u9NwYu7lbgxe https://share.gemini.google/u9NwYu7lbgxe
by argee 3mo ago
This repros nearly 100% of the time on most LLMs, even the most advanced ones: https://share.gemini.google/u9NwYu7lbgxe https://share.gemini.google/u9NwYu7lbgxe
- imajoredinecon 3mo agon=1 but I gave this to Sonnet 5 medium effort (free model) and it had no trouble with it
- argee 3mo agoTry it without "reasoning". As you can see in my example (and GP), it meanders to correctness eventually after emphatically being wrong, and most reasoning modes hide that from you. If LLMs worked the way people want to believe they do, there’d be no reason to start in the wrong place — a computer should have the facts!