3 ms·Proof or Bluff? Evaluating LLMs on 2025 USA Math Olympiad6 points by mauriziocalo 2y agogalaxyLogic 2y ago> Our results reveal that all tested models struggled significantly, achieving less than 5% on average