3 ms·
Llama 3.1: Our most capable models to date
- deleted 2y ago[deleted]
- msoad 2y agoMMLU PRO is the benchmark I trust the most. I noticed they are using 5 shots and CoT. Is that true for GPT4 and Sonnet as well?
- sagz 2y ago405B is already being served on WhatsApp! https://ibb.co/kQ2tKX5 https://ibb.co/kQ2tKX5
- ChrisArchitect 2y ago[dupe] More discussion: https://news.ycombinator.com/item?id=41046540 https://news.ycombinator.com/item?id=41046540 https://news.ycombinator.com/item?id=41046773 https://news.ycombinator.com/item?id=41046773