4 ms·
Luna is one of the budget lower-end last generation models. It'd be useful to at least try to verify the present before being bearish about the future. For Open
by meowface 17d ago
Luna is one of the budget lower-end last generation models. It'd be useful to at least try to verify the present before being bearish about the future. For OpenAI, the best publicly available model is GPT-6 Astra with XHigh or Max reasoning, and for Anthropic it's Claude Fable 5.1 with XHigh or Max reasoning.
- XMPPwocky 17d agoout of curiosity, do you think fable would get this right? (I'm not sure myself, and haven't tried yet.)
- zahlman 17d agoElsewhere in the thread there are reports of Astra on xhigh playing at what I would characterize broadly as a competent casual level, at least given occasional prodding (which a human of that skill level would basically only require when trying to play unreasonably quickly). There seems to be a pattern (even after correcting for relative ELO systems that aren't calibrated) of the LLM bots demonstrating stronger play against traditional bots than against humans.