3 ms·
> I could write in brainfuck with ai, but I presume, wouldn’t get the same results than if going with python. https://esolang-bench.vercel.app/ https://esolang
by robot-wrangler 5mo ago
> I could write in brainfuck with ai, but I presume, wouldn’t get the same results than if going with python.
https://esolang-bench.vercel.app/ https://esolang-bench.vercel.app/
- _boffin_ 5mo agoand this sums it up right here.
- Tarq0n 5mo agoThe conclusions seem overly broad. Just because these languages are Turing complete doesn't mean they aren't massively hampered by expressiveness and amount of batteries included. To attribute all of this to training data memorization is premature.
- robot-wrangler 5mo agoOh this is a very damning paper. Using simple languages from their definitions alone is a great proxy for studying truly out-of-distribution reasoning. Also just for following simple rules/instructions correctly, because a simple enough language is practically just a grammar. This paper is terrible for anyone who wants to make the case that models can do those things well. To the extent today's AI can reason, add this to the pile of evidence that you definitely need a harness. Counter to what you hear.. that seems true for SOTA and frontier, not just toy models. Lots of people were saying many years ago someone should test exactly this, because it's obvious. Someone at megacorp probably did try and decided not to publish because they thought it was bad optics.