4 ms·
No, we definitely should continue to quiz LLMs on mathematics and absolutely any other topics. Otherwise how do we know and understand limitation of the system?
by deelly 3y ago
No, we definitely should continue to quiz LLMs on mathematics and absolutely any other topics. Otherwise how do we know and understand limitation of the system?
- muzani 3y agoWe should also test its capabilities on cooking steak, flying rockets, and making love. Only then will we know if AI can be superior to humans on all things.
- falcor84 3y agoI know you're being facetious, but I'd absolutely be in favor of having AI benchmarks for any and all of these.
- muzani 3y agoI realized that halfway through typing that too. As well as the absurdity of trying to create AIs that exceed a human in all tasks. On a serious note, most of these AIs are bad at math but good at writing code for calculators. So what you'll be benchmarking is their ability to create and use tools.
- ineedasername 3y agoWell, if there was an API hooked up to a webcam & 6-dof arm that would be an interesting task. (The steak)
- BeFlatXIII 3y agoLovemaking, too.