4 ms·
The author claims 61.0% on WinoGrande vis-a-vis GPT-4's 87.5%.
by macrolocal 4y ago
The author claims 61.0% on WinoGrande vis-a-vis GPT-4's 87.5%.
- pffft8888 4y ago"you can fine-tune RWKV into a non-parallelizable RNN (then you can use outputs of later layers of the previous token) if you want extra performance." Is that 61% using the non-parallelizable RNN mode or the standard mode? I wonder if it's the latter. This new model may be a viable alternative to ChatGPT, which is not only closed sourced but can be shut down in the future just as they did with the older text-davinci models. Plus, the alignement and safety has rendered ChatGPT useless for helping with areas such as critical analysis of social issues (that go against the aligned views) and any and all critical thinking that goes against the aligned views of those who own and program ChatGPT. This could a viable free (as in freedom) alternative.
- macrolocal 4y agoI think the Cambrian explosion is just beginning.
- mach1ne 4y agoI hope not but day by day it seems more likely. If text-generating LLMs can reach superhuman cognition they will so so in a matter of a few years. At that point a Waluigi prompt will be like arming a virtual nuclear missile.
- macrolocal 4y agoNuance: computers have been accumulating superhuman cognitions for half a century. But most people are bad at recognizing intelligence they don't relate to.
- MaxikCZ 4y agoI can't seem to find it in GitHub repo, do you know the value for ChatGPT before it switched to GPT-4?
- macrolocal 4y agoHere are a few benchmarks: https://paperswithcode.com/sota/common-sense-reasoning-on-winogrande https://paperswithcode.com/sota/common-sense-reasoning-on-wi...
- akavi 4y agoHow’d GPT-3/3.5-turbo do?
- cyanf 4y agoLooks like 81.6%. macrolocal linked this below: https://paperswithcode.com/sota/common-sense-reasoning-on-winogrande https://paperswithcode.com/sota/common-sense-reasoning-on-wi...