3 ms·
This is not the consensus among ML researchers. Transformers are showing strong generalisation[1] and their performance continues to surprise us as they scale[2
by moconnor 4y ago
This is not the consensus among ML researchers. Transformers are showing strong generalisation[1] and their performance continues to surprise us as they scale[2].
The Socratic paper is not about “higher intelligence”, it’s about demonstrating useful behaviour purely by connecting several large models via language.
[1] https://arxiv.org/abs/2201.02177 https://arxiv.org/abs/2201.02177
[2] https://arxiv.org/abs/2204.02311 https://arxiv.org/abs/2204.02311