3 ms·
Unless you look at the actual objective functions for training: primarily, produce human like text, and second, produce answers that "feel good" to a human via
by tensor 2y ago
Unless you look at the actual objective functions for training: primarily, produce human like text, and second, produce answers that "feel good" to a human via reinforcement learning.
Neither function optimizes for an understanding of what truth means, generating confidences, reasoning, let alone higher level objectives like "doing PhD level research".
Given that none of these things are changing I wouldn't expect any major difference in its ability to provide truthful reasoned responses. It make continue to pick up some of these abilities as a coincidental byproduct of learning to generate text, but it will still be a glorified text generator without these changes.
That said maybe they will have some new training objective function, we'll see. Either way these systems are a far cry from human-level when it comes to having actual goals beyond being a toy text generator.