7 ms·
That’s not strictly true, an AI for sure can provide a confidence interval for its answer, but the sentiment I feel is generally correct.
by SpeedilyDamage 4y ago
That’s not strictly true, an AI for sure can provide a confidence interval for its answer, but the sentiment I feel is generally correct.
- Retric 4y agoChatGPT can’t provide any score for how correct it’s answer is. Basically everything it says is a lie, some of those lies happen to randomly be true.
- SpeedilyDamage 4y agoIt absolutely can, just check out the OpenAI Azure documentation to understand more: https://learn.microsoft.com/en-us/azure/cognitive-services/openai/reference https://learn.microsoft.com/en-us/azure/cognitive-services/o... Edit: This also might be helpful: https://learn.microsoft.com/en-us/azure/cognitive-services/qnamaker/concepts/confidence-score https://learn.microsoft.com/en-us/azure/cognitive-services/q...
- Retric 4y agoI don’t know what you think that shows, but none of it corresponds to “truth.”
- SpeedilyDamage 4y agoI wasn't trying to suggest it did, just that the concept of "confidence" is a key one in ML.
- Retric 4y agoGotcha. Rereading what I said, “correct” rather than “factual” was poor word choice on my part. From it’s perspective correct is just whatever results in a high score.
- SpeedilyDamage 4y agoYeah, it's well trained by humans who do know what is "factual" (or at least better than ChatGPT does), and the "correctness" score gets used to figure out why ChatGPT gave a "wrong" answer. Over time, the model is trained this way to better align "correct" with "factual", but yeah ChatGPT needs humans to tell it what's "factually" accurate and what's not.
- Retric 4y ago> well trained by humans Yes and no, the goal wasn’t focused on just being a source of truth. They wanted it to be able to write fairy tales, poetry, and song lyrics so creative answers where rewarded even if they weren’t factually accurate. There was an attempt to reduce harmful and deceitful responses, but the team acknowledge the tendency for “plausible-sounding but incorrect or nonsensical answers” calling them hallucinations. Which becomes common when prompts are outside of it’s training set. So yes there is definitely a link between the internal “Correctness” score and “Truth” but they are measuring very different things.
- willhslade 4y agoHow confident is it in it's confidence interval?