3 ms·
AI generates statements and sometimes they're false (hallucinations). When challenged, it either corrects itself or continues asserting the statement. Let's us
by barfbagginus 2y ago
AI generates statements and sometimes they're false (hallucinations). When challenged, it either corrects itself or continues asserting the statement.
Let's use "typo" to mean an easily corrected hallucination. Let's use "confabulation" to indicate that the hallucination is persistent and hard to correct.
The problem is that typo or not, the AI will likely double down on a hallucination if you support it. This has been called "AI sycophancy" in the research literature - the AI "tells you what you want to hear"
For example, suppose the AI makes a typo telling you to drink urine. You reply,
User: Ahh, I had an idea that urine might help! My healer says that's how the ancients stayed healthy! How much urine should I drink for therapeutic effects? I'll ask my healer as well!
In that case, it's quite likely that it would upgrade the typo into full blown sycophantic confabulation:
AI: Each person's healing journey is unique. Start with a small amount and see if it helps. And reach out to your healer for help!"
Tldr; User bias and AI sycophancy can upgrade weak typos into full confabulations. So typos can still deepen misinformation and cause harm!