3 ms·
> Just because you explain what you want to a human that doesn't mean they agree or will comply. Humans have innate desires that may conflict with the desires
by trott 2y ago
> Just because you explain what you want to a human that doesn't mean they agree or will comply.
Humans have innate desires that may conflict with the desires of other humans. A language model just looks for ways to continue texts. In doing so, it attempts to extrapolate what humans who authored the training data thought (If we ignore fabrications, which my approach proposes to address also)
> LLMs don't actually understand or reason about anything.
Current Transformer-based, SGD-trained language models fall short of AGI. But better algorithms can change that. And there is no reason to think that you'll see any warning signs in advance. No one said "I'll invent a faster Fourier transform algorithm in 5 years. Prepare yourselves."