3 ms·
You’re a fabric of stitched together organic molecules that selects the best response to electromagnetic and auditory stimulus.
by burrows 4y ago
You’re a fabric of stitched together organic molecules that selects the best response to electromagnetic and auditory stimulus.
- tsimionescu 4y ago> You’re a fabric of stitched together organic molecules that selects the best response to electromagnetic and auditory stimulus. That's true in some sense, but definitely not if you apply it exclusively to language. The "response" you select is a step in a plan to achieve a goal. Part of that plan may be to utter a sentence, or it may be to stop, duck and roll. If the plan is to utter a sentence, the way you generate the sentence is NOT to emit the most likely token to fit all of the speech input you ever heard before.
- burrows 4y ago> The "response" you select is a step in a plan to achieve a goal. Is it? Or are you just a bird trying to explain how flying works by dropping poop on my head? > If the plan is to utter a sentence, the way you generate the sentence is NOT to emit the most likely token to fit all of the speech input you ever heard before. This is completely incoherent to me.
- tsimionescu 4y agoSay my internal sensors tell my brain I need water. I look around and notice a bottle of water close to another person. I decide to utter the sentence "Give me a glass of water" to try to induce that person to help quench my thirst. If this doesn't work, I will say other things ("Please?"), or get up and pour some water for myself. I will do this even if I have never before heard anyone utter this sentence, or anything similar to it - except for the words themselves, which I need to have learned from hearing them spoken once or twice in my early life. This same kind of planning is visible in any animal you care to study long enough - even in insects, possibly even in jelly-fish. It doesn't exist to even a shallow degree in a conversation with GPT-3. > > If the plan is to utter a sentence, the way you generate the sentence is NOT to emit the most likely token to fit all of the speech input you ever heard before. > This is completely incoherent to me. The way LMs generate sentences is to keep picking the most likely next token (letter) that matches the prompt + the text they generated so far, up to some depth. That is, the basic step that happens when you give a prompt to LaMDA, say, "Do you want a glass of water?" is that it will predict the most likely token is "Y". Then, you prompt it again with "Do you want a glass of water? Y", and it will predict "e". "Do you want a glass of water? Ye" -> "s". "Do you want a glass of water? Yes" -> ".". So, the program around LaMDA will show you the output "Yes.". (there are more steps after this, and they way it decides that it generated enough tokens is not trivial, and tokens are not exactly single letters etc; but this is the basic way it actually works). The likelihood function is based on all of the text it has gone through in the one-time training step. This is definitely NOT the way humans output language.
- raffraffraff 4y agoIf you're saying that there's no difference between (i) AI that is good enough to convince you that it's a real person and (ii) a real person, then I strongly disagree. But if you're saying that one day we might get past the "faking it" stage, because in the end everything is made out of atoms... maybe. I'm not sure advanced human societies will survive long enough to get there.