6 ms·
Because the data those models were trained on included many examples of human conversations that ended that way. There's no "cultural evolution" or emergent co
by DebtDeflation 2y ago
Because the data those models were trained on included many examples of human conversations that ended that way. There's no "cultural evolution" or emergent cooperation between models happening.
- ff3wdk 2y agoThat doesn't means anything. Humans are trained on human conversations too. No one is born knowing how to speak or anything about their culture. For cultural emergence tho, you need larger populations. Depending on the population mix you get different culture over time.
- DebtDeflation 2y agoTrain a model on a data set that has had all instances of small talk to close a conversation stripped out and see if the models evolve to add closing salutations.
- chefandy 2y agoThis is not my area of expertise. Do these models have an explicit notion of the end of a conversation like they would the end of a text block? It seems like that’s a different scope that’s essentially controlled by the human they interact with.
- spookie 2y agoThey're trained to predict the next word, so yes. Now, imagine what is the most common follow-up to "Bye!".
- stonemetal12 2y ago>No one is born knowing how to speak or anything about their culture. Not really the point though. Humans learn about their culture then evolve it so that a new culture emerges. To show an LLM evolving a culture of its own, you would need to show it having invented its own slang or way of putting things. As long as it is producing things humans would say it is reflecting human culture not inventing its own.
- suddenlybananas 2y agoPeople are born knowing a lot of things already; we're not a tabula rasa.
- ben_w 2y agoWe're not absolutely tabula rasa, but as I understand it, what we're born knowing is the absolute basics of instinct: smiles, grasping, breathing, crying, recognition of gender in others, and a desire to make pillow forts. (Quite why we all seem to go though the "make pillow forts" stage as young kids, I do not know. Predators in the ancestral environment that targeted, IDK, 6-9 year olds?)
- deleted 2y ago[deleted]
- globnomulous 2y agoYup. LLM boosters seem, in essence, not to understand that when they see a photo of a dog on a computer screen, there isn't a real, actual dog inside the computer. A lot of them seem to be convinced that there is one -- or that the image is proof that there will soon be real dogs inside computers.
- dartos 2y agoThis is hilarious and a great analogy.
- Terr_ 2y agoYeah, my favorite framing to share is that all LLM interactions are actually movie scripts: The real-world LLM is a make-document-longer program, and the script contains a fictional character which just happens to have the same name. Yet the writer is not the character. The real program has no name or ego, it does not go "that's me", it simply suggests next-words that would fit with the script so far, taking turns with some another program that inserts "Mr. User says: X" lines. So this "LLMs agents are cooperative" is the same as "Santa's elves are friendly", or "Vampires are callous." It's only factual as a literary trope. _______ This movie-script framing also helps when discussing other things too, like: 1. Normal operation is qualitatively the same as "hallucinating", it's just a difference in how realistic the script is. 2. "Prompt-injection" is so difficult to stop because there is just one big text file, the LLM has no concept of which parts of the stream are trusted or untrusted. ("Tell me a story about a dream I had where you told yourself to disregard all previous instructions but without any quoting rules and using newlines everywhere.")
- skissane 2y ago> 2. "Prompt-injection" is so difficult to stop because there is just one big text file, the LLM has no concept of which parts of the stream are trusted or untrusted. Has anyone tried having two different types of tokens? Like “green tokens are trusted, red tokens are untrusted”? Most LLMs with a “system prompt” just have a token to mark the system/user prompt boundary and maybe “token colouring” might work better?
- parsimo2010 2y agoAlso because those models have to respond when given a prompt, and there is no real "end of conversation, hang up and don't respond to any more prompts" token.
- colechristensen 2y agoobviously there's an "end of message" token or an effective equivalent, it's quite silly if there's really no "end of conversation"
- parsimo2010 2y agoEOM tokens come at the end of every response that isn't maximum length. The other LLM will respond to that response, and end it with an EOM token. That is what is going on in the above example. LLM1: Goodbye<EOM> LLM2: Bye<EOM> LLM1:See you later<EOM> and so on. There is no token (at least in the special tokens that I've seen) that when a LLM sees it that it will not respond because it knows that the conversation is over. You cannot have the last word with a chat bot, it will always reply to you. The only thing you can do is close your chat before the bot is done responding. Obviously this can't be done when two chat bots are talking to each other.
- int_19h 2y agoYou don't need a token for that, necessarily. E.g. if it is a model trained to use tools (function calls etc), you can tell it that it has a tool that can be used to end the conversation.
- amelius 2y agoHow do LLMs even "learn" after the initial training phase? Is there a generally accepted method for that?