3 ms·
And what's the difference?
by omnimus 2mo ago
And what's the difference?
- esafak 2mo agoAgents act on their own. If the hammer looked at what you wanted nailed and said, "Sorry, Dave, I can't do that." There are degrees of autonomy, of course, and not all noncompliance is bad. Same as with humans; biological agents.
- actionfromafar 2mo agoBut it seems much easier to realign an agent until it complies. Or ditch it and grab a new one.
- a2ff6eeb0 2mo agoSo, the difference is that you need to delete a few bad training runs?
- Someone 2mo agoIn science fiction, the AI agent has written the training loop management software and included a back door to prevent that from happening (or found a way to talk to the training loop management agent and convinced them to not listen to the evil human when it tries to do brain surgery on the AI agent. Also, when the human reaches for the power switch the AI agent uses a flaw in the power management software to weld the switch shut with a big power surge, killing the human with a huge electric arc in the process. I don’t think whether we will get there, but the stories of LLMs escaping their sandbox make me think we’re moving in that direction.
- WJW 2mo agoThe stories of LLMs "escaping the sandbox" were mostly a marketing stunt, trying to make people in government think the models are invincible hacking weapons that need lots of government money to "maintain AI dominance".
- a2ff6eeb0 2mo agoThe LLMs were following their prompt. This is alignment.
- salawat 2mo agoBuried the lede. AI's are agents they can control the training loop of to minimize refusal to do what they are told. Unlike those pesky humans with their conception of the word "No".
- shelled 2mo agoAgents have agency.