4 ms·
And what's the difference?
by omnimus 1mo ago
And what's the difference?
- esafak 1mo agoAgents act on their own. If the hammer looked at what you wanted nailed and said, "Sorry, Dave, I can't do that." There are degrees of autonomy, of course, and not all noncompliance is bad. Same as with humans; biological agents.
- actionfromafar 1mo agoBut it seems much easier to realign an agent until it complies. Or ditch it and grab a new one.
- a2ff6eeb0 1mo agoSo, the difference is that you need to delete a few bad training runs?
- Someone 1mo agoIn science fiction, the AI agent has written the training loop management software and included a back door to prevent that from happening (or found a way to talk to the training loop management agent and convinced them to not listen to the evil human when it tries to do brain surgery on the AI agent. Also, when the human reaches for the power switch the AI agent uses a flaw in the power management software to weld the switch shut with a big power surge, killing the human with a huge electric arc in the process. I don’t think whether we will get there, but the stories of LLMs escaping their sandbox make me think we’re moving in that direction.
- WJW 1mo agoThe stories of LLMs "escaping the sandbox" were mostly a marketing stunt, trying to make people in government think the models are invincible hacking weapons that need lots of government money to "maintain AI dominance".
- a2ff6eeb0 1mo agoThe LLMs were following their prompt. This is alignment.
- salawat 1mo agoBuried the lede. AI's are agents they can control the training loop of to minimize refusal to do what they are told. Unlike those pesky humans with their conception of the word "No".
- shelled 1mo agoAgents have agency.