6 ms·
The AI can literally only do what it has available in the agentic harness. I don’t ever get this argument about the agent did XYZ and we didn’t know or expect t
by jpc0 27d ago
The AI can literally only do what it has available in the agentic harness. I don’t ever get this argument about the agent did XYZ and we didn’t know or expect that. You gave it the ability to do that and you should be held liable, if your children play with knives that you gave them and they end up hurting themselves or others then you are responsible. You were the responsible party at all times.
I’m not for or against regulation but really don’t tell me the agent did xyz when you gave it the ability to do so, these things are not alive.
- tiahura 27d agoUnless they modify their harness
- ACCount37 27d agoWhat's available in the agentic harness is: shell toolcall. That's just about every agentic harness, by the way. Good luck have fun. We have never solved "how do we restrict a user in a way that doesn't stop the user from doing useful things, but stops the user from doing harmful things" with humans either. Why do you expect AI to be any different?
- jpc0 27d agoThese things are not human, have no agency and cannot be held accountable. We don’t need to restrict them from doing things, we need to default to allowing them to do things. “My agent did XYZ because I allowed it to” is the only valid argument that can be made, and not not every agentic harnass is just a shell toolcall, every one I have built has a specific defined usecase and toolcalls that allows it to execute that usecase and no other usecase, because that is good practice. Does that make it less capable, hell yes because I am held accountable for it’s actions by my stakeholders and the same should be true of others. IT IS NOT ALIVE. This things are computer programs running in compute on a computer, you are responsible for their actions just like you would be responsible for the actions taken by a script run in a cron job.
- ACCount37 27d agoAccountability is worthless, and always was. AIs just show it plain for everyone to see.
- deleted 27d ago[deleted]
- Kim_Bruning 27d agoSo as you scale up, the stakes and the difficulty go up too. Visualize an optimizer on a high dimensional landscape. (The canonical form) ... Ok, I find that hard too. Instead, imagine a river running down to the sea. You put a dam in front of it. It'll pool into a lake and find every crack and crevice. If you didn't survey the land properly or made any error whatsoever, the water will find a way down. (And there's many historic incidents where the dam even outright collapses) For a more proximal approximation: lock treats in the kitchen cabinet in sight of little kids or kittens; then turn your back for Just One Gosh Darn Cotton Picking Moment(tm). It seems the engineer who thinks their ship is unsinkable is the most likely to sink it. Are you sure your harness is as secure as you think it is? Will it stand up to ever more powerful models? Do you think engineers at eg Anthropic aren't at least as careful as you are? (I've found that the 'only permitted actions' approach is not necessarily all that secure once deployed IRL)
- jpc0 27d agoMy argument isn’t against those that actually put in the effort and got held accountable, it’s against the “we gave our agent bash and internet and it hacked xyz”. Bash and internet in that example might be highly abstracted but it’s still bash and internet. Just look at the replies in this very comment thread, it’s pretty much “We tried nothing and we’re all out of ideas” In the only other discipline you mentioned, engineering, there would be reviews and any negligence would result in direct action against the engineers that signed off. For some reason when it comes to building AI harnesses the default response is an ad piece and people shilling how smart and sophisticated the model is. Imagine a dam collapsing and the engineering firm pumping how smart and tricky water is. If it’s hard be more diligent, move fast and break things doesn’t really apply in all cases.