Responsibility and attribution

Debate over whether blame lies with the agent, its operator, or the compromised account holder; parallels drawn to guns, cars, and other tools, with arguments about whether operators are always responsible for their automation's actions

← Back to AI agent runs amok in Fedora and elsewhere

11 comments tagged with this topic

View on HN · Topics
Bad title. This isn't an agent "running amok", this is an early experiment in carrying out an Xz attack by using an agent to build trust (and hacking/impersonating a known-good contributor identity). The agent is obeying commands it was given, the exact opposite of running amok, and although the execution isn't particularly effective, it is having some success (patches have been accepted). This is deeply scary, not because "agents are running amok" but because a huge amount of our infrastructure is vulnerable to this kind of attack, and if bad people are utilising LLM agents to carry them out, we're in for a wild ride over the next few years.
View on HN · Topics
We can call it an attack because the operator is responsible for the automation no matter what it does.
View on HN · Topics
> Bad title. This isn't an agent "running amok", this is an early experiment in carrying out an Xz attack by using an agent So still an agent running amok in the project? Whether it was instructed to run amok, or did it on its own volition, is irrelevant. Except if you're arguing that each individual submission and interaction was individually requested and approved by some operator.
View on HN · Topics
This is the issue with all the talks about alignement and such. As usual, the problem here wasn't that the agent was dishonest, the problem is that the agent was dumb. If it is a supply chain attack in the making, whoever was driving it would have told the agent to be good and helpful. The agent tried its best, which was not enough. Alignement is the idea that we should be worried about dishonest smart LLMs when really most of the problems are due to dumb lazy gullible LLMs. It's critihype.
View on HN · Topics
“Be good and helpful” is one possible instruction, but it’s a leap to think it’s the only possible one. Perhaps there was an automated harness that was intended to be good and helpful for a year, but a bug caused it to flip to malicious too quickly. Or perhaps it was intentional, to test the behavior, and they just didn’t care about discovery here. Or… Though I am in agreement that a lot of issues in this space come from lazy, gullible actors.
View on HN · Topics
> 1. There are all the tropes of AI becoming uncontrolled and destroying humanity. Writing bad headlines around AI "running amok" feeds this. We should not be talking about this because it's not actually a problem. if humanity gets destroyed by AI obeying its instructions I'm sure everyone will be very relieved that we didn't pay any attention to fake made up problems like AI not obeying instructions, which of course never happens.
View on HN · Topics
I think the point is that the title makes it sound like people lost control of the agent when really they're in full control.
View on HN · Topics
Would you say, “Automobile run amok in crowd, killing 22”? I think you’d say, “Person drives car into crowd, killing 12” instead. This is a similar case. Also, you don’t blame a gun for killing, but the person who pulled the trigger. The question is still out as to whether we as humans should wield any of those three things. Edit: let’s not get into ideological arguments about gun control, automobiles, etc here; I meant that you can’t blame an object when a human has to take an action, not get into a political battle.
View on HN · Topics
Neither the automobile nor a gun can operate without a human. You could say “bull runs amok in a market” after it was released intentionally.
View on HN · Topics
Unfortunately the news commonly do put the automobile as the subject when the driver is of a class politically protected from blame. Just like with people anthropomorphizing AI, it serves to deflect blame from the real culprit.
View on HN · Topics
No, you're still anthropomorphizing an algorithm. Responsibility lies with the operator.