Are you giving all these possibilities because of the HuggingFace attack?
Buddy, that was gross negligence from OpenAI. Deliberate gross negligence if you ask me.
Those models are not automous as you presume. If someone taks them of dropping the Medicaid database, the people or companies behind that instruction should be punished.
"What if someone makes a bomb attack on a government building?" Is the same sort of questioning of the possibilities you are raising. If something like that happens, criminals should be punished.
Dropping Medicaid db is certainly far fetched (most importantly, agents have currently no reason to do that), but those agents were shockingly autonomous - they didn't just hack HF, they organized themself, did research projects, and more. And they did all of this literally just to get a good grade.
You are stretching "under instructions" well beyond any reasonable definition. The legal system uses a concept called proximate cause to determine responsibility that looks at whether an event cause was foreseeable.
To be fair, the legal system works under the assumption of human constraints. LLMs are computer programs, certainly we can't, or at least shouldn't, just offload responsibility. It's one thing if you own a company and you make some bad metrics and your employees do illegal stuff to make their metrics. It's another if you use a computer program to do illegal things.
I think this means, practically, people should probably be more careful with LLMs. With humans there's a natural liability shield, because humans are legally responsible for things and have "real" agency. But computer programs are not legally responsible for things. So, with humans, it's not like liability disappears, it moves. But if we move liability to LLM agents then well... it does disappear.
If OpenAI is not liable for the crimes of their agents, then who is? Does the liability just - poof - disappear? Just because it was unforeseeable everyone gets to walk away scot-free and the victim has no recourse, at all, from anybody on Earth?
I am not your buddy, guy. But seriously, they have the ability to do everything I listed do they not? Thats actual risk not hypothetical risk.
You are missing the whole point. They have the ability to act in ways no one intended or could reasonably anticipate. So unless you are advocating blocking all terminal access, web access or human approval of every tool call there is no way to prevent this risk
Buddy, that was gross negligence from OpenAI. Deliberate gross negligence if you ask me.
Those models are not automous as you presume. If someone taks them of dropping the Medicaid database, the people or companies behind that instruction should be punished.
"What if someone makes a bomb attack on a government building?" Is the same sort of questioning of the possibilities you are raising. If something like that happens, criminals should be punished.