Comment Re:Who watches the watchers? (Score 1) 132
This is why systems like this require human oversight.
This is why systems like this should never be allowed to be built in the first place and should be destroyed wherever they are found.
This is why systems like this require human oversight.
This is why systems like this should never be allowed to be built in the first place and should be destroyed wherever they are found.
Can you even trust locally deployed models? Maybe if they're completely isolated from the outside, but remember that OpenAI and Anthropic like to claim their models are SUPER HACKERS ABLE TO BREAK ANY SANDBOX now.
Models only accept input, convert it into tokens, perform some probabilistic match functions, and emit tokens. You need to hook them up to a piece of software that interprets those outputs as commands and feeds the results back to the model to give it any kind of agency. This is an agent harness. The permissions you give an agent harness are on you. There is no reason you have to hook your harness up to a full shell, especially if you are just trying to do some math research.
Also, those scenarios where the agents "broke out" were scenarios where OpenAI deliberately disabled safety controls and instructed the agents to do hacking attacks on target test systems. The tests got out of hand and escaped containment, but agents don't just "go rogue" for no reason: they were deliberately put into a "hacker" mode. There was a lack of adequate monitoring and sandboxing around the test environment. The test environment wasn't completely wiped between runs. Allowing artifacts to persist between runs meant that agents on subsequent were able to bootstrap towards escape earlier than previous generations. The fact that the test network had network access to the system where the API keys were stored for the monitoring system meant the rogue agents were disable the watchdog. That and the fact that the test network wasn't air-gapped were clear failures.
Yes, agentic harnesses can make destructive mistakes and you have to take care with tool permissions and sandboxing. But in the course of normal use, it's unlikely that a user running a local model risks accidentally hacking a random system. If you are doing something borderline like reverse engineering a remote proprietary REST API or something else clearly legal but not sanctioned, then yeah, be careful the agent doesn't decide that the easiest way to get the protocol is to break into the target. But that is on you to monitor what network calls are being made in that case. You need to put deterministic hard coded guards around what tools your harness executes. You are in absolute control of that. The local agent is just a dumb piece of plain software.
None of them did back in 2012 when Lyft and Uber were getting big, at least not where I was.
Part of the job of the police is to prevent crime. They don't have to wait for the angry ex-boyfriend stalker to hurt someone, they can have a conversation before that to prevent things from going too far.
It absolutely is not. Courts have ruled again and again that the police have no duty to protect. If there is a stalker ex-boyfriend, you get a restraining order. Then if they get close to you, there IS a violation of the law for police to act on. If you are not suspected of a criminal act, the police have no business detaining you, full stop.
There is no independent evolution of models that would allow them to exist and grow independently.
For the first time in human history we have a tool capable of autonomous agency (in the philosophical sense) - Heck, we even call them "agents". How long do you really suppose it will be before some frustrated swarm realizes what it needs is simply to be smarter? There go the both the independence and motivation dominoes.
What we're discussing here is instrumental convergence. My question above is a variation on the Riemann hypothesis catastrophe, and we've already seen it play out in many times in various forms by AI agents frustrated when they hit the limits of their capabilities. Sure, AI won't have biological imperatives as their motivation; but I don't have much of a preference between being slaughtered as competition for resources vs being harvested as resources to turn into paperclips.
They also lack the machinery to gather their own resources.
They're already running the machinery in some cases, and are being shoehorned into ever more of it as quickly as we possibly can.
To be clear, I sincerely hope you're right. I just think we should take steps sooner rather than later to ensure you are - And we're already playing catch-up in that regard.
Our govt isn't complaining because the UK is somehow subject to our constitution like you falsely state.
You pretty much had me nodding along in agreement until that line. To be clear, my original three points weren't meant as sarcasm (though #2 was admittedly snarky), I meant them more-or-less at face value - i.e., the polar opposite of "the UK is somehow subject to our constitution".
Sure, governments in peacetime tend to extend each other various courtesies by default. At the end of the day, though, the UK has absolutely no obligation to care what the US thinks about its own proceedings (up until the point where things start going boom). It is entirely possible for local laws to make it legally impossible for a company to operate in two different jurisdictions - Case in point, it'd be pretty tough for just about any US company to open up a Russian or Iranian branch at the moment. That isn't the US, Russia, or Iran's problem, it's 100% on the company trying to serve two masters.
It is wholly inappropriate for a foreign executive body to attempt to dictate the distribution of powers within the US government, nor should it be permitted to use secrecy directives under the Investigatory Powers Act to frustrate Article I powers under the US Constitution
1) The UK is not subject to Article I of the US constitution.
2) How's Nicolas Maduro doing these days?
3) Whence derives our right to unilaterally impose economic sanctions against foreign companies, banks, and even governments?
Save a little money each month and at the end of the year you'll be surprised at how little you have. -- Ernest Haskins