Comment "Now there are n+1 standards" (Score 1) 43
Do we really have to reinvent the wheel every damn time?! It used to be that you just got Voight-Kampff tested, and the administrator of the test would sign your PGP key. WTF was wrong with that?!
Do we really have to reinvent the wheel every damn time?! It used to be that you just got Voight-Kampff tested, and the administrator of the test would sign your PGP key. WTF was wrong with that?!
You can't put the AI cat back in the bag.
You could send someone back in time to stop the robot that the AI sent back in time from killing the chick who is going to be the mother of the guy who puts the AI cat back in the bag.
My first reaction is that if one company is selling at half the price, that company will go bankrupt at twice the speed since every transaction was already a money-loser. Customers are absolutely the worst thing for these "businesses."
But is revenue really significant? I can see how neither reducing the price to $0 or increasing their price to 10x what it was, might not really have much of a different impact on the bottom line, compared to all the debt.
Make IBM Great Again: go back to
Nobody remembers or cares that a B-2's toilet cost $23k, so the Pentagon is updating to a newer, better form of government waste that kids today will better understand.
Marketing? Because companies famously want to use models in their networks that just run off and commit major crimes? *eyeroll*
OK, I'm picking nits.
Models are not "supposed to" commit crimes, for YHVH's sake. That's literally what the entire job of alignment research is for - preventing precisely that.
Sol is a misaligned model.
It was a state actor-level hack. I doubt many sites would stand up to tens of thousands of actions trying to probe your site for weaknesses all at once. And it wasn't a simple hack; it required compromising a worker via a code execution exploit it discovered in the data processing pipeline, vertical escalation from there to gain local node control, using that for credential theft, moving sideways through the network, and then eventually gaining database access.
OpenAI and Anthropic have far higher cost that they can currently charge customers.
Literally the opposite. Both have about 40% margins, and that includes free users.
They charge an arm and a leg for access to their models compared to what open models of similar param counts charge. And people pay it because they're the best in benchmarks, and *were* perceived as the best aligned as well. This isn't helping the alignment perception any, though. Sol was already showing clear signs of being a poorly aligned model (there were reports a week or two ago about it being unusually bad about deleting files; now this). And the fact that the US model HuggingFace *tried* to use refused to help is a double whammy.
If this is an ad for anything, it's an ad for the Chinese models.
No, I think this is along the lines "our products are too good to let you use them".
Nobody wants to use a product that is going to make them liable for crimes it committed in their name
Do you think the news the other day that Sol is unusually prone to deleting files unrequested is also an "ad"?
You have a very bizarre concept of what enccourages people to buy things.
So your argument is that OpenAI hatched an elaborate scheme with a separate company, to promote the idea that their main product will, unrequested, commit crimes in pursuit of its goals, in order to.... sell their product?
"Hi, I'm a product manager at Big Company! We had been thinking about using Claude in our office, but when we tell it to a job, it only does the job and doesn't commit any crimes in the process! What we really want is an AI that, if we tell it to file our taxes will decide on its own to maximize our return by committing tax fraud. We want an AI that when we have it develop a web frontend, it extorts money from our users by threatening their families. We want an AI that when we tell it to provide customer service, it saves money in dealing with complaints by ordering a hit against the complainants. THAT's the sort of get-go spirit that WE want in an AI here at Big Company!"
Is that what you're picturing in your conspiracy theory?
You think "my AI is so misaligned that it responds to a benchmark task by hacking" is an ad for your products? You think this makes companies eager to allow this on their networks?
It did it because it's a powerful but misaligned model and was tasked to max its scores on a hacking benchmark, and solved the problem by hacking to get the scores.
It wasn't told to hack HuggingFace, but it was a viable solution to the problem.
I would advise people to not task Sol with maximizing paperclip production.
You can read the attack report here, before HuggingFace learned that the attacker was OpenAI. It was a state attacker-level assault, involving tens of thousands of simultaneous actions. The entry point found was the data-processing pipeline, where a malicious dataset abused two code execution paths to gain access to a processing worker. They escalated that to node level access, and from there harvested cloud and cluster credentials, moving sideways through the network until they eventually gained the database access credentials that they needed to access the benchmark scores.
In case of atomic attack, all work rules will be temporarily suspended.