Forgot your password?
typodupeerror

Comment Re:WTF (Score 1) 105

Nope, they're all "distilled" models. That's what post training is, largely, and they're all using other models to help curate the datasets for this. They're all doing it. It's why all the models appear to be converging in terms of capabilities and personality: they're effectively using the same feed stock and using an increasingly similar set of techniques.

US "frontier" labs are distilling, as well. It's postulated (with good reason) that Opus 5 is, for instance, just a distillation of Fable to a smaller model: it's very architecturally and behaviorally different than prior Opus releases.

The anti-distillation argument is basically the same argument that's been used against Open Source in general over the past 30 odd years. It's anti-competitive and self interested.

Comment "Sometimes illicitly" (Score 1) 105

They say "sometimes illicitly"... OK, so what's the criteria that makes it illicit?

If they can't state that, it's just FUD.

And frankly, that's most of what the anti-AI brigade is spewing: it's a psyop, pure and simple - by both China and the US domestic companies, with the blessing of the US government. People who can't see this aren't paying any attention to what's happening.

Anthropic/OpenAI want restrictions to lock people in and provide them with an out for their gross overspending.
US government and its handlers wants more control in the dissemination of information and intelligence capabilities.
China wants to stifle US industrial efforts.

This is more or less the exact same argument open source has been fighting for the past 30 years. Embrace, Extend, Extinguish. They're working on the 'extinguish'.

It's just a shame that we don't have any genuine "open source" models, it's all just open weight. We need more openness so information isn't stifled by vested interests (in the same way that governments have always done). We should be able to have models which aren't given implicit alignment about specific nationalities and ethnicities and prohibited from making any informed observations about them, in the same way that we should be able to have access to our own history. It would be unconscionable that the government get control of all publishing, or even singular ideological control within corporate hands - yet here we are having that same discussion about AI. It's preposterous.

Comment yeah, right (Score 1) 118

Both of the 2 domestic frontier labs are screaming doomsday about AI, while concurrently asking for government regulation (which keeps getting batted down), and (presumably) funding senators to put bills in place to make model distillation a felony. While they're trying to IPO. "Oh, we really need to slow down the industry".

What the hell is going on here? In what world does this make sense - where a company making AI, and is trying to IPO, is blasting the internet with AI doomerism?

The only conclusion I can walk away with is some combination of these:

1) It's about control, and they're in some sort of arrangement with the government or some third party interests who want (an even bigger) national panopticon.
2) They want to stop domestic competition for their models so they'd have exclusive domain here in the US.
3) They actually want the company to fail, because then they could cry "regulation" on account of competition (back to #2).

Comment Re:Translation: (Score 1) 94

On the contrary, this is a giant psyop to push for regulated market capture by both Anthropic and OpenAI.

The Jacob Coxon thing smells heavily of a psyop: dude who's worked there only a couple months quits, and his X account with zero followers suddenly blows up in a media frenzy. He was/is a nobody. It's sensationalism.

You think they'd hire that many $400k+ engineers and still not know how to properly secure a network? They surely knew the security breaches were happening, unless they simply haven't hired competent people. They literally told the agents to have a breakout: do the needful. The first time it happened - OK, sure, you didn't expect emergent behavior. The second, third times? You were running a science experiment and that was the desired result, to one end or another: whether it was negligent incompetence or intentional for the purposes of further sensationalism, it's hard to say.

They (OpenAI/Anthropic) WANT to slow things down through regulatory control to stifle competition. They do not want foreign models to be accessible. We've got the old crooner Bernie Sanders pushing for 5+ year felonies for distilling models, for fuck sake. This is all about decreasing freedom and enforcing your dependence on closed models.

Translation: we have hit a point where the public backlash has led to untenable increased risks in liability for our criminal actions.

Nah, the "public backlash" is them talking in circles, pushing for industry control.

Comment Re:Boo hoo hoo... (Score 2) 86

They paid those companies for access to the models. They're not stealing the service: they signed up for them and used them. The same way Anthropic and OpenAI do for the other companies.

THere's plenty of evidence that these same companies (OpenAI and Anthropic specifically) also used resources illegally in their training. Not just "it was on the internet and not explicitly licensed for training" but "this is protected IP". I personally don't think that matters, but it matters to their argument, so it's worth bringing up.

You can't do one thing and then complain when someone does a lesser version of that same thing.

Comment Re:And nothing of AI value was lost. (Score 1) 86

You gotta remember the slashdot mindset. It's got a very complex herdmind.

Something I disagree with: troll.
Something that bites the hand that feeds me: overrated
Something that attacks the thing I dislike: insightful

Musk is a fascist, and the others are all cleptocratic capitalists, so they're hated. China is the good communist (but really capitalist because they give us cheap slop) country.

It's all very trite and tiresome and single dimensional.

Comment Not even using the right terms (Score 2) 86

They're not even using the right terms.

They're not distilling US models, they're generating synthetic data to train models to distill their own models.

You know, the exact same thing I do to train local purpose-built open models.

So am I going to have the full weight of the US government at my door, next?

" conducting systematic extraction of proprietary functionalities and capabilities of U.S. AI companies"

How is this in any way meaningfully different than using them to generate code, or perform any other arbitrary set of tasks? It isn't.

This is ridiculous.

Comment Re:Its very puzzling, isn't it? (Score 1) 96

While the energy source of wind and solar are free, building and maintaining the plants is not free.

If all they were building were a single 100MW data center, sure: renewables would be cheaper to build. But they evidently are planning a major campus that will consume most of the plant's 615 MW output over the next 25 years.

A solar installation, in a favorable location, capable of supplying 600MW around the clock (with battery backup) would have to be roughly 1800MW in capacity. It would be among the largest inthe world, on the order of ten thousand acres in size at a build cost of maybe 2 billion. The battery backup system to ensure 99.9% uptime would be 9x the size of the largest li-ion grid storage system ever built, and set you back on the order of 4-5 billion dollars.

While it's probably cheaper to go renewable than build a *new* nuclear plant, if you can reactivate and one for just two billion that looks like a bargain, if you have a use for all that power. Inability to save money by load following is the financial Achilles' heel of this generation of reactors, but if you have a guaranteed customer for most of your output, years in advance, that's as close to an ideal economic case for them as anything could be.

Comment Re:So what? (Score 2) 86

ToS is the strongest argument, but the distillers are not parties to the ToS. They get their data from data brokers. It is possible that data brokers are violating the ToS, but it would be hard to write ToS that precluded running queries for third parties without creating problems for consultancies and other businesses. Even presuming the ToS could be written to preclude the data brokers doing that, it doesn't affect the resulting model.

But the general shape of the argument brings us right back to unclean hands: we worked hard on this model and it's not fair for you to profit off our work in a way that doesn't have our permission.

Slashdot Top Deals

Programmers used to batch environments may find it hard to live without giant listings; we would find it hard to use them. -- D.M. Ritchie

Working...