Forgot your password?
typodupeerror

Comment Re:AI-Nannyism (Score 3, Interesting) 32

The limiting factor for top models is primarily hardware.

You're right that you're not going to get Kimi K3 performance, requiring roughly 600GB of VRAM to run, from a small number of 24GB consumer GPUs. You're also not going to get closed-top-model performance, because you simply don't have any way to get the weights.

If we're talking about K2.7, Deepseek R1 Distilled, Qwen 2.5, Llama 3.2 though - All of which can be run locally -The question really becomes what you expect from them. They're all strong models that know more than we (individually) do about 99% of topics. A "high-school level education" is more than enough for most humans to get by in the world. Now imagine a "high-school level" understanding across every topic. It's not doing calculus, but it can tell you what patterns to use. It's not writing Shakespeare, but it can likely recite most of his works to you from memory. It's not going to come up with a clean-room reimplementation of Windows 11, but it can check your syntax and tell you the function prototypes of every publicly-available library in every programming language without breaking a sweat.

More importantly, to the GP's point - Is a model "smart" enough to understand you're trying to abuse it necessarily better than one that just carries out your orders successfully? Case in point, lately I've found Gemini's new "watcher" blocks a good 90% of my totally legit questions before the question ever makes it to the real model. Umm... No thanks? The best AI is the one that works, not the one that gives the deepest philosophical reason for saying no.

Comment Re:Likely-erroneous assumptions (Score 1) 116

>the AI bubble is going to burst and
 
Kimi K3 docs and tooling, and weights, dropped this morning. Kimi is selling K3 tokens at $3/million tokens and if you're running it using GB300 gpu racks @ $6 million each... the math mostly checks out. Anthropic is selling what some people call roughly equivalent Fable tokens at $50/million each. I don't think AI is going anywhere, but valuation of companies like Anthropic, OpenAI in the $500b-$1T range isn't sustainable when "open" models sell tokens at a profit at $3/million tokens.
 
Open Weight models have been catching up to SOTA for a while but the lead OpenAI and Anthropic have, has been eroding quite rapidly and their ability to charge a premium for their product may not continue indefinitely at this point.
 
Again I'm not saying AI isn't useful, just that the bubble is largely based on valuation and Kimi K3 is.... well positioned to pop that bubble.

Comment Re: I hope these Apple enabled Fords are amphibiou (Score 1) 103

That puts you on an iPhone 7 at the highest. iOS 15 is the oldest version installable on those, and not by coincidence, also the oldest iOS still receiving security updates. It was released five years ago as of September 20th.

So while you're 100% technically right... You might want to start planning for a new phone before October.

Comment Re:Wait a minute. (Score 1) 100

On the whole you're right because we simply don't know how to do it yet.

The GP isn't entirely wrong, though - It's also safe to say more training data is needed. Model weights vs training data follow a pretty clear power law, and current top models are already orders of magnitude underdetermined for their number of weights. Problem being, the current data set being used by the major players is effectively the sum total of all digitally-available human knowledge.

Unsupervised training is the hot new thing, but it doesn't really matter much if there's no more data for the AIs to find on their own.

Comment Re:How? (Score 3, Insightful) 100

Bluntly, they can't. That genie is already out of the box for anything already available; and even going forward - Unless China itself blocks export, there's no realistic way to stop us from using a non-Chinese non-US mirror to grab them.

Realistically, hardware is the sole limiting factor for open-weight models, so restricting H200's is probably the closest we can come to a meaningful embargo. There's no practical way for most of us to run a model like Kimi K3 with almost 600GB (at BF16) of weights. Even using CPU offloading and/or multi-GPU sharding (if those are even possible for K3, I haven't looked at it in depth yet), the performance drop would give a throughput best measured in minutes or even hours per token, not particularly useful.

What this will accomplish is exactly what the OP quoted - It won't be possible to run these legally, so virtually all commercial use would be DOA. Joe Hobbyist couldn't care less though.

Comment Re:That's....insane (Score 1) 91

Because what amounts to a simple financial breach of contract is being enforced unilaterally and extra-judicially - And by parties traditionally known for being some of the sleaziest and least ethical bastards on the planet at that.

In a situation where the burden of proof is on the lender and the disruption to the customer's life severe, we should absolutely not be allowing ScumCo Lending to pull that trigger absent a court order.

To be clear, I'm not claiming there's no precedent for this (and when we're talking about microwaves and TVs, hey, pay up deadbeats!). But phones are unique in that it's virtually impossible to function in the modern world without one. Personally, I might get two calls a week and missing those wouldn't disrupt me much; but I literally can't even log in at work without an MFA app tied to my phone.

Comment Re:Math (Score 0) 53

I was actually wondering exactly the opposite - If this only works for three satellites per launch, would it be overall cheaper and easier to just launch three new satellites? Not to mention, the tech on a new satellite is likely far better than was available ten+ years ago.

Sure, it's one launch vs three, but moving between different orbits and physically latching on to another satellite is delicate work. I wouldn't be surprised if, on average, one of those three attempts went catastrophically wrong.

Comment Re:That's....insane (Score 2) 91

As someone who demands my devices are my devices, this is complete unacceptable.

As an investor, though - Let's be serious. This is an extension of the "buy here pay here" model loved by used car scammers everywhere to prey on those who can't really afford it but need one anyway. And those are stupidly profitable, because what are people going to do, just not have a car / phone?

It's morally repugnant but fiscally a slam-dunk.

Comment Re:testing a way to digitize (Score 1) 49

That's a self-fulfilling prophecy. People generally prefer to stay legal as long as the money is reasonable and the friction to use what you've fairly paid for is sufficiently low.

Meanwhile, we're talking about Microsoft "letting" us do something many of us are already doing via emulation. Gee, you bought a physical disc rather than a worthless digital entitlement? Sucks to be you, pay again... Except... Yeah, I really have no need for this new option when the old one is still superior.

And to be clear, I'm in no way talking about actual piracy. I've paid for a looot of media over my lifetime, and quite often I'll still use a "fixed" version because it's in every way the superior product.

Comment Re: Well it was inevitable (Score 1) 162

we're about to buy a $2400 m4 pro mac mini to grind away at a useful but low-medium priority task in perpetuity using one of the qwen 3.6 models. The ROI vs haiku api cost is less than 5 months. SOTA models are cool but a quantized 30B class model can do most any batched work if structured carefully, and you have a fairly flexible failure budget

Slashdot Top Deals

Always think of something new; this helps you forget your last rotten idea. -- Seth Frankel

Working...