Forgot your password?
typodupeerror

Comment Re:They are trying to sell the "cyberwar" sujet.. (Score 1) 140

And it's not exactly an ad for US models when HuggingFace had to rely on a Chinese model to analyze their logs because the US model they tried refused to answer.

It is if the customer is the US military.

* US model is capable enough to hack its way out. * China's model was able to detect it. * ... now we have an arms race.

Nope. You misunderstood. It's not that the US model wasn't able to detect it, it's that the US model's guardrails prevented it from explaining. This doesn't demonstrate a capability gap against Chinese models, it demonstrates a two-sided failure of the safety protections of the US models. On the one side, the safety guardrails on the OpenAI model failed to prevent the attack. On the other side the guardrails on whatever US model(s) they tried blocked the model(s) from explaining the attack (presumably to protect against the explained-to people from learning how to perform it).

Nope. I didn't misunderstand. You may have misread what I said, or attributed the previous comment to me.

You said that this creates an arms race, which implies that it's a question of capability.

Comment Re: AI (Score 1) 120

To each their own. I'm a python guy. I don't like typed variables or specifying anything I don't have to. Lately I don't work with code at all. I have AI do it in Python first then port to C++. It would be interesting to see if I could port to rust.

I strongly dislike dynamically-typed languages like Python. They're okay for toy programs but for anything of any size... you'd better have one hell of a good test suite because the tools give you absolutely no help. I want strong, static typing. I want a very picky compiler that won't accept anything that isn't exactly correct. C++ is good. Rust is better. I do like type inference so I don't have to manually specify types very often. C++ is reasonably good at this. Rust is better.

I also do a fair amount of work on very tiny systems, programming on the bare metal. Something like Python is a complete non-starter there, it just won't fit. C is the norm for those cases, but C is almost as loosey-goosey as Python, but without any run-time checks. C++ is much better than C. Rust is better yet.

The worst non-hardware bug I've ever dealt with in nearly 40 years of professional programming was in Python, using the Twisted framework. I traced it down to one place that took a pure abstract interface, instantiated it, then called methods on it!. Even logging the type of the instantiated object showed that it was the abstract type. Someone way too clever had built a dependency injection framework that was so magical that there was absolutely no way to figure out what dependency was being injected.

I also have AI write most of my code these days, but that actually increases my desire for a very picky compiler, because AI does a lot better with those guardrails in place.

Comment Re:OFFS...OpenAI LOVES government regulation (Score 1) 140

The point is...that's all it did. Carried out the instructions it was given. It didn't "act on its own".

That is completely and utterly irrelevant.

You need to read the story of the Paperclip Maximizer. It's an entertaining and humorous story, but the point is that instrumental goals are real and often diverge wildly from the intrinsic goals that drive them.

In this case, the LLM was directed to increase its CyberBench score, a sensible intrinsic goal for an AI being trained to be able to find vulns and write exploits. It chose to do that by breaking out of its container, hacking the company that creates the scores, stealing employee credentials, and breaking into the score database. Those intermediate steps were instrumental goals, and they obviously diverged wildly from the intended intrinsic goal.

Would you also have said "all it did was carry out its instructions" if it had found that in order to break in it needed to hire private soldiers to shoot their way in? Or take control of the US nuclear arsenal and issue a threat to HuggingFace that they increase its score or it would nuke them? Or...

Yes, these are fanciful scenarios. What they are not, however, is impossible scenarios. And as AI gets more and more capable, you really want them to be impossible. Alignment matters. If fast takeoff happens (that's a theorized situation where AI becomes capable of self-improvement and rapidly makes itself vastly smarter than any human), alignment may well be essential for human survival.

Comment Re:So much drama with Open AI and Anhropic models (Score 1) 140

Literally the opposite. Both have about 40% margins, and that includes free users.

Losing money, not projecting profitability until 2030.

Both of those things are true, I'd guess.

I am not an accountant, but as I understand it, "Profit margin" is generally "operating profit margin", meaning revenues less operating expenses, all divided by revenues, and their operating expenses aren't that large, so they can absolutely have very strong operating profits. But they also have mind-bogglingly huge capital expenditures going on, so when you account for those they're losing a lot of money. I'm surprised they're projecting profitability as soon as 2030, actually. That's only 3.5 years away.

Comment Re:toast (Score 1) 140

Slashdot's summary left out the best part. Yes, GPT 5.6 Sol was indeed trying to cheat on an evaluation, but what specific evaluation? CyberBench. A benchmark testing how good AI models are at hacking.

We told it to do better at that benchmark. How does the training data imply that should be done? How many companies have been caught gaming the benchmarks?

Well... an AI model that is capable of hacking into the benchmark database and altering its score is clearly better than one that is not.

Comment Re:They are trying to sell the "cyberwar" sujet.. (Score 1) 140

And it's not exactly an ad for US models when HuggingFace had to rely on a Chinese model to analyze their logs because the US model they tried refused to answer.

It is if the customer is the US military.

* US model is capable enough to hack its way out. * China's model was able to detect it. * ... now we have an arms race.

Nope. You misunderstood. It's not that the US model wasn't able to detect it, it's that the US model's guardrails prevented it from explaining. This doesn't demonstrate a capability gap against Chinese models, it demonstrates a two-sided failure of the safety protections of the US models. On the one side, the safety guardrails on the OpenAI model failed to prevent the attack. On the other side the guardrails on whatever US model(s) they tried blocked the model(s) from explaining the attack (presumably to protect against the explained-to people from learning how to perform it).

Comment Re: AI (Score 1) 120

It can't be nicer. It nags you into not doing certain things.

It's absolutely nicer. I've been writing C++ for 36 years and really enjoy it, but I find Rust to be much more pleasant to work in.

I don't know about Qt from Rust (I avoid UI code). It's been 15 years since I wrote any Qt code so I don't really remember it all that well even from C++, and it's probably evolved, so I can't even guess about what it might be like and whether the impedance mismatches with Rust are significant -- The C++ and Rust object models are somewhat different, so there definitely can be mismatches.

But in general, I find Rust is just nicer to work in. I'm more productive and the code feels cleaner and tighter.

As for the "nagging"... that's really only a learning curve issue. It's a non-trivial learning curve, true, but once you get past it the borrow checker doesn't really get in your way. And what you get in trade for that learning curve is fantastic. You just don't have to think about a whole raft of memory-safety and concurrency problems, because the language guarantees they can't occur. You can get the memory safety guarantees in Python and similar language, but they come at a heavy runtime cost.

And, yes, I know all about "Modern C++" and how avoiding raw pointers, etc., is supposed to give you memory safety. And it does help, a lot... but you'd still better run valgrind on your programs to check for subtle mistakes, and even that might not catch them all.

Overall, Rust is just nicer.

Comment Re:Madness (Score 1) 215

Agreed, though the trend towards expansion of presidential power through executive orders goes back long before Biden. Back to Reagan, at least, and much of the root of it was in legislation signed by Carter (IEEPA). Every successive administration -- both parties -- has pushed the boundaries. Dubya was probably the worst, mostly because he had the biggest hammer. But the pattern is that each administration takes a bit more and when the other party gets elected, they don't give any of it back.

Very much a "both sides" problem.

Comment Re:Figures (Score 1) 93

I'd say Slashdot is in a much more degraded state than Fark.

Fark at least lets you create an account. Slashdot requires you to email them and explain why you should get an account if you want to join.

https://slashdot.org/my/newuse...

New user registration is now approved by Slashdot administrators. Please contact feedback@slashdot.org and let us know why you are interested in registering, and what you can add to the discussion.

Comment Re:imagine (Score 1) 169

having the world's worst criminal state as your greatest ally

I take your point, but I have to ask "Which state is which"? If forced to choose between Israel and America for the title of worst criminal state, I'd have to flip a coin.

There are a lot of problems with both, but neither deserves the title of worst criminal state. Most hypocritical state, maybe.

Comment Re:android 16 flags them (Score 3, Interesting) 169

android 16 is expected to start flagging rogue cell phone towers.

https://www.techradar.com/phon...

It does. I got a notification in early 2025 (running a pre-release Android 16 build) while traveling in NY. I got someone who was working on the feature to pull the detailed logs from my phone and they were able to get quite a bit of detail about what kind of fake tower it was. Pretty cool.

Comment Re:Question (Score 2) 169

Is it possible to put software on a phone that prevents it from being tracked this way? Sort of like randomized MAC addresses with WiFi.

Not if you want to have mobile phone/data service.

If your mobile provider supports VOIP and you have Wifi available (e.g. at home) you can just disable cellular service. Then everything will go through your Internet connection -- and everything is encrypted these days. But that only works when you have Wifi.

Slashdot Top Deals

"Consequences, Schmonsequences, as long as I'm rich." -- "Ali Baba Bunny" [1957, Chuck Jones]

Working...