Forgot your password?
typodupeerror

Comment Re:Sounds about right (Score 1) 127

In the case of Erdos 1196, https://www.erdosproblems.com/1196, the AI solved it by constructing a Markov chain on the integers weighted via the von Mangoldt function. Not an idea in the literature. If a human grad student had done that they would be immediately considered an up-and-coming bright star. For the Unit Distance Conjecture, https://arxiv.org/abs/2605.20695, the AI constructed a clever tower of number fields. Tim Gowers, who is a Fields Medalist and has worked on related problems, said this construction if done by a human, would have been a shoe-in for the Annals, which is widely considered the most prestigious journal in math. If you are claiming that all this sort of thing is is putting puzzle pieces together, then you are using that phrase to be so incredibly broad as to be absolutely meaningless.

Comment Re:Read Bubeck's response, not just Bumaster versi (Score 2) 127

If you have people working on a problem for decades, and there are rumors that they have made progress, what is the motivation to independently launch your millions of dollars of AI inference at that same problem?

In a normal academic context, none and moreover this sort of deliberate scooping would be considered incredibly bad behavior. When you have two giant corporations fighting to be able to have results they can showcase to the public and their investors? Then it makes a lot more sense, even as it is incredibly damaging to the academic math culture. This is really toxic behavior and deserves pushback. But it is a mistake to go from there to think that therefore any copying or plagiarism happened.

Comment Read Bubeck's response, not just Bumaster version (Score 5, Informative) 127

We've discussed this already on Slashdot here https://science.slashdot.org/story/26/09/08/2228220/openai-says-it-has-cracked-one-of-maths-millennium-problems. Bubeck, the research at OpenAI has a very different take on this entire situation here https://www.linkedin.com/feed/update/urn:li:activity:7503148080888893440/. I will quote his reponse:

I would like to clarify a few things:

1) The screenshot is my reaching out to Levent to coordinate our releases. I hope it’s clear from the message that we came in with the best possible intentions.

2) I never ever asked for Levent to be removed from authorship of his own work (as indicated by my text). I was surprised to learn during the call with Tristan that they had only solved Euler and not Navier-Stokes; after learning this we brainstormed possible paths forward. One option we discussed was that Tristan could be the lead author on a rewrite of OpenAI’s Navier-Stokes proof. It is in that context that I said “it would be simpler if Levent was not an Anthropic employee” because I felt it would be inappropriate for an Anthropic employee to author OpenAI’s work. Importantly it was admitted that internal Anthropic models had been used in their proof of Euler blowup; I therefore felt I could not consider Levent to be an independent academic. Another option I wanted to propose (but got cut short) is to offer access to our internal model so that they could try to finish their proof and bridge the gap between Euler and NS. Again I did not know how to navigate giving access to internal OpenAI IP to an Anthropic employee.

3) To reiterate it plainly: as my text clearly indicates, and as I said during our call, OpenAI's intention was to do everything possible to celebrate their mathematical achievements and the heroic efforts that they made on Euler. In the call I was immediately met with a litany of slander, including direct threats that if we were to announce Navier-Stokes he would immediately go to the press with a barrage of unfounded accusations. I refuted all these accusations but he replied “there is nothing you can do, I simply do not trust you”. I was confused why one would turn an incredible source for celebration (of their achievements!) into such bickering, which is when I said that I did not understand why one would risk their career [over unfounded accusations]. Genuinely, at that moment, I was trying to care for him and do a last ditch attempt to get a chance to give them all the credits that they deserve. I deeply apologize for this extremely poor choice of words, it is the opposite of what I was trying to convey. (I should say that I retracted them on the spot by the way.)

4) Overall, on a personal level, it was incredibly difficult to have these conversations. Levent refused to attend any of the meetings despite my repeated asking. As Sholto Douglas said, there will need to be coordination between Anthropic and OpenAI in the future; I felt I was doing a proxy negotiation with Anthropic while the Anthropic employee refused to directly participate.

As I said in that thread, I don't know whose version of events is accurate here, but some of the details due support Bubeck's version. I suspect that a breakdown of communication occurred where both then misinterpreted what the other saying in a more hostile way than it was intended until the conversation then became genuinely hostile. In terms of the math, although I'm a mathematician, differential equations is pretty far from my expertise, so I cannot deeply evaluate how close the methods were. However, it is the case that the both were building on the methods of Cordoba-Martinez-Zoroa. In this context, my being far from this sort of work is relevant, because this was a well known enough approach that even though I'm in a pretty different subfield, I had heard of CMZ's work as an approach and that this was considered a promising approach to Navier-Stokes. Given that, my inclination is that the case that any theft occurred here is very weak.

And since that thread, I've become more convinced, as experts reading both papers point to substantial differences in the details of their approaches. One of the ironies here is that a lot of the anti-OpenAI views are coming very loudly in part due to a general anti-AI attitude. But Bumaster and Alpoge, the people who are claiming to have had work stolen (well, primarily Bumaster, Alpoge is mostly staying out of the fray) were heavily using both Claude and Codex in their work.

Comment Re:Goal (Score 1) 33

Once again: training a model with the reward being "does it solve the task?" without looking at how it solves the task is very, very dangerous.

So, people are able to recognize this, and yet in the other threads here about Anthropic calling for a slow down of AI research, everyone seems convinced that this isn't about the risks really at all.

Comment Re:We are going so fast we need to slow down! (Score 0) 118

You can if you want, claim that OpenAI has had serious issues with honesty, especially with Sam Altman, their CEO. But Anthropic's record is a lot better. More to the point, the Theranos comparison doesn't work. The fundamental object Theranos was selling just didn't exist. AI systems do, and they are getting millions of people to use them. The situations are wildly different.

Comment Re:Or, counterpoint, it stole somebody else's work (Score 1) 97

You have never even met a practicing theoretical mathematician let alone participated in a conversation with one, have you? How would you know what constitutes a faux pas in their community, or how they collaborate?

Did you see in the comment where you are replying to where I said I'm a mathematician? My own primary areas of research are number theory and graph theory. It is pretty easy to find who I actually am and verify that yes, I am a mathematician and am in a pure field. So yes, maybe I do know something about the discipline.

Comment Re:AGI is a threat. LLMs aren't (Score 1) 124

This is pretty clearly not overhyping. First, it involves the fourth breach by an Anthropic AI. There's no hype at this point from identifying more breaches. Second, Anthropic, more than any company involved, doesn't want to be legally liable, since the current US federal government is looking for any excuse to go after them. Third, we've seen similar breaches by open weight models by third party security groups such as with Kimi 3. No one had any incentive there to claim it was hype https://www.wired.com/story/moonshot-kimi-k3-ai-model-escape-sandbox/ Fourth, in the case of OpenAI, we have a detailed third party investigation https://metr.org/blog/2026-08-26-openai-hugging-face-incident-investigation/ by METR which is a group which has a fair bit of reason to be skeptical or neutral to OpenAI. Overall, the pattern of these being a problem is real. One also doesn't need to think there's an existential risk here to see serious problems. If an LLM tries to crash an airplane or make a nuclear power plant go-offline, we have a problem even if that isn't an existential risk threat.

Comment Re:could have obtained internet access (Score 1) 124

The problem here is a bit more subtle. A lot of these tests and training are being done in data centers being used for other things also. So keeping the system physically airgapped is difficult. Complicating matters, they want people to be able to analyze and poke these systems at offices or homes in other locations. Now, if it were up to me, all of these breaches would be more than enough to insist on physical airgapping. But apparently the other issues and convenience are enough for them not to. It will be interesting to see if these issues are enough to change that. I suspect that as long as there are no consequences to Anthropic, or OpenAI or anyone else for these issues, the answer will be no.

Comment Re: This is just a distraction (Score 1) 167

Again, dealing with one risk doesn't mean concern about other risks goes away. And whether scientists are "tone deaf" shouldn't be remotely relevant. If something is a concern, it should be worked on whether or not the public is worried about it. In fact, part of how the public gets understanding is for researchers to go investigate something and then tell the public about it. As for "delusional," I would disagree and would these researchers even more so. But more to the point, the existential risk concerns are very explicitly not about "Terminator" with researchers concerned about this (like Yudkowsky) explicitly saying that they consider Terminator to be a bad analogy for what they are considering. So it suggests one should maybe grapple more with their concerns.

Slashdot Top Deals

"We don't care. We don't have to. We're the Phone Company."

Working...