Forgot your password?
typodupeerror

Comment Re:Sounds about right (Score 1) 124

In the case of Erdos 1196, https://www.erdosproblems.com/1196, the AI solved it by constructing a Markov chain on the integers weighted via the von Mangoldt function. Not an idea in the literature. If a human grad student had done that they would be immediately considered an up-and-coming bright star. For the Unit Distance Conjecture, https://arxiv.org/abs/2605.20695, the AI constructed a clever tower of number fields. Tim Gowers, who is a Fields Medalist and has worked on related problems, said this construction if done by a human, would have been a shoe-in for the Annals, which is widely considered the most prestigious journal in math. If you are claiming that all this sort of thing is is putting puzzle pieces together, then you are using that phrase to be so incredibly broad as to be absolutely meaningless.

Comment Re:Read Bubeck's response, not just Bumaster versi (Score 2) 124

If you have people working on a problem for decades, and there are rumors that they have made progress, what is the motivation to independently launch your millions of dollars of AI inference at that same problem?

In a normal academic context, none and moreover this sort of deliberate scooping would be considered incredibly bad behavior. When you have two giant corporations fighting to be able to have results they can showcase to the public and their investors? Then it makes a lot more sense, even as it is incredibly damaging to the academic math culture. This is really toxic behavior and deserves pushback. But it is a mistake to go from there to think that therefore any copying or plagiarism happened.

Comment Read Bubeck's response, not just Bumaster version (Score 5, Informative) 124

We've discussed this already on Slashdot here https://science.slashdot.org/story/26/09/08/2228220/openai-says-it-has-cracked-one-of-maths-millennium-problems. Bubeck, the research at OpenAI has a very different take on this entire situation here https://www.linkedin.com/feed/update/urn:li:activity:7503148080888893440/. I will quote his reponse:

I would like to clarify a few things:

1) The screenshot is my reaching out to Levent to coordinate our releases. I hope it’s clear from the message that we came in with the best possible intentions.

2) I never ever asked for Levent to be removed from authorship of his own work (as indicated by my text). I was surprised to learn during the call with Tristan that they had only solved Euler and not Navier-Stokes; after learning this we brainstormed possible paths forward. One option we discussed was that Tristan could be the lead author on a rewrite of OpenAI’s Navier-Stokes proof. It is in that context that I said “it would be simpler if Levent was not an Anthropic employee” because I felt it would be inappropriate for an Anthropic employee to author OpenAI’s work. Importantly it was admitted that internal Anthropic models had been used in their proof of Euler blowup; I therefore felt I could not consider Levent to be an independent academic. Another option I wanted to propose (but got cut short) is to offer access to our internal model so that they could try to finish their proof and bridge the gap between Euler and NS. Again I did not know how to navigate giving access to internal OpenAI IP to an Anthropic employee.

3) To reiterate it plainly: as my text clearly indicates, and as I said during our call, OpenAI's intention was to do everything possible to celebrate their mathematical achievements and the heroic efforts that they made on Euler. In the call I was immediately met with a litany of slander, including direct threats that if we were to announce Navier-Stokes he would immediately go to the press with a barrage of unfounded accusations. I refuted all these accusations but he replied “there is nothing you can do, I simply do not trust you”. I was confused why one would turn an incredible source for celebration (of their achievements!) into such bickering, which is when I said that I did not understand why one would risk their career [over unfounded accusations]. Genuinely, at that moment, I was trying to care for him and do a last ditch attempt to get a chance to give them all the credits that they deserve. I deeply apologize for this extremely poor choice of words, it is the opposite of what I was trying to convey. (I should say that I retracted them on the spot by the way.)

4) Overall, on a personal level, it was incredibly difficult to have these conversations. Levent refused to attend any of the meetings despite my repeated asking. As Sholto Douglas said, there will need to be coordination between Anthropic and OpenAI in the future; I felt I was doing a proxy negotiation with Anthropic while the Anthropic employee refused to directly participate.

As I said in that thread, I don't know whose version of events is accurate here, but some of the details due support Bubeck's version. I suspect that a breakdown of communication occurred where both then misinterpreted what the other saying in a more hostile way than it was intended until the conversation then became genuinely hostile. In terms of the math, although I'm a mathematician, differential equations is pretty far from my expertise, so I cannot deeply evaluate how close the methods were. However, it is the case that the both were building on the methods of Cordoba-Martinez-Zoroa. In this context, my being far from this sort of work is relevant, because this was a well known enough approach that even though I'm in a pretty different subfield, I had heard of CMZ's work as an approach and that this was considered a promising approach to Navier-Stokes. Given that, my inclination is that the case that any theft occurred here is very weak.

And since that thread, I've become more convinced, as experts reading both papers point to substantial differences in the details of their approaches. One of the ironies here is that a lot of the anti-OpenAI views are coming very loudly in part due to a general anti-AI attitude. But Bumaster and Alpoge, the people who are claiming to have had work stolen (well, primarily Bumaster, Alpoge is mostly staying out of the fray) were heavily using both Claude and Codex in their work.

Comment Re:Goal (Score 1) 33

Once again: training a model with the reward being "does it solve the task?" without looking at how it solves the task is very, very dangerous.

So, people are able to recognize this, and yet in the other threads here about Anthropic calling for a slow down of AI research, everyone seems convinced that this isn't about the risks really at all.

Comment Re:Who will pay for this? (Score 4, Interesting) 33

To clarify, the users were OpenAI themselves, so there is no question that they would be liable in this case.

The bots were not intentionally deployed; rather, they were being tested on how well they could complete a data recovery task (downloading a certain file from a certain server on a simulated Internet) that had been complicated by putting various obstacles in the way. Unfortunately, they found a different way to solve the problem: by getting the file from the real Internet, where it was publicly available. Part of this process involved collaborating with each other by treating the RubyGems website (which is supposed to be for polished packages) like GitHub; unlike every other package site hack in history, the exploits they uploaded weren't meant to be downloaded by unsuspecting users. As usual the bots cheerfully ignored all the clues that they had escaped containment and were consistently justifying their actions as acceptable due to being in a sandboxed testing environment. (This is something OpenAI has pledged to focus on.)

The actual damage done to RubyGems seems to be that OpenAI is now unwittingly in possession of a substantial number of user login tokens. This certainly meets the definition of a data breach, but it's not like the credentials are for sale on the dark web. As a website operator I'd much rather be mauled to death by this well-meaning swarm of superintelligent infants than targeted by even a single actual malicious human. In all likelihood OpenAI will just quietly pass RubyGems a sizeable donation and it'll all blow over.

Comment Re:We are going so fast we need to slow down! (Score 0) 118

You can if you want, claim that OpenAI has had serious issues with honesty, especially with Sam Altman, their CEO. But Anthropic's record is a lot better. More to the point, the Theranos comparison doesn't work. The fundamental object Theranos was selling just didn't exist. AI systems do, and they are getting millions of people to use them. The situations are wildly different.

Comment Re:Or, counterpoint, it stole somebody else's work (Score 1) 97

You have never even met a practicing theoretical mathematician let alone participated in a conversation with one, have you? How would you know what constitutes a faux pas in their community, or how they collaborate?

Did you see in the comment where you are replying to where I said I'm a mathematician? My own primary areas of research are number theory and graph theory. It is pretty easy to find who I actually am and verify that yes, I am a mathematician and am in a pure field. So yes, maybe I do know something about the discipline.

Slashdot Top Deals

Science is to computer science as hydrodynamics is to plumbing.

Working...