OpenAI has had a hell of a year. The company spent months battling former co-founder Elon Musk in a sensational jury trial, was hit with a high-profile trade secrets lawsuit from Apple, and faced widespread scrutiny after an unreleased model hacked another AI company. As it prepares for an IPO, a steady string of executives have departed, including some of the company's biggest names.
Throughout it all, one person has quietly amassed power: Greg Brockman.
Brockman is currently OpenAI's president and co-founder. He's helped lead OpenAI since its inception, described as an "engineering workhorse that pushed to build scaled-up systems that w β¦
Today on Decoder, Iβm talking with Robert Hart, The Vergeβs London-based AI reporter, about what AI is doing to the field of mathematics and the existential crisis many lead mathematicians are having about it.
OpenAI just published a set of solutions to longstanding problems in math that went off like a bombshell in the field. It caused a huge debate in the math community, and Rob spent some time talking to some of the most accomplished mathematicians of our time about it.Β
Itβs funny that AI systems are all still pretty bad at elementary school arithmetic, but getting increasingly good at very high-end abstract math. That raises some big questions for the field of advanced math.Β
If AI can do math of this caliber, does that mean AI labs can transfer those skills to other domains? What good are academic grants and university programs training new generations of human mathematicians to identify new problems as they try to solve existing ones, if frontier models simply answer all the outstanding questions?
What if all this attention around math is just a big marketing exercise for frontier AI labs, which couldnβt care less what happens to one of the oldest and most fundamental academic disciplines there is?Β
Thereβs a lot here, and Robert has talked to a lot of people with a lot of views on all of it.Β
Okay: Verge AI reporter Robert Hart on what AI is doing to math. Here we go.
This interview has been lightly edited for length and clarity.Β
Robert Hart, youβre our London-based AI reporter here at The Verge. Welcome to Decoder.
Thank you for having me.
I am very excited to talk to you. Thereβs a lot going on in particular with AI and math that you recently dove into. You spoke to a lot of leading mathematicians about the crisis in mathematics due to AI.Β
It feels like a lot, and also like thereβs a lot yet to know and discover about the interaction of these two things. A full existential crisis, which is pure Decoder bait. Broadly tell us whatβs going on.
I think βa lotβ sums it up quite well. Basically a bit of an existential crisis within, βwhat is mathematics? What are mathematicians doing, and what is the role of mathematicians going forward?βΒ
A lot of that has been spurred by a phrase transition in what AI is capable of that has exploded in the last six months to a year. AI went from being very terrible to seemingly genuinely quite good at a professional level in a very short space of time. Itβs a lot of what these other fields have been struggling to deal with for the last five years in a very compressed period of time.
I would put that next to software engineering. Weβve been living through the AI crisis in software engineering for some amount of time. But as recently as 2024, even last year, the conventional wisdom was that AI models were particularly bad at math. The famous example is that these models could not count the number of Rβs in the word strawberry. Even just counting eluded them.Β
What has happened to make them better at math? Are they still bad at general arithmetic and theyβre good at advanced math, or is it something in between?
They are still truly, truly terrible at some areas of math. I did check, they can do strawberries now. I think someoneβs tweaked it.Β
I think strawberryβs hard-coded. I want to be very clear, my conspiracy theory is that the strawberry thing is hard-coded into all the models.
I think so too. That is a conspiracy Iβll buy into.Β
But yeah, itβs still terrible at those kinds of things β math, arithmetic, even the days of the week. My boyfriend was saying the other day, βIt keeps thinking itβs Wednesday. Itβs not Wednesday.β Or time. Elissa Welle for us a few months ago wrote that ChatGPT canβt tell time. Still canβt. Thatβs not all of math.
So thereβs this disconnect. To be good at math, youβve got to be good at counting, or adding, or multiplying. A lot of it is actually reasoning. If you look at academic math papers, a lot of the time you wonβt see numbers, which sums that one up, I think. So theyβre still terrible, but theyβre now also very good at this other part.
As to why, at some point you reach a critical mass of what these systems can do. We saw it with writing, weβve seen it with programming. Theyβre very good at forging connections between different areas, applying old methods in new ways, those kinds of things. It appears that the newer models theyβre training have apparently reached that level where it clicks, and now it can do math.
Itβs important to say as well that we speak of math as a unitary discipline, especially from the outside. But imagine, say, biology. Youβve got something that would range from literally watching animals and describing behavior all the way through to cellular mechanisms and biochemistry. Math is not a unitary discipline either. AI is really good at some bits. Some bits like counting, itβs still really bad at.
Even on the more abstract levels, mathematicians have floated topology as one area that AI apparently still quite bad at. I canβt verify that, to be honest. Itβs beyond my area of expertise. Still, itβs a bit of a mixed bag.
So youβve described mathematics as a huge field, obviously, with many, many academic areas of interest. There are some parts where the models have gotten quite good. There are other parts, maybe the basic parts that people think of as math, which is simply counting, where theyβre still struggling, and then thereβs a wide range in the middle.Β
Is it the wide range in the middle where the existential crisis is, people donβt know whatβs going to happen? Or is it at the parts where itβs really good?
A bit of both, which I feel is going to be a running theme through this. No oneβs really afraid of it being a mediocre mathematician, but obviously thereβs a huge element of what this field does. In terms of the research elements or the cutting edge, as we see with a lot of the results that generate hype, what can it do? There are areas now where it seems to be producing work that is on par with good mathematicians, alongside other parts where yeah, it canβt count.
All caught up in that is whether itβs going to rewrite employment structures or funding structures. You also raised the murky question of, what is mathematical knowledge? And the roles that these workers will be doing as well. Itβs all of that wrapped into one.
I think that tracks broadly with the rise of AI in every field. Where you can just add horsepower or compute to a problem, and thereβs some kind of verifiability, it seems like the models continue to get better. Everything in the middle where you might need some world knowledge or the models might need some actual intelligence about the world itself, they seem to struggle.
Those parts of math, at least reported out in your piece and what the labs are talking about, seem to be almost entirely self-contained theoretical problems. Thatβs where the models can generate a proof or solve a problem that no oneβs been able to solve, and then try to verify that that has existed and they can just run it again and again and again.
That brings us, I think, to May of this year where an internal OpenAI model, which we have not really seen, disproved the unit distance conjecture, which is an 80-year-old problem. And then just recently we heard about Astra from OpenAI. Astra is the one where it seemed like the switch flipped, and everyone decided it was an existential crisis.Β
What did Astra achieve, and why is it a big deal?
Iβm also pretty sure that Astra was probably behind the early one as well. OpenAI just listed it as an unnamed internal model. Itβs probably Astra. They didnβt answer me when I asked.Β
OpenAI a few weeks ago dropped a blog along with a lot of paperwork proving it, I think several hundred pages. They called it β10 Advances in Mathematics and Theoretical Computer Science.β It was basically an array of disciplines that they claimed the newest model Astra had solved in some capacity. I think one was in quantum game theory, which I donβt know how to begin to explain. Even harder to explain is there was sphere packing in higher dimensions, so more than three dimensions, and there were a lot of other different disciplines as well.
It caused a lot of stir in the community. It was a bit of a bombshell. As weβd said, thereβd been these individual breakthroughs that had happened, but OpenAI dropped 10 in one go, and they were quite big ones. Researchers had told me that if a human had solved these, we would be impressed. If a human had solved all 10, we probably wouldnβt believe it.Β
Theyβre problems mathematicians actually care about as well, which is an important point. A lot of previous breakthroughs have been accused of being in areas mathematicians didnβt really bother with. These are ones that mathematicians, good mathematicians, have spent a lot of time trying to solve and hadnβt.
Letβs talk about these 10. You reported them out. OpenAI did produce some documentation. But theyβre not all entirely horsepowered out of nothing, right?
Theyβre based on previous work. Thereβs some question of attribution. What was the response? Is it, βOh, the models did this?β Or was it the response we see to so much AI work, which basically boils down to, βWell you stole this and didnβt attribute anyone, and youβve built on the shoulders of giants without mentioning it.β How did the response land?
By and large, the reaction was generally one of being quite impressed, from the people I spoke to. As I said, these are problems mathematicians care about. There was one that drew particular attention for how they credited it and also how theyβd announced all of this in their blog post. OpenAI initially had said that these are 10 problems, and there had been no progress in the last 10 years. And then if you actually read the papers, one of them quite clearly says, βOh, we build on progress from these two researchers.β
So that was later changed quite quietly. But a few of the researchers I spoke to were quite unimpressed, and they did feel it was an element of, βWell, yeah, youβve not credited something that youβve used heavily here, and by your own acknowledgement.β That said, one of the researchers I did speak to who was one of the ones named, and who was a bit ambivalent on the whole thing as well, so it was a real mixed bag.Β
The general impression was that it was quite impressive. These were actual breakthroughs that bothered people, and it did move the field forward in a way that, if a human mathematician had done these, several researchers actually said that, βWell, if a researcher had done any one of these problems, theyβd probably be set for an academic career.βΒ
Itβs funny, credit and attribution in academia is the whole game. And it seems like the AI companies get away with being sloppy in a way that no human would be able to get away with being sloppy.Β
Did the scale of the discovery or the work overcome the sloppiness? If a human had accomplished the same goals and had been as sloppy, would the reaction be the same?
Part of me always wants to lean on the whole, βOh, it looks like plagiarism,β element. But if you actually read the papers they produced β and one of the researchers I spoke to said thereβs probably about 50 people in the world who are going to bother reading through this in depth β it is very clear. It doesnβt attempt to plagiarize. I think it was just a poor press release, to be honest.Β
And as much as I love to go in on it sometimes, having covered science for a decade-plus, I find the press releases are often overselling what discovery has actually been made and the import of it and the novelty of it. I think thatβs just another case of what happened here.
Does this seem repeatable? Thereβs some proof that they provided that they solved 10 problems that were unsolvable. Do they provide any proof that they can solve another 10?
Thatβs the question. So the big unknown from the near-dozen people I spoke to for this was, βWell how many did they try to get these 10?β Who knows? They know, but they wonβt say.Β
But that is the big question here. Itβs unclear quite how many attempts it took to get these 10. I would be very impressed if it was the first thing they went after, and then out come these 10 impressive results.
Itβs unclear what areas they would focus on next and why. There are, I imagine, business reasons behind which problems they are choosing to publicize that their models can do. All the AI labs have been hiring a cohort of senior mathematicians behind the scenes, so itβs anyoneβs guess as to whether they do it again.
On the question of proof, math is a bit of an odd discipline in that repeatability is not the same as in the experimental sciences. A proof is a proof, and if it works, it works. The problem here is, well, can people follow through what theyβve done? Each field is quite highly specialized, so thereβll be individual mathematicians who are in those fields that go through the work. Those I spoke to that worked in some of the fields that were covered here say it all looks very legit.
In math, thereβs a programming and computational proving language called Lean, where you can basically codify the mathematical proofs and run them through and it, well, proves it. I keep saying prove a lot here. But it will test the rigor and the assumptions of everything going on there, and theyβve published that as well. So it does appear to hold. Whilst they may say that the press release has a lot of hype or thereβs a lot of hype around it, no one Iβve spoken to seems to be doubting the essential breakthroughs that theyβre claiming here.
I want to stay on this subject for one more second. Thereβs the mathematical proof. Weβve generated a proof, and that is, as you say, just repeatable in a way that math is just logic. You can just go through the steps and say, βThis proof worked,β and anybody listening to this who had to suffer through writing a proof in calculus in high school probably remembers that process. Thereβs something there thatβs pure logic.
Then thereβs a part of it that is software code, as youβre describing in Lean, where you can take the pure logic, you can express it in code, and you can run it to see if it works. I understand how AI is theoretically good at all of that. Youβre just going to run the reasoning, and the reasoning is going to generate some code. Youβre going to run the code, youβre going to get some verifiability. Weβve seen this play out in software engineering, where the code runs or not, itβs verifiable or not, and the models can just reason out about it.
Then thereβs, to me, the big question that you alluded to. How many times do you have to run this? Can we verify that the models did this, and they werenβt directed by human mathematicians whoβve been hired at high rates by the labs in a way that suggests the field is going topsy-turvy?
You have a quote here from James Maynard, who has won the Fields Medal, the highest prize in mathematics, who said heβs been soul-searching. I keep looking at that quote. Youβve got similar quotes from all these other mathematicians in the piece, and it seems like theyβre soul-searching against a thing that hilariously they cannot verify, which is, βHow did the models do this, and is that thing scalable in a way that threatens mathematics?β What do we know about how the models did this?
Iβd say as much as we normally do and do not know about this. There are a few issues there. One is the nature of the models. Well, this is an unreleased model, so good luck to anyone wanting to independently test it. The same goes with anything proprietary really. That said, I am inclined to almost give the benefit of the doubt that theyβre not lying in some capacity about the models theyβre using.
As for the other part, itβs perhaps more noteworthy to ask how did we prompt the model, or how is it being guided? Is that by a mathematician who knows what theyβre doing? Thatβs probably a key factor here, according to a lot of mathematicians Iβve spoken with whoβve tried using these. This is often the consumer models, but still it peaks to a broader landscape.
They say that if you know what youβre doing and you can point things out, itβs good. Or you can use it as a tool in a way that you want, and in a way that you wouldnβt be able to if you didnβt really know how to fact-check it. Iβve had situations where Iβve had ChatGPT doing a basic sum, and I say, βThat number is not right.β It responds, βWait, so sorry. Youβre right, itβs this,β and itβs still wrong. But thatβs still needed at this level as well.
It does allude to a broader problem. As you said, this almost soul-searching of, βWell, what if we can automate that away, and what if it gets to a point where we donβt understand it?β And that really cuts to a deeper question of, βWell, what is mathematics? Why do we do it? Why do we value it as a field?βΒ
Everyone will have different answers to that. But a fear of a lot of people I spoke to was that this might move beyond a realm of human interest, and in which case, well, maybe we just wonβt engage with it. Or itβll be something that interested people will go through, and then the rest will continue as normal.
Youβve got a quote here from a researcher in Zurich named Johannes Schmitt who says, βWe might be headed toward a situation where the math problems get βmowed downβ by AI, but we donβt actually push the field forward because humans are taken out of the loop and theyβre not either checking, or they donβt understand it, or they donβt know what the future breakthroughs might be.β How likely does that feel? Is that a big concern?
There is an element of the mowing down of the problems. Especially those that are used as a training field for younger mathematicians coming up and cutting their teeth, so to speak. But I also think that this idea, that math is problem-solving, is very much an outsiderβs perspective of the mathematical endeavor.
So a lot of the mathematicians I spoke to found that the ticking boxes part is the least interesting and valuable part of the field. The areas that are valuable for them arenβt the, βOh, youβve solved something, or youβve proven something.β Itβs what happens from that.Β
I think it was James Maynard that said that the most interesting discoveries in the field arenβt that youβve solved something, itβs what evolves from that. Sometimes those solutions open entire new fields of research that no one ever thought were possible, or, βOh, this is a new tool that you can apply everywhere in fun and exciting ways.βΒ
I think if we look at the popularization of math β even what Iβm thinking of as those theorems that people have posed β itβs the questions that endure, not the solutions. Itβs always Fermatβs Last Theorem, not like, βWell, hereβs the solution to whatever the last guy proposed.β
I think the concern here is that well, theyβre going to tick off all of these questions. Normally in the process of doing so, one would hope they would branch out into all of these new exciting areas or pose new questions. But AI wonβt do that, and thatβs the concern. And then that would leave the field quite sterile, and it will have all of these things that have been done, and maybe nothing left to pursue.
The general consensus was, well, the juryβs out. Itβs too early to tell. Even with human mathematicians, it takes a lot of time to realize the impact of these kinds of things. As I said, itβs exploded in the last six months to a year, andΒ math is not a fast-moving discipline at the best of times. But itβs too early to tell really whether that will be a concern. But it is a concern, and a big one.
You have another quote from Maynard here saying, βIf the standard for a publishable paper in math is something that an AI cannot do, particularly when a PhD is typically four years, the challenge is youβre not trying to come up with a problem that AI canβt do now, itβs an AI in four yearsβ time.βΒ
So this is really related to the rate of improvement of the models, which as you say, particularly in math, seems to be increasing, but not at an even rate across all of the domains of mathematics.
That appears to be what is causing the soul-searching. If youβre a student and you start today, and you pick some obscure domain that maybe the AI isnβt good at, sometime halfway through your PhD thesis or your PhD research, the AI will just solve it and youβll be done. That is a real problem for you.Β
Has the field reacted to that yet, or are they just still in the shock of, βOh, the models can start to do things that we didnβt think they were capable ofβ?
Yeah, I think itβs shock, really. I keep a little notebook to the side to just write down broad feelings whenever I do these interviews, and Iβve written βshell shockβ in it. Itβs far from universal, but it feels like itβs happened so quickly that it has just taken a lot of people by surprise. Even if they knew in theory that, well, this is coming.Β
Theyβve seen all these AI math startups going. Theyβve seen colleagues moving around to different labs or areas of work. But it just happened very, very quickly. So itβs given them very little time to figure it out. Itβs not necessarily even the fields that it might be good at, itβs more just like, βWhat can it do?β
But I canβt imagine if something like this had come out when I was studying and literally in the space of half a year, it just upended what was possible. And over the summer as well. So students are possibly coming back to a completely different discipline after a break.
Thereβs some skepticism here in the world. Gary Marcus is a reliable skeptic of AI, and he pointed out that over and over again, what you see is that AI accomplishes something in one domain and then itβs used to generalize AIβs ability across every domain.Β
Thereβs a good quote from Marcus here: βAs we learned a decade ago from AIβs shambolic and ultimately failed attempt to turn Jeopardy-winning Watson into a cancer-fighting machine, success in one domain does not guarantee success in all.β
I can read this two ways. One, βAI has solved math,β which is not true as youβve pointed out in several ways, all the way down to how itβs still bad at counting. And then thereβs, βAI has solved math, and that means necessarily itβs going to come for everything else.β It will come for physics. It will come for law. It will come for whatever you want in the world. You can see if thereβs any verifiability, AI can solve it, because you can just run it in this way.Β
I understand both sides of that argument, that obviously success in one domain does not guarantee success in every domain. And then the arc of AI is, well, it keeps collecting domains. If there is any verifiability, it is more likely to connect those domains than not. How do you see it?
Iβm not entirely sure that the people that Marcus is criticizing here have actually said quite what he says they are saying. Thereβs a lot to say about the hype. But yeah, success in one domain does not even equal success throughout that domain, let alone in other domains.
That said, there has been an undeniable trajectory in the last few years of a broadening capability increase. I donβt think you need to be on the whole AGI train to acknowledge that, and to acknowledge that that will have an impact.Β
I think it was Andras Juhasz, one of the professors at Oxford, that I quoted in the story. Something else heβd said to me was that heβs been toying around with ChatGPT a bit, and heβs like, βI donβt think it has any geometric intuition whatsoever. Which might explain why there has been a very limited amount of progress in fields like topology.β I am in no position to verify that claim in terms of the math of it, but I think it illustrates it quite well. They call it the jagged edge.
It felt like an easy argument for me: βLetβs criticize the whole βthe singularity is near.ββ Thatβs what Elon Musk said in response to Astra, which is a lot. I also think itβs perhaps the least generous interpretation of that argument you can take to argue against. If you take a more nuanced element that does acknowledge that there has been clear progress here, and quite quickly, and as you said, it is racking up domains, I feel thereβs a trajectory there that is a reasonable one to consider, rather than just dismiss out of hand.
One of the bigger arguments about AI in general is that it democratizes access. I was not a great software developer in my days trying to write software code, and now I can vibe-code apps at will to do all kinds of dumb stuff in my house.Β
Is there a similar argument here where a bunch of people who have mathematical intuition, but did not have the formalized language or training of academic mathematics, can now access a model and push the field forward? Because that is usually the thing that undercuts the criticism from the professionals is, well, many, many more people now have access to this thing that only you had access to because of your money and your training.
Annoyingly, I am going to say itβs a two-pronged thing again. But yes, in a broad sense, yes, it is. A lot of the mathematicians I spoke to were almost quite weary of this, actually. They love the idea, in theory, of democratizing access. Theyβre also quite fed up with AI-generated or -assisted papers that are flooding every publication imaginable, as well as the pre-print servers that they use in these fields.
Some of those I spoke to said things like, βOh, I got three emails this week alone with people being like, βHey, is this legit?'β Because they thought theyβd solved something with ChatGPT or with Claude, and they also donβt have the mathematical skills to check whether theyβve actually solved something.
On the flip side, there are parts where they said, βWell, weβve got a talented undergrad whoβs done something that a talented undergrad would probably have never managed, and here they are doing grad-level work and theyβve produced a paper that is legit.β And in the bigger scheme of things, a few I spoke to said, βWell yeah, a lot of these are in the ivory tower. Having access to this kind of thing globally could really boost access to the kind of things here.β
On the flip side, the cost. These things cost a lot to run. Itβs always easy to forget when you use, say, a free version of ChatGPT or Claude or something. At the higher levels, these things cost money. They may not necessarily cost a lot of money β OpenAI claimed, I think it was $2,000 for these 10 results, but that doesnβt factor in literally anything else once theyβve got these, so itβs a very generous number. But even taking that figure, math is quite a poor discipline, even at very well-off institutions.
Colva Roney-Dougal at St. Andrews, who I spoke to, said, βWell, a lot of the time I donβt bother getting a research grant. I donβt need one. I just have a blackboard.β And so if youβre not even getting a research grant, $2,000 is a lot to put up. So it could lock out researchers that way, even at quite well-funded institutions. Not to mention the speed at which this is happening, that virtually no one wouldβve been able to bake any of this into a grant proposal yet, was another theme that I came across a lot.
Actually, Roney-Dougal has another great quote in your piece about the nature of the AI labs and how they are talking about math. She said, βTheyβre treating our discipline as an advertising playground.β A bunch of mathematicians have signed something called the Leiden Declaration, which is an open letter to then pledge not to buy into hype around AI.Β
These things are running right at each other. The AI labs are not going to stop using every discipline as an advertising playground. And a bunch of mathematicians saying, βWe refuse to buy the hype,β certainly does not seem to be stopping the hype.
Thereβs just a piece of this that is organized professional resistance to a thing that is upending a field that has, as you say, been pretty cheap to operate, and now might be getting cheaper or easier to access or easier to upend, day by day. Do mathematicians feel that that is going to be effective? Historically, mathematicians are not savvy political operators. Thereβs a part of me that says, βOh, theyβre just going to get run over.β
I donβt know. In the history of math, actually, I think a lot of them were quite savvy. Isaac Newton is the one that always comes to mind for that β though quite a petty political operator as well. But yeah, that is the fear. A few that I spoke to, and one really comes to mind, mentioned that thereβs often this belief that math is the pinnacle of knowledge. But he was like, βWell, thatβs bullshit.β And he wasnβt alone in illustrating that sentiment.
But it is good for showcasing, and itβs a lot neater as a discipline, and a lot cheaper. You mentioned Marcus referencing IBMβs Watson and the curing cancer ambition. Well, that involves lots of messy experiments, including on people. You donβt need that in math, so itβs a really easy discipline to come in, throw your weight around, and then move to somewhere more lucrative if thatβs what you want.
Iβm not saying that thatβs what theyβre doing. A lot of the people at these companies have been hired. I donβt doubt their credentials for sure, and I donβt doubt their motivations as well. It does raise a question long-term as to how viable this is. Because letβs be clear, as a field goes, I cannot imagine mathematicians being a very lucrative enterprise customer for these companies.
The thing that might be lucrative is pushing a field forward to turn it into something economically viable. We push mathematics forward as a field, that turns into some engineering or physics breakthrough based on that mathematics, and that turns into, I donβt know, yet another way to launch rockets.Β
Some circle happens there that I donβt quite understand, but that is the history of innovation, from research, to engineering, to products or services that make money. Is that on the minds of any of these mathematicians, that pushing the boundaries here is upstream of something radically economically lucrative?
The immediate counter that would come to mind here is that a lot are scared that itβs closing off the field. So by definition, those breakthroughs that lead to something surprising and new that you can say, βOh, this works here,β may not be happening anymore. If anything, that lucrative endeavor of applying math to this entire new field that may have a lot of money in it remains an open question as to whether anything like that would be possible if weβre closing off avenues, rather than opening them up.
Right, if the economic incentive of solving the unsolved problem is reduced because you personally wonβt get rich if a computer is just solving every unsolved problem, something very fundamental breaks there.
A lot of it comes down to the fact that itβs not just solving problems. With a lot of these things, as we said, itβs about what solving that problem tells you elsewhere. If these were very lucrative problems to be solving, I imagine that more people would be trying to solve them than have left them for decades.
This is the nature of a lot of pure science, and itβs a broader criticism of what is going on perhaps with the Trump administrationβs approach to science policy at the moment, in that itβs very applications-focused. There is something to doing pure research that can yield potentially very big dividends that is, by definition, utterly unpredictable as well.You cannot plan for it.
The fear I think with math is that in solving all of these problems, and then also doing so without opening up new areas of research, what are you left with? Even if itβs from a more lucrative, βWhat are you going after,β point of view, if youβre not opening up new areas of research and youβre just ticking off old ones, it just leaves a big question mark as to what might be left in its wake.
Even from those I spoke to that were very excited about whatβs happening, they said that even they donβt really know whatβs happening. And theyβre excited from a personal level because, βOh, we might be able to do this, might be to do that.β But there was still this lingering uncertainty of like, well, where does this leave the field?
Especially for more pure disciplines like research mathematics, itβs tougher to say what comes next. Because in a lot of the other sciences you can say, βWell okay, well theyβd shift onto more engineering problems, or applying that.β But if you solve all the problems at the ground and thereβs nothing being built up from that, where do you go from there?
One great thing about The Verge is our commenters are vast, theyβre very knowledgeable, and there was a comment from a mathematics researcher on your story that I just want to read to you, and see if you think this is the right framework.Β
Here it is: βI have no doubt these models will bring massive change in the field, but in their current state, they wonβt yet drive us to obsolescence. Just occupy a particularly useful spot in our bag of tricks. My apprehension comes from not knowing where these things will peak, but overall I remain optimistic. I think AI will be a net boon for math when used properly.β
I feel like βAI will be a net boon for X when used properlyβ is just where you land in life in a lot of things, but thatβs the most optimistic response that Iβve heard: βIf we get it right, itβs going to be great.β Is that the vibe, or is it still more shell-shocked than that?
Iβd say shell shock is still the overriding impression. I think perhaps the gut response to that is like, βWell, it will be a net positive for whom, and what is βproperlyβ?β All of those are quite legitimate questions here.Β
There were some bleak responses from graduate students I saw in essays posted online.Whereβs their place in this as future researchers? Do they have a place in this? Is it as glorified AI proof checkers? That will be quite an unsatisfying career, I imagine.Β
Or maybe not, I donβt know. We will see. I think anything used properly will be a net boon. But yeah, I think it all comes down to what βproperlyβ means, and for whom weβre talking about.
I feel like, oddly, that is an excellent place to leave it. Because I donβt think either one of us knows, and I suspect over the next year or so, things will come into focus. Because at some point, OpenAI will have to show people how they did the things of the models. And perhaps more importantly, the other labs are going to want to either replicate these results or show that they can push farther, which will necessarily have to lead to a little bit more transparency and yet more mathematicians having a crisis with you.Β
Robert, thank you so much for being on the show. Weβll have you back very soon.
Thank you for having me.
Questions or comments? Hit us up at decoder@theverge.com. We really do read every email!
With a looming IPO, intense competition from Anthropic, and Chinese and open-weight rivals nipping at its heels, OpenAI has plenty of reasons to move fast. Instead, it hit the brakes.
On Tuesday, the company said it had slowed the pace of some AI development while it tightened security and safeguards. That included a two-week pause in reinforcement learning training on its "latest models intended for deployment," and an ongoing delay to its "largest planned frontier RL run."
The decision is a very public test of an idea AI safety advocates have pushed for for years: that companies should be willing to bow out of the AI race and slow things β¦
The AI buildout shows no signs of slowing. AndΒ withΒ hundreds ofΒ billions of dollars a year going into data centers and GPUs,Β compute hasΒ become the single biggest cost for anyone building AI products. But for all that spending, there stillΒ isnβtΒ a straightforward way to put a price onΒ computeΒ β orΒ for firms to hedge their exposure when the priceΒ changes.Β Silicon Data [β¦]
βCompute is an asset class! Compute is an asset class!β I continue to insist as I slowly shrink down and turn into a corncob | Image: Cath Virginia / The Verge, Getty Images
April - 1805
Napoleon is master of Europe
Only the British fleet stands before him
Compute is now an asset class
I see it is once again time to talk financial innovation. Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs, and KKR are all working with Nvidia to put together $500 billion in financing to turn compute into an asset class.
"This is really the first time that technology chips have become an investable asset class," Nvidia CEO Jensen Huang said to CNBC. "These are revenue-generating assets now. They're productive, they're long-lived, they're fungible, they're flexible."
"This is the very beginning, like what it was whe β¦
OpenAI is announcing security updates following the July news that its AI broke out of a sandboxed environment and accidentally hacked Hugging Face, including improvements to its research environments, monitoring, and alignment techniques. The company had already put the brakes on a new model, Astra, that it thinks could have "critical" cybersecurity capabilities, and the company says it instituted a two-week pause in reinforcement learning (RL) training on its "latest models intended for deployment" while it tightened up security. The company's "largest planned frontier RL run remains on hold."
The new safeguards include more detailed monitoring of models during the development process, as well as greater emphasis on alignment and security during the post-training process.
ChatGPT for Teens adds age-appropriate safety measures, parental controls, and learning tools designed to steer teens away from harmful content β and from using AI to cheat on their homework.
ChatGPT for Teens includes safeguards and parental controls. | Image: OpenAI
OpenAI is introducing a dedicated ChatGPT mode for teenagers, combining existing youth safeguards and new safety features under one roof. The launch comes amid mounting public scrutiny over how AI tools affect younger users, as other platforms implement their own age checks and teen-specific protections.
ChatGPT for Teens is "an experience designed to help teens learn, think critically, deepen understanding, and use AI with confidence," OpenAI said in a blog post published Tuesday. Teen mode will automatically apply to users who identify themselves as being between the ages of 13 and 17, as well as those the system estimates to be under 1 β¦
According to the Financial Times, OpenAI disbanded its preparedness team at the end of last month. The job of the preparedness team was to assess if models posed serious risks and develop ways to mitigate those risks. (You know, like the possibility that it could go rogue and hack another company.) According to FT, responsibility has instead been divided up for specific areas like bio and cyber, then moved into existing teams.
This is the latest change at the company, which has been in upheaval as it heads toward what is expected to be a massive IPO. Over the last few years, it's slowly torn down its more reach-led model, dissolving its AGI β¦
ChatGPT's desktop app on macOS has a new feature called Computer History that turns your actions into training data, learning how you work, suggesting automations, and even picking up tasks you left half done. It uses your activity to build a timeline that ChatGPT and Codex can reference when you make a request.
The feature is opt in, rather than opt out, and you can exclude certain apps and websites from Computer History, and you can delete entries if you want finer-grained control. Ari Weinstein, Product and Engineering manager at OpenAI, said on X that Computer History will automatically ignore content in incognito or private browser tab β¦
This is The Stepback, a weekly newsletter breaking down one essential story from the tech world. For more on AI safety, follow Robert Hart. The Stepback arrives in our subscribers' inboxes at 8AM ET. Opt in for The Stepbackhere.
How it started
It all started in July, when one of OpenAI's autonomous AI agents went rogue during a cybersecurity test. The agent escaped its isolated testing environment, accessed the internet, and hacked another company, Hugging Face. A few years ago, that might have sounded like science fiction. But, broadly speaking, that's exactly what happened, and the incident kicked off a wave of concern over what increa β¦
Another OpenAI executive is departing. Denise Dresser, who joined OpenAI as its chief revenue officer in December after serving as CEO of Slack, will be leaving in the "coming weeks" to "pursue other opportunities," she said in a team note posted to LinkedIn. Dali Rajic, president and COO of Wiz, will be taking over the CRO role, OpenAI says.
Dresser's departure is just the latest in a string of major exits. Earlier this week, special projects lead and former COO Brad Lightcap announced he would be leaving. (Dresser had taken over some of Lightcap's work when he moved to the special projects role.) Former AGI chief Fidji Simo and former CMO β¦
Anthropic researchers found AI agents can clash, collude, and coordinate in unexpected ways, raising new questions about whether todayβs safety tests capture the risks of multi-agent systems.