Posts / artificial-intelligence

Did AI Just Solve Navier-Stokes, Or Did We Just Solve a Very Expensive Homework Question?


Spent a chunk of my Sunday morning down a rabbit hole that started with r/singularity and ended with me reading a Millennium Prize problem statement I barely understood, which is roughly how most of my weekends go these days.

The short version: OpenAI apparently threw a swarm of AI agents, something like ten thousand of them, at the Navier-Stokes existence and smoothness problem. That’s one of the seven Clay Millennium Prize problems, a million dollars each, and only one has ever been solved (Poincaré, back in 2003, by a Russian mathematician who then refused the prize money and the Fields Medal because he’s apparently made of sterner stuff than the rest of us). The agents reportedly burned through millions of dollars in compute and came back with a proof that disproves the smoothness conjecture, meaning under certain very specific starting conditions, the equations can blow up to infinity in finite time.

The subreddit’s reaction split roughly into two camps, and it’s a genuinely interesting split, not just AI-hype-bros versus AI-sceptics.

Camp one: this is a historic moment, the day AI surpassed the collective mathematical achievement of humanity, we are so back, etc. Camp two, led by an actual aerospace engineering professor who turned up to hose things down: take a breath, this is a narrow result about a specific pathological edge case, it doesn’t mean anything for how planes fly or how weather models work, and headlines saying “AI solved Navier-Stokes” are doing a lot of misleading work with that word “solved.”

Here’s the thing that got me though. Reading through the comments, a mathematician showed up and made a point that I think both camps needed to hear: the professor’s dismissal (calling it a “mathematical curiosity”) undersells how significant this is as mathematics, even while being completely correct that it changes nothing about aircraft design or your Bureau of Meteorology forecast tomorrow. Both things are true at once. It’s a genuine, first-rate achievement in a field that has resisted the best human minds for over a century, and it has approximately zero practical consequences for anyone building anything. Pure maths people already knew this tension; it comes with the territory. The rest of us aren’t used to holding both ideas in our head at the same time, especially not when there’s a press release involved.

That’s sort of the story of AI coverage generally, isn’t it. Something genuinely remarkable happens in a lab, it gets compressed into a headline for people who don’t have the background to parse the actual claim, and then a fight breaks out between people defending the headline and people defending the nuance, and everyone leaves more confused than when they started. I do this for a living, sort of, in a much smaller way. I’ve watched a demo of some internal tool get described in a company all-hands as “revolutionary” when what it actually did was save four people twenty minutes a day. Both descriptions were technically defensible. Neither was the full story.

What actually stuck with me wasn’t the maths, it was the money. Someone did the back-of-envelope calculation: if the compute cost was somewhere around fifteen to twenty million dollars, and the prize is one million, and OpenAI has already said they’re not going to bother claiming it, then the entire exercise was never about the money. It was a demonstration. A very expensive, deliberately public flex, aimed less at the mathematics community and more at investors, competitors, and anyone still wondering whether the next funding round is justified. I don’t say that cynically, exactly, I say it because I think it’s just true, and worth noticing separately from whether the maths itself is impressive. It clearly is.

And that’s where my worry lives, the one that doesn’t go away no matter how many times I read these threads. Fifteen million dollars of compute to disprove one conjecture about one class of edge-case fluid behaviour is an amount of electricity and water and silicon that has a footprint, and we’re going to see a lot more of this kind of thing: agent swarms thrown at problems not because it’s the most efficient way to solve them, but because it’s a good story, and because whoever gets there first gets to write the headline. I don’t know how to feel about that. I’m fascinated by what these systems can do, genuinely, unreservedly fascinated, and I also think we’ve built an entire economy around demonstrations that cost more than they’re worth in any conventional sense, and we’ve decided that’s fine because the demonstration itself is the product.

Maybe it is fine. Maybe this is just what progress looks like from up close, messy and overhyped and expensive and real all at once. The mathematician was right about the achievement. The professor was right about the plane. I don’t think either of them needed to lose the argument for the other to be correct, and I reckon that’s the actual lesson buried in that thread, more than anything about fluid dynamics: hold two true things at once, and be suspicious of anyone who only wants you to hold one.