The question is: are those real proofs, or just AI reasoning bugs again?
Ok so? Do the proofs hold up or not?
We’re reaching a point where the interesting part isn’t just whether an AI found the proof: it’s whether anyone outside the company can reproduce the result. Publishing Lean proofs is great. Keeping the model closed means the process stays a black box.
I don’t understand enough about math and lean proofs - do you mean it’s possible nobody can interpret and explain how the proof goes in “human language”?
The impressive part isn’t that an AI produced a proof, it’s that Lean lets everyone verify it. The frustrating part is the model stays closed. Science advances fastest when others can reproduce both the result and the method, not just inspect the finished homework.
But they do plan to release the Astra model to the public “after it’s finished” I assume?
And yeah this Lean software sounds amazing, didn’t know this was possible.
I misread your last sentence in a way that create a different problem: Imagine AI starts producing more and more proofs at a pace that humans can’t keep up with. So science advances, and humans can even make use of it, but can’t verify or really understand the theory any more. Like we get better working machines, materials and processes, but do not understand why because we can’t keep up. If that happened then that I guess would a significant stage in the singularity.
It would be surprising if they didn’t release Astra. And yes, we’re living in crazy times.
In other words its all marketing bullshit.
So youre suggesting they secretly have Einstein standing in the server rack, making loud fan noises with his mouth and just typing really fast?
Either the problems were solved or they weren’t. If they were, then that’s evidence enough; doesn’t matter if they model isn’t publicly available.
Having those sudden breakthroughs come from a person or even a large group of mathematicians, suddenly and all at once, would be more surprising than a well-harnessed LLM figuring it out.
With that said, I’m curious about peer review of the actual proofs. Just because Lean builds them doesn’t mean they’re materially valid. It could very well be that it’s completely wrong and it just hallucinated well enough to fool OpenAI into publishing it, which would be a hilarious egg-on-face moment
I agree with you. If it solved it, it solved it. The title is an odd phrasing.
The thumbnail itself is the silliest ai shit. So his shirt becomes his arm then turns into his shirt at the wrist. And thats why these idiots are so mad. They are too stupid to see how actively bad it all is, but that’s your fault for being perceptive, not theirs for being a compulsory idiot. These people are so lucky breathing doesn’t require critical thought.
To believe anything these companies claim about their model is to demonstrate you’re a gullible simpleton. You’re exactly the type of non-critical thinker they need to thrive.
Holy shit. If an AI actually solves a Millennium Prize Problems we’re seriously entering superintelligence (but not AGI) territory.
In any case this should disprove the “AI is just a statistically typing monkey” theorem some AI skeptics are fond of.
The current examples of AI could never be described as a general intelligence let alone Superintelligence. These LLMs are definitely just a typing monkey (that was eloquently put).
They are literally generating text word by word with a next-word-prediction methodology.
With some tweaking they can be made to produce mostly coherent code and apparently also equations that balance. This does not make these algorithms geniuses. A calculator can also balance an equation, so that alone is not impressive.
The techniques LLM designers use to leverage statistical models in order to generate coherent sentences are actually rather impressive. It is legitimately quite a feat to make LLMs work as well as they do. But that does not mean that these outputs are anything like “thinking”, they’re just outputs. Data goes in, the machine churns, does some mystery math, and different data comes out.
Ok, not sure I understand “AI is just a statistically typing monkey”? As far as I know AI simplified is just predicting what is statistically the most likely next word or is that wrong? I mean sure the way it predicts is complicated and word vectors are fascinating etc but highly simplified its a statistically typing monkey?
It’s a “category error” like saying hydrodynamics are just quantum mechanics. It does not describe the complex emergent behavior and doesn’t really tell you anything. Or it’s like saying psychology is how the neurons and synapses connect with each and exchange signals and that is all - this would not explain the mental processes between a marmot and a human. So it’s not an accurate simplified description, just a description of the basic building blocks.
Afaik nobody knows how LLMs actually think, we only know sort of how to train them. They are researching how LLMs think with an equivalent to MRI brain scanning: Vista do LLM-MRI Python module: a brain scanner for LLMs
The Decoder reported on August 1 that Brown said the lab had not spent much on each problem and that there were “no Millennium Prize Problems (yet)”.
Just to be clear.
But yeah, if that happens, it will certainly be a major milestone.
Nothing can disprove that to that crowd, I’m afraid.
They think the problem is AI, when it’s actually capitalism and securities fraud.
It’s much simpler to fixate on the technology than to have to understand complex things like economics and history.



