Gary Marcus
@GaryMarcus
The real news here isn’t the result; it’s what we were not told.
1. AI once tried to be a science. Now we get stuff like the completely vague report from OpenAI below, and a lot of ignorant questions from people who don’t know how to think critically.
“Same procedure”? “using an unreleased model”?
This would never pass peer review.
We don’t know what the procedure was.
We know nothing about the architecture (e.g., were proofs generated in one shot, and then verified by Lean? was there an iterative process?).
We know nothing about the failure rate. We know nothing about the training/post training/data agumentation.
2. As a result we have zero idea of how generalizable the result is outside math.
3. A lot of X has been reduced to an ignorant cheering section that applauds without knowing what it is applauding or what it might mean — without ever asking basic scientific questions.
The new system could be a legitimate step towards AGI or just a clever leveraging of Lean and synthetic data in a verifiable domain with no generality whatsoever; from this report we can tell almost nothing.