
OpenAI Forms Independent Math Advisory Group After Claiming 100+ Open-Problem Solutions
The AMW Read
OpenAI's case-study capability claims (Navier-Stokes, 100+ solved problems) draw a governance response to Fields Medalist pushback, extending a pattern of contested benchmark claims already visible in the ARC-AGI dispute.
OpenAI Forms Independent Math Advisory Group After Claiming 100+ Open-Problem Solutions
OpenAI has established the Advisory Group on Mathematics and Artificial Intelligence, hosted at the Institute for Advanced Study in Princeton, giving outside mathematicians a channel to weigh in on the company's math-research output. The move follows OpenAI's abrupt claim that its internal model solved the Navier-Stokes Millennium Prize problem, plus more than 100 additional previously open problems across most areas of mathematics. Nine mathematicians were named as initial members; unpaid but independent, they can publish their own views and control future membership, though the group has no authority to slow or redirect OpenAI's internal research pace.
The advisory structure is a direct response to friction rather than a routine research update. Earlier this month, 25 Fields Medal-winning mathematicians signed an open letter warning that AI labs racing to claim priority on famous problems are straining the field's normal verification norms — and only one of OpenAI's nine appointees, IAS's Camillo De Lellis, also signed that letter, underscoring how contested the roster itself is. It also lands days after ARC Prize's own verification showed GPT-6 Astra's headline 99.9% ARC-AGI-3 score collapsing to 62.7% once OpenAI's proprietary evaluation harness was removed — a second recent case of a marquee OpenAI capability claim drawing independent scrutiny rather than acceptance. OpenAI-linked coverage in AI Market Watch's pipeline rose to 340 items over the past 90 days from 289 prior (name-matched, pipeline-ingested sources only), consistent with a company generating an unusually high rate of contested claims.
For builders and investors, the signal is procedural: an advisory board with no veto power buys reputational cover without ceding control over publication timing or which results get emphasized, so external benchmark and proof verification — not the lab's own framing — remains the credible checkpoint before treating any single math or capability claim as settled.

