OpenAI drops 372 math results, nearly all from one prompt

Share
OpenAI drops 372 math results, nearly all from one prompt

A dense news day across three fronts: OpenAI published a mountain of mathematics from an unreleased model, SpaceX reportedly wants the debt market to fund its Nvidia orders, and Anthropic restructured its cyber program around three tiers of trust.

OpenAI published 372 mathematical result families from an internal frontier model — and says nearly all of them came from a single prompt handed to a single AI agent. The release comprises 722 manuscripts touching hundreds of open questions, plus Lean formalizations that let a computer check the proofs. In a nod to the Advisory Group on Mathematics and Artificial Intelligence at the Institute for Advanced Study, OpenAI also shipped process details it usually keeps quiet: ten summaries of the model's reasoning, compute estimates (the average result consumed roughly three hours of ChatGPT Pro thinking), and statistics on attempted problems. The headline artifact sits in the repository as preprint number 109: integer multiplication below n log n, which the authors say disproves the Schönhage–Strassen optimality conjecture — a benchmark of fast multiplication for decades. Skepticism is running ahead of celebration: MIT mathematician Andrew Sutherland told Scientific American to treat claims of one-shotting problems with a single agent as unverified and to "ask for receipts," and OpenAI published only average compute, not the prompts. Our take: every number here is self-reported and unreplicated, but if even a fraction survives review, this is the first large batch of research results where a lab's unreleased model — not its shipping product — is the protagonist.


SpaceX is reportedly seeking to raise $40 billion — about $10 billion in bank loans and $30 billion in investment-grade debt — to buy Nvidia chips, with Apollo leading the deal. The Financial Times broke the story, and Reuters and Bloomberg both carried it within hours, each citing FT; SpaceX, Apollo, Nvidia and Pimco have not confirmed it. If the number holds, it would be the largest AI-chip financing yet, and the structure is the story: the compute buildout is being funded by debt rather than equity, with the deal expected to close in 2027.


Anthropic folded Project Glasswing into an expanded, three-tier Cyber Verification Program that gives vetted security teams its most capable models with progressively fewer safety blocks. Defense Access covers incident response and malware analysis with a review turnaround of days; Red Team Access adds authorized offensive testing; the top tier — Specialized Access, for testing systems like flight operations and power grids — is vetted in collaboration with the US government, and existing Glasswing members move over without reapproval. All tiers get Claude Opus 5.5, Sonnet 5.5 and Mythos 5.1. Anthropic says Glasswing partners found 129,000 verified vulnerabilities between April and July, and its own benchmark draws the tier line sharply: every cyber task was blocked on the first prompt for users without program access, while top-tier access completed 34 of 50 runs with zero blocks.

What to watch: whether any mathematician starts replicating the OpenAI results — and whether the SpaceX figure survives confirmation beyond FT's sources.

Three hours of ChatGPT Pro thinking per theorem, and the prompts stay private — would you trust a proof you can't audit? Tell us in the comments.

Read more

Akhetonics says its all-optical CPU reaches a customer in 2026

Akhetonics says its all-optical CPU reaches a customer in 2026

Light-based computing keeps promising more than it delivers — but one Munich startup has just put a date on its bet, and the interview laying it out is doing the rounds on Hacker News this week. Akhetonics says it will deploy its first commercial machine with a major customer by the end of 2026, with several more planned for 2027. The company, founded by Michael Kissner and Leonardo Del Bino, is building a computer where data enters as light, is switched as light, and circulates through memory

The Week in AI — October 5–11, 2026

The Week in AI — October 5–11, 2026

Every big claim this week turned out to rest on fine print more interesting than the headline: revenue only the company reporting it can define, safety tests sandboxed while the product keeps the web, and a Pentagon phase-out nobody would confirm until reporters kept asking. The week's top 5 1. OpenAI's revenue was $20 billion below the numbers everyone quoted — and the gap was definitional. The Financial Times reported Thursday that OpenAI's annualized revenue runs roughly $20 billion unde

Drone strike shuts a third Yandex data center, taking YandexGPT offline

Drone strike shuts a third Yandex data center, taking YandexGPT offline

Russia's largest tech company is learning what the AI era's infrastructure war looks like from the receiving end — three data centers in four days, and with them much of the cloud layer Russian businesses run on. A Ukrainian drone strike knocked out Yandex's data center in Vladimir early Sunday morning, the third of the company's facilities hit since October 8. The site — reported at roughly 50 MW and designed for about 2,880 server racks — stopped operating completely after the attack, Yandex

Agent teams cost up to 5x more, barely score higher

Agent teams cost up to 5x more, barely score higher

The multi-agent hype train hit a benchmark this weekend — and the grid and the trucking regulators had quiet weeks of their own. Vals AI put agent teams head-to-head with single agents on its Vibe Code Bench, and the teams cost between 1.8 and 5.1 times more for almost no extra quality. The evals company ran GPT-6 Sol and Claude Opus 5.5 solo and in teams across 50 apps at two reasoning efforts; out of four comparisons, only one was statistically significant — Sol at medium effort, where the t