OpenAI drops 372 math results, nearly all from one prompt

Share
OpenAI drops 372 math results, nearly all from one prompt

A dense news day across three fronts: OpenAI published a mountain of mathematics from an unreleased model, SpaceX reportedly wants the debt market to fund its Nvidia orders, and Anthropic restructured its cyber program around three tiers of trust.

OpenAI published 372 mathematical result families from an internal frontier model — and says nearly all of them came from a single prompt handed to a single AI agent. The release comprises 722 manuscripts touching hundreds of open questions, plus Lean formalizations that let a computer check the proofs. In a nod to the Advisory Group on Mathematics and Artificial Intelligence at the Institute for Advanced Study, OpenAI also shipped process details it usually keeps quiet: ten summaries of the model's reasoning, compute estimates (the average result consumed roughly three hours of ChatGPT Pro thinking), and statistics on attempted problems. The headline artifact sits in the repository as preprint number 109: integer multiplication below n log n, which the authors say disproves the Schönhage–Strassen optimality conjecture — a benchmark of fast multiplication for decades. Skepticism is running ahead of celebration: MIT mathematician Andrew Sutherland told Scientific American to treat claims of one-shotting problems with a single agent as unverified and to "ask for receipts," and OpenAI published only average compute, not the prompts. Our take: every number here is self-reported and unreplicated, but if even a fraction survives review, this is the first large batch of research results where a lab's unreleased model — not its shipping product — is the protagonist.


SpaceX is reportedly seeking to raise $40 billion — about $10 billion in bank loans and $30 billion in investment-grade debt — to buy Nvidia chips, with Apollo leading the deal. The Financial Times broke the story, and Reuters and Bloomberg both carried it within hours, each citing FT; SpaceX, Apollo, Nvidia and Pimco have not confirmed it. If the number holds, it would be the largest AI-chip financing yet, and the structure is the story: the compute buildout is being funded by debt rather than equity, with the deal expected to close in 2027.


Anthropic folded Project Glasswing into an expanded, three-tier Cyber Verification Program that gives vetted security teams its most capable models with progressively fewer safety blocks. Defense Access covers incident response and malware analysis with a review turnaround of days; Red Team Access adds authorized offensive testing; the top tier — Specialized Access, for testing systems like flight operations and power grids — is vetted in collaboration with the US government, and existing Glasswing members move over without reapproval. All tiers get Claude Opus 5.5, Sonnet 5.5 and Mythos 5.1. Anthropic says Glasswing partners found 129,000 verified vulnerabilities between April and July, and its own benchmark draws the tier line sharply: every cyber task was blocked on the first prompt for users without program access, while top-tier access completed 34 of 50 runs with zero blocks.

What to watch: whether any mathematician starts replicating the OpenAI results — and whether the SpaceX figure survives confirmation beyond FT's sources.

Three hours of ChatGPT Pro thinking per theorem, and the prompts stay private — would you trust a proof you can't audit? Tell us in the comments.

Read more

Underdog launches a private on-device AI assistant, backed by a16z

Underdog launches a private on-device AI assistant, backed by a16z

The privacy split in consumer AI got a new entrant tonight, GitHub turned code-review benchmarks into a vendor-bias argument, and OpenAI opened its oddest API to everyone. Underdog launched in invite-only beta: an on-device AI assistant that keeps your data on your machine and never charges you a subscription. Self-taught coder and Thiel Fellow Sigil Wen — who moved to Silicon Valley at 17 and lived in an AI hacker house with Andrej Karpathy — built his own inference engine, Husky, to run a 2

Today in AI — October 6, 2026

Today in AI — October 6, 2026

A day of second-order moves: the labs are buying task data from software vendors instead of scraping the web, Waymo is borrowing to fund the robot world, and the biggest bank in the US just put a number on what Anthropic's latest model cost it in risk. Models & Research * OpenAI is training GPT-6 Astra on Ironclad's real contracting work. Ironclad staff helped turn 11 tasks across legal, commercial and procurement work — setting up NDAs, approval workflows, clauses that change by jurisdict

Meta, Walmart and Stripe publish the Personal Agent Protocol

Meta, Walmart and Stripe publish the Personal Agent Protocol

The agent economy is writing its rulebook tonight: one open standard for AI bots at the checkout, a cheaper image model from Google, and Anthropic turning bug-hunting into a tiered product. Meta, Walmart, Stripe and Sierra are publishing an open "personal agent protocol" — a standard that defines how personal AI agents interact with businesses online. The group behind it reads like a cross-section of agentic commerce: Meta, Sierra, Walmart, Stripe, Shopify, Genesys, Rocket, NiCE, Decagon and I

Lambda raises up to $4B from Blackstone ahead of its IPO

Lambda raises up to $4B from Blackstone ahead of its IPO

The neocloud money is consolidating fast, and today's inbox shows both ends of the market: a heavyweight pre-IPO round on one side, and a Google open model you can run on a phone on the other. Lambda is raising up to $4 billion led by Blackstone and Coatue at a $14.5 billion pre-money valuation — its last private round before a planned IPO. The Wall Street Journal reported the scoop from a letter to limited partners, and Reuters independently confirmed the headline terms: the round is led by t