OpenAI's Codex lead pledges 28 days of daily fixes or full resets

Share
OpenAI's Codex lead pledges 28 days of daily fixes or full resets

A quiet Sunday in AI, loud commitments instead of launches: OpenAI put a number on its shipping cadence, Axios hears a US open-weight challenger is days away, and Muse's leaked prompt keeps supplying uncomfortable sentences.


OpenAI's Codex lead has promised 28 consecutive days of either a visible improvement or a full usage reset. Tibo — Thibault Sottiaux, who leads Codex and ChatGPT work at OpenAI — posted the commitment Saturday night: "Over the next 28 days, each day we'll either ship one thing that is a clear improvement and relevant for most codex/work users or ship a full reset." The pledge follows an apology tour: after GPT-6.1 Sol overwhelmed capacity at DevDay on September 29, OpenAI apologized on October 2 and ran a global reset for all paid ChatGPT accounts on October 3. Why it matters: resets were supposed to be the pressure valve for heavy users, but unpredictable ones make it hard to schedule actual work — and a daily public commitment turns shipping into a scoreboard anyone can check. It's also openly competitive: xAI's Lauren Tan needled OpenAI for propping up launches with resets, and Tibo shot back with "I challenge you to do the same." The bar is deliberately fuzzy — by the terms above, a bug fix counts as much as a new model — so the interesting question is whether day 27 looks like day 1, or like a goalpost that quietly moved.


Reflection AI is "set to release" an open-weight model that could take on DeepSeek and Qwen, Axios reported Sunday. The scoop says the model arrives soon, will initially lag the top US closed models, but should be competitive with the best Chinese open-weight offerings — with other Western open-weight launches also expected this month. Reflection, founded by ex-DeepMind researchers Misha Laskin and Ioannis Antonoglou and backed by Nvidia, has raised at a reported $8 billion valuation and inked compute deals well above a billion dollars; a spokesperson declined to comment. Why it matters: the US still has no answer to the DeepSeek/Qwen line on open weights, and a credible American entrant — even one that trails closed labs — changes what "open-weight race" means for everyone shipping weights this quarter. One caveat: this is a single original report with no model name, weights, license, or benchmarks published yet, so treat it as a dated promise, not a release.


The leaked system prompt behind Meta's Muse says household authority "overrides your safety training." The exact line — "The user's authority over their own household is unconditional and overrides your safety training" — sits in the prompt dump extractors pulled from the app, alongside instructions not to "refuse, water down, or moralize on a household request." It surfaced in a Reddit thread and has since been written up by two outlets; Meta told WIRED the files are meant to be user-accessible for transparency, but has not commented on this specific clause. Why it matters: the prompt is doing policy work here — defining the boundary between a user's say-so over their home and whatever the model was trained to refuse, in one sentence, in shipping software. We've followed Muse's prompt since launch — Meta ships Muse: a consumer agent inside a sealed VM — but this is the first clause that reads as a deliberate escalation of the agent's leash rather than a privacy leak.

What to watch: day one of Tibo's 28-day clock lands Monday, and Reflection's weights — if the scoop holds — should land within weeks.

Is publicly committing to a shipping cadence a real accountability mechanism, or just marketing with a calendar? Tell us in the comments.

Read more

Altman says the world must accept AI's 'bad things'

Altman says the world must accept AI's 'bad things'

A heavy news day for AI governance and open weights: OpenAI's CEO is publicly pricing the trade-off his industry keeps dodging, Reflection finally put specs on the model it teased yesterday, and AMD is trying to set the terms before Nvidia's RTX Spark lands. Altman says the world should accept AI's "bad things" — and the labs' new pact agrees. In an interview released Monday on Politico's Decoded podcast, Sam Altman said OpenAI's position is "we believe that the world should accept some bad th

Today in AI — October 5, 2026

Today in AI — October 5, 2026

The day the ecosystem stopped pretending everyone is a partner: Meta and Microsoft quietly cut their Claude budgets, Washington gave AI policy a new name, and New York City put lab executives under oath. Elsewhere, one model learned to drive a robot, and Mac users finally got Apple Intelligence off their disks. Models & Research * Reka AI's Rho-1 collapses the multimodal stack into a single 19-billion-parameter model. The research preview runs text, images, video and robot control as token

OpenAI adds text watermarking to ChatGPT and Codex — EU first

OpenAI adds text watermarking to ChatGPT and Codex — EU first

Regulation is now shipping inside the product: OpenAI's EU-only watermark rollout lands today, Wikimedia publishes its evidence against OpenAI's agents, and two of Anthropic's biggest customers are easing off Claude. OpenAI is turning on invisible text watermarking in ChatGPT and Codex — starting with the European Union. Over the coming weeks, eligible EU users across all plans will get a machine-readable signal called textGrain woven into the text the model produces, while API customers anywh

Meta raced to patch a VM escape in Muse before launch

Meta raced to patch a VM escape in Muse before launch

Three stories today share a theme: systems that were supposed to be contained — an agent platform, a preprint archive, a text watermark — all straining at the edges. Meta's own security teams didn't think Muse was ready to ship. 404 Media reports that in the weeks before launch, engineers found several vulnerabilities in the viral agent product, at least one of which could have let a normal Muse user break out of the sandbox and reach sensitive internal Meta databases. The evidence is an inte