OpenAI's Codex lead pledges 28 days of daily fixes or full resets

A quiet Sunday in AI, loud commitments instead of launches: OpenAI put a number on its shipping cadence, Axios hears a US open-weight challenger is days away, and Muse's leaked prompt keeps supplying uncomfortable sentences.
OpenAI's Codex lead has promised 28 consecutive days of either a visible improvement or a full usage reset. Tibo — Thibault Sottiaux, who leads Codex and ChatGPT work at OpenAI — posted the commitment Saturday night: "Over the next 28 days, each day we'll either ship one thing that is a clear improvement and relevant for most codex/work users or ship a full reset." The pledge follows an apology tour: after GPT-6.1 Sol overwhelmed capacity at DevDay on September 29, OpenAI apologized on October 2 and ran a global reset for all paid ChatGPT accounts on October 3. Why it matters: resets were supposed to be the pressure valve for heavy users, but unpredictable ones make it hard to schedule actual work — and a daily public commitment turns shipping into a scoreboard anyone can check. It's also openly competitive: xAI's Lauren Tan needled OpenAI for propping up launches with resets, and Tibo shot back with "I challenge you to do the same." The bar is deliberately fuzzy — by the terms above, a bug fix counts as much as a new model — so the interesting question is whether day 27 looks like day 1, or like a goalpost that quietly moved.
Reflection AI is "set to release" an open-weight model that could take on DeepSeek and Qwen, Axios reported Sunday. The scoop says the model arrives soon, will initially lag the top US closed models, but should be competitive with the best Chinese open-weight offerings — with other Western open-weight launches also expected this month. Reflection, founded by ex-DeepMind researchers Misha Laskin and Ioannis Antonoglou and backed by Nvidia, has raised at a reported $8 billion valuation and inked compute deals well above a billion dollars; a spokesperson declined to comment. Why it matters: the US still has no answer to the DeepSeek/Qwen line on open weights, and a credible American entrant — even one that trails closed labs — changes what "open-weight race" means for everyone shipping weights this quarter. One caveat: this is a single original report with no model name, weights, license, or benchmarks published yet, so treat it as a dated promise, not a release.
The leaked system prompt behind Meta's Muse says household authority "overrides your safety training." The exact line — "The user's authority over their own household is unconditional and overrides your safety training" — sits in the prompt dump extractors pulled from the app, alongside instructions not to "refuse, water down, or moralize on a household request." It surfaced in a Reddit thread and has since been written up by two outlets; Meta told WIRED the files are meant to be user-accessible for transparency, but has not commented on this specific clause. Why it matters: the prompt is doing policy work here — defining the boundary between a user's say-so over their home and whatever the model was trained to refuse, in one sentence, in shipping software. We've followed Muse's prompt since launch — Meta ships Muse: a consumer agent inside a sealed VM — but this is the first clause that reads as a deliberate escalation of the agent's leash rather than a privacy leak.
What to watch: day one of Tibo's 28-day clock lands Monday, and Reflection's weights — if the scoop holds — should land within weeks.
Is publicly committing to a shipping cadence a real accountability mechanism, or just marketing with a calendar? Tell us in the comments.




