← All posts / Tools

Ship or Reset: OpenAI's Codex Lead Bets 28 Days on Daily Improvements or Full Quota Resets

Tibo Sottiaux, who runs Codex and ChatGPT Work at OpenAI, has pledged that each of the next 28 days brings either one clear improvement for most users or a full usage reset — turning weeks of quota complaints into a public daily scoreboard.

Ship or Reset: OpenAI's Codex Lead Bets 28 Days on Daily Improvements or Full Quota Resets

Developer tools rarely make promises with deadlines attached. On October 5, 2026, OpenAI did exactly that. Tibo Sottiaux — the OpenAI executive who runs Codex and ChatGPT Work, and one of the most visible product voices at the company since the DevDay barrage — posted a pledge on X with a built-in scoreboard: “Over the next 28 days, each day we’ll either ship one thing that is a clear improvement and relevant for most codex/work users or ship a full reset. Let the improvements begin.”

Within hours the post passed a million views, and the replies split almost evenly between cheering and sarcasm. Depending on who you ask, it is either the most refreshing piece of accountability a major AI lab has offered its developers, or a sign that OpenAI has been paying down trust with so many usage-limit resets that it finally had to make them official policy.

What was actually promised

The pledge came in two posts, roughly a day apart. The first set the tone: Sottiaux said the team was “locking in,” and that the only things being worked on are “simplifications, more efficiency for more usage, groundbreaking features or new models.” He added that “sometimes you have to invest ahead of the curve, but feedback is clear that you all want things to get simpler.”

The second post, quoting the first, is the commitment itself: 28 days, one improvement or one full quota reset per day, “relevant for most codex/work users.”

Three details matter more than the headline. First, “relevant for most users” sets a real bar — a niche CLI flag does not qualify under the text of the pledge. Second, the reset is a fallback, not a bonus; the phrasing implies the team expects some days to have no qualifying improvement ready. Third, “the only things being worked on” reads like an internal scope freeze — an implicit admission that feature sprawl and unclear limits are the main sources of user pain.

One correction that traveled less far than the hype: this is not “28 resets.” Multiple viral summaries flattened the either/or structure into a month of free refills. It is one or the other, each day, and resets are one-time refills of your weekly bucket, not a higher ceiling.

Why this landed now

The pledge is a response to a brutal month, and the timeline explains it. On September 25–26, Codex and ChatGPT Work suffered an outage, followed by a limits reset for paid accounts. Days later came a GPT-6 Astra quality postmortem and another reset. At DevDay on September 29, OpenAI shipped Dots (the always-on agent), the Pro 500 tier, and an Ultrafast speed tier, while reopening Pro 200 with a smaller usage formula. From October 1 to 3, the GPT-6.1 Sol launch drove a massive load spike, and Sottiaux announced a global reset for all paid ChatGPT accounts. On October 3, Claude Opus 5.5 took first place on the Epoch Capabilities Index, 0.84 points ahead of GPT-6 Astra — a narrow margin, but a very quotable one for Anthropic. And on October 4, Sottiaux teased “6.1 coming soon.”

Read in order, it is the story of a team repeatedly paying down user frustration with resets. The 28-day pledge turns that habit into explicit policy — and into a public commitment to ship something every single day.

The math, and the October 30 problem

Nobody outside OpenAI knows the mix of improvements versus resets. The scenarios range from reset-heavy (roughly 20 resets and 8 improvements, which makes usage feel generous while the product barely moves) to improvement-heavy (roughly 4 resets and 24 real changes, where efficiency gains have to do the work). Even the generous case is not unlimited use; resets refresh what you have, they do not raise the ceiling.

The sharper issue is that the pledge window — October 5 through roughly November 1 — contains October 30, the scheduled Pro 200 cut. From that date, Pro 200 Codex and Work usage drops from 20x to 10x the Plus allowance, and GPT-6 Pro chat drops from 200 to 100 messages per week, per the subscriber email OpenAI already sent. Existing seats keep old limits through October 29. Nothing in Sottiaux’s posts reverses this. The realistic read: resets and efficiency wins will cushion the cut for a few days, not cancel it. If Codex capacity is the reason you hold a $200 seat, the number to price is the post-October 30 plan, not the current one.

How developers reacted

The reaction splits into four camps. The cautiously optimistic welcome a daily commitment over silence — “all the best” is a fair response until a week of results is in. The structural skeptics make the sharpest critique: tying quota to a daily lottery makes planning harder, not easier; the thread’s best joke was that Sottiaux “turned usage into gambling: burn your Codex usage every day and pray it’s a reset day.” The churn camp claims to have already moved subscriptions to Claude Max or shifted API spend to Anthropic — individual statements, not data, though similar threads appear every time either vendor tightens limits. And the hype camp outran the facts, including a widely shared and unverified boast that an overnight Claude Opus 5.5 run decompiled the Linux kernel into full source.

One reply argued the pledge reads like a “death rattle” unless a major model ships before day 28. That is opinion, but it identifies the real test: resets buy goodwill; shipped models and genuine efficiency buy retention.

What to watch

The signal that matters is whether the first improvements are subtractive or additive. Today a Codex user must hold an extraordinary amount in their head: GPT-6 Astra versus GPT-6.1 Sol versus Luna, the Ultrafast tier, Work meters versus Codex meters, Pro 200 versus Pro 500, a first-month Dots allowance, and a pile of overlapping features. “Simplification” that removes, merges, or renames would match what users actually asked for. More modes would not.

For day-to-day users, the practical advice from the thread is simple: stop hoarding quota, since resets have typically replaced the remaining balance; front-load big refactors early in your window; keep a short daily log so that by day 28 you know whether the pledge changed your actual capacity or just your mood. And re-price your seat before October 30 — because when the scoreboard goes dark, the underlying plan formula is what you are left with.