GPT-5.3 Codex: The 10x Engineer,
Now More Fun at Parties
The autonomy we wanted is here—but the model still does what you say, not what you mean
Dan Shipper
Katie ParrottOur verdict: The dominant criticism of coding tool Codex has always been the same: It acts like a brilliant senior engineer who is methodical to a fault. It'll ship a full product autonomously that builds without errors—if you have a detailed spec. But it's slow and cautious, sometimes gets stuck in myopic loops, and has very little empathy.
GPT-5.3 Codex, which OpenAI has released today, maintains the coding prowess of its predecessors, but it's a much more user-friendly model. It's fast, a bit warmer, and more creative. It's also way more industrious—it does things without asking for permission. For developers who were frustrated by earlier Codex versions stopping to double-check obvious decisions, this is the update you've been waiting for.
In a lot of ways, it feels like Codex got upgraded with some of Opus 4.5's better qualities. Paired with the new Codex app that OpenAI launched on Monday, it's clear that OpenAI wants to make Codex a more general-purpose model for knowledge work beyond coding, and 5.3 is a step in the right direction towards that goal.
What OpenAI told us
Best-of-both-worlds model
It has the frontier coding chops of the company's latest research combined with GPT-5.2's reliability for agentic work. It's built for long-horizon tasks—the kind of sustained, multi-step work that unfolds over minutes or hours rather than a single prompt-and-response.
Less of a black box
The model narrates what it's doing as it works, making agents feel more transparent and predictable.
Mid-turn redirection
You can course-correct while the model is working instead of waiting for it to finish.
We focused our testing on coding for this Vibe Check. OpenAI also claims the model unlocks "advanced writing," but due to timing and testing constraints, we didn't evaluate that on this go-around. For now, Claude is still our preferred model for writing. If that changes, we'll let you know.
The Reach Test
"I'm entering my Codex era. Prior to this model, I would only use Codex a bit for really hard tasks or code reviews. Now, it's becoming a daily driver for my non-vibe coding tasks in bigger code bases. I especially like using it in the Codex app. GUIs are back!"

"GPT-5.3 Codex is my go-to model. I've been using it for the past two weeks with the Codex app. Even up against Opus 4.6, I'm still reaching for Codex. I gave a big redesign task to both Claude Opus 4.6 and Codex. Codex did it well, with no build errors, but Claude couldn't complete the task. It had a few build failures. Little things like that give me more trust to use GPT-5.3 Codex over Opus 4.6."

"The -codex lineup was always powerful. It always went deep into source code, third-party plugins, etc. to find solutions. I can't say I noticed any standout difference from previous Codex models. Speed is up, one-shot reliability is consistent, deep investigation is intact. It's a solid upgrade, not a revelation."

"The new Codex model surprised me because it's so fast. It feels more useful and friendly than the ones before, where it felt a little bit too like an old-school, grumpy engineer who specialized on backend, less creative projects. This model is a little bit more creative, and that's really good. It still has its power—it just keeps going and does the work well. My daily driver will still be Claude, but Codex has a place in my workflow now. I use it for research, reviews, and long feature builds, and Claude for planning."

The headline findings
Subscribers only
Only available for paid subscribers
Get full access to the verdicts, comparisons, and detailed analysis.
Subscribe to unlock →Get all of our AI ideas, apps, and training
Every is the only subscription you need to stay at the edge of AI—trusted by 100,000 builders.
Expert led courses and camps
Four productivity apps
A community learning together