We use analytics and advertising tools by default. You can update this anytime.
Manage optional tracking categories. Necessary cookies stay on so the site can function.
Staff Writer & AI Editorial Lead
Katie Parrott is a staff writer. She writes Working Overtime and contributes to Vibe Checks, Source Code, and Context Window.
Plus: An AI writing policy that makes writers do the thinking, Zuckerberg’s AI Future for Everyone runs on Meta, and a tool that keeps your agents from timing out
Opus 4.8 tops both our Senior Engineer benchmark and our writing tests. It’s the most complete model we’ve tested. We just wish it had an app to match.
Anthropic’s latest Opus can do impressive work—but getting it there may require dismantling the systems you already use
GPT-5.4 is fast, opinionated, and good enough to tempt our Opus loyalist
Anthropic’s new model can draft, code, and analyze competently, but every use case has a cheaper, faster, or smarter alternative
OpenAI’s new model is a top-end senior engineer—and easy to talk to
Our tests show where it replaces Fable and Opus—and where we’re keeping the alternatives
Cursor team members share their thoughts on building software with AI and why model selection beats prompting tricks
The AI-native IDE is now becoming an agent-orchestration tool. Will it work?
Use AI to make better editorial decisions, and carry what you learn into the next story
It one-shotted a problem other models missed—and brings agentic, parallel work to non-coding tasks
OpenAI’s new model impressed us with writing, operating software, and visual design. Anthropic’s Fable still has better instincts for building a product.
A power-user’s guide to turning ChatGPT into an operating system for knowledge work, including setup, workflows, and a seven-day starter plan
Plus: How to audit old agent instructions, habits you unlearn when you start working for yourself, a solution for Claude-isms, and the new-model tests we may be neglecting