We use analytics and advertising tools by default. You can update this anytime.
Manage optional tracking categories. Necessary cookies stay on so the site can function.
All about the newest tools from Anthropic
Faster than GPT-5 Codex, smarter and more steerable than Opus 4.1
After 24 hours of hands-on testing, we found a model that’s fast, reliable, and surprisingly funny—but still prone to overreaching and not yet a writing champ
It feels less like learning something new than a browser that has caught up to how we already want to work with AI
Five tests across blind comparisons, editorial standards, and deadlines—here's what changed our setup
Switching browsers is a pain. Here are the ones that our team deemed worth it.
It launches today—here’s our day-zero vibe check
The smartest model isn’t always the most useful one
Why Google might quietly win the race to be AI’s top backend provider
This one’s for the developers
OpenAI's latest model update excels at instruction-following and extended tasks, but don't expect it to surprise you
The asynchronous, agentic workflow developers love is finally accessible to everyone—but the polish isn't there yet
OpenAI nailed the interface. But it's built for hardcore engineering.
Anthropic's latest Opus is more precise, more literal, and the best coding model we've tested on well-specified tasks—but it won't fill in the gaps for you anymore
Sora 2 removed every creative barrier, but our feeds tell a different story about human imagination
Sonnet 4.6 delivers Opus-close performance at half the price—but speed didn't come along for the ride
Opus 4.8 tops both our Senior Engineer benchmark and our writing tests. It’s the most complete model we’ve tested. We just wish it had an app to match.
Anthropic’s new model can draft, code, and analyze competently, but every use case has a cheaper, faster, or smarter alternative
GPT-5.4 is fast, opinionated, and good enough to tempt our Opus loyalist
We’ve tested both models thoroughly—here’s our head-to-head Vibe Check