
by Katie Parrott in Vibe Check Midjourney/Every illustration. Was this newsletter forwarded to you? Sign up to get it in your inbox. Anthropic surprised us yesterday by dropping Opus 4.7—so we did what we do: We went live on X and YouTube with five testers and figured it out in front of 10,000 people. Anthropic researcher Alex Albert even joined the stream to explain what had changed. Two hours of live testing and an afternoon in Slack later, here’s the short version: This model rewards people who write tight prompts and frustrates everyone who doesn’t. Here’s our complete Vibe Check on Opus 4.7. The highlights from five testers across coding, writing, and agentic work: Kieran Klaassen ran it on our hardest coding benchmark and called it the best model he’s ever tested—the first to nail a full e-commerce website build, including a custom product designer and dependable shopping cart performance. Dan Shipper watched it write a senior-engineer-quality diagnosis of a messy codebase, then refuse to execute the solution. Mike Taylor got consulting copy so sharp he said it might be better than his own writing and the best slide deck design he’s seen. Katie Parrott ran the model head-to-head with its predecessor on a personal essay and picked 4.6. 4.7’s draft was competent but rhythmically flat. Brandon Gell had it do his monthly P&L analysis and found 4.7 missed a data error that 4.6 caught unprompted last month. The pattern underneath all of it: Anthropic is tuning Claude’s eagerness like a dial between releases, and 4.7 is a hard dial-back from 4.6’s gap-filling intuition. Your old Opus prompts probably won’t deliver the results you’re used to, so you need to tweak them for this release, if 4.7 is what you want to use. Click here to read the full post Want the full text of all articles in RSS? Become a subscriber, or learn more.
No discussion yet. Be the first to share your thoughts!