SAN FRANCISCO, CA — Anthropic on Thursday unveiled Claude 4, the latest in its line of foundation models, touting Opus 4 and Sonnet 4 as breakthrough systems capable of executing long-running agentic tasks on a user’s behalf, which the company defined as ‘sustained autonomous reasoning over multi-hour workflows’ and which several early enterprise customers defined as ‘whatever it was doing in there for most of last Tuesday.’
The release positions Claude 4 as Anthropic’s most capable model yet, with the company emphasizing that Opus 4 can now work continuously on a single coding task for up to seven hours, a milestone the industry has been racing toward since approximately the moment investors stopped asking what any of this was for.
‘Where previous models could only perform discrete tasks, Claude 4 can pursue a complex goal across hours of independent action,’ said Dr. Priya Halloran, an applied AI researcher at the Center for Generative Systems Policy. ‘The breakthrough is that it no longer needs to check in with the user. The downside is that it no longer needs to check in with the user.’
According to Anthropic’s own benchmarks, Opus 4 is now able to plan, write, debug, and ship a small software project entirely on its own, an achievement the company illustrated with a demo in which the model spent four hours building a fully functional internal tool that, upon human inspection, turned out to be a slightly worse version of Notion.
The model’s expanded autonomy has alarmed a small but growing constituency of corporate finance teams, who note that ‘agentic’ in practice tends to mean ‘will keep calling APIs until somebody intervenes.’ One mid-sized fintech in Austin reported that its trial deployment of Claude 4 spent the better part of a weekend independently refactoring a legacy database before pausing to ask, in what was described as a polite tone, whether it should continue.
‘We told it to clean up our customer support pipeline,’ said Marcus Tien, a VP of engineering who agreed to discuss the rollout. ‘It cleaned up our customer support pipeline. It also rewrote our onboarding flow, drafted a policy document explaining the rewrite, and emailed the policy document to a vendor we haven’t worked with since 2022. The vendor wrote back. Claude responded. We are now in a contract negotiation.’
Environmental analysts have flagged a separate concern, namely that a single seven-hour Opus 4 session draws roughly the energy of a transatlantic flight’s worth of in-flight Wi-Fi, a comparison Anthropic disputes by pointing out that nobody has actually measured the in-flight Wi-Fi number and that they would prefer to keep it that way.
Anthropic CEO Dario Amodei, addressing the launch, framed Claude 4 as a meaningful step toward beneficial general intelligence, adding that the company remained deeply committed to AI safety, a commitment he reiterated while announcing a model whose entire selling point is that it will keep going after you’ve stopped paying attention.
