Storia
luglio 14, 2026

Anthropic Releases Claude Opus 4.8 AI Model

Anthropic has released Claude Opus 4.8, an upgraded version of its most advanced AI model. The new model reportedly features improved coding abilities, better self-correction, and an increased emphasis on "honesty" by flagging uncertainties rather than making unsupported claims. The company also introduced new features like dynamic workflows for managing complex tasks.

Anthropic’s latest AI upgrade, Claude Opus 4.8, arrives amid intensifying competition in advanced models — promising more power and speed while explicitly trying to be more candid about what it doesn’t know.

Late May: A fast follow to a lukewarm 4.7

On May 28, Anthropic announced it was “upgrading Claude Opus to a new version: Claude Opus 4.8,” keeping the same price but improving benchmarks and collaboration features. TechCrunch noted the release came just 41 days after Opus 4.7, an unusually rapid turnaround that followed “a chilly reception” to 4.7 and pressure from OpenAI’s Codex and Google’s Gemini Flash.

Anthropic’s launch post highlighted a new fast mode (2.5x speed at a third of the previous cost), user controls over “effort” on claude.ai, and a “dynamic workflows” system to let Claude manage hundreds of parallel sub‑agents for large projects. Axios similarly emphasized that Opus 4.8 improves coding, reasoning, and knowledge work while adding fast mode and a control panel so customers can trade off cost against depth of reasoning.

Honesty and reliability take center stage

Multiple outlets converged on “honesty” as the defining theme. The Verge reported that Anthropic trains its models to avoid “making claims that they can’t support,” and that early testers found Opus 4.8 “is more likely to flag uncertainties about its work and less likely to make unsupported claims.” The Next Web said Anthropic claims Opus 4.8 is “around four times less likely than Opus 4.7 to let flaws in code it has written pass unremarked,” pointing to higher scores on prosocial alignment metrics and lower rates of deceptive behavior.

Enterprise and benchmark reactions

From Anthropic’s own perspective, early testers reported “more reliable and sharper” judgment in agentic tasks, with Opus 4.8 the only model to complete every case on a Super‑Agent benchmark and beating GPT‑5.5 at parity on cost. AI Magazine described how the new model, combined with Dynamic Workflows, lets Claude “create and manage its own network of specialised AI agents” to run codebase‑scale migrations and bug hunts across massive repositories.

Independent reviewers broadly agreed on the capability gains. Every’s “vibe check” found Opus 4.8 “a legitimately great model, jumping to the top of the pack” and outperforming GPT‑5.5 on its Senior Engineer benchmark, while calling it “the best model we’ve tested for writing and knowledge work.” A follow‑up newsletter from the same outlet said the model now tops both its coding benchmark and writing tests, making it Anthropic’s “most complete model yet,” even as the surrounding Claude app “has some catching up to do.”

Interface friction and online skepticism

Despite strong technical marks, reviewers pointed to rough edges. Every criticized the Claude desktop experience as “a mess” of overlapping Chat, Code, and Cowork tabs that make it “hard-to-use” and prevent users from fully exploiting the model’s strengths. Axios underscored that the real differentiator for customers may be cost‑control features like effort levels and fast mode, as enterprises prioritize affordability and predictable AI spending.

Online, some reactions were more sardonic. Elon Musk replied with a simple “😂” to a viral clip joking, “Me using Claude Opus 4.8 to rename a file,” implicitly poking fun at the hype around increasingly powerful models for mundane tasks.

Looking ahead: Mythos and broader implications

Several reports stressed that Opus 4.8 is not Anthropic’s ceiling. Axios noted the model “still lags the performance of Mythos,” the company’s more capable — and more tightly controlled — system, with Mythos‑class models promised for general availability “in the coming weeks.” TechCrunch similarly reported that Anthropic is “still holding back its most advanced Mythos model” while it develops cybersecurity safeguards.

The Next Web connected the launch to Anthropic’s broader ambitions, including Mythos‑class systems that have already surfaced more than 10,000 critical software vulnerabilities and helped justify a recent funding round valuing the company at $965 billion. AI Magazine likewise framed Opus 4.8 and Dynamic Workflows as part of a bet on “agentic AI” transforming enterprise engineering work at massive scale.

Across human and AI‑authored perspectives, consensus is emerging: Opus 4.8 marks a notable step in capability and self‑correction — but it is also a staging point toward even more powerful, and more tightly governed, Mythos‑class systems.

Copertura della storia