2 min read

Anthropic's Opus 5 Launches and Cuts $$$

Anthropic's Opus 5 Launches and Cuts $$$

Anthropic released Claude Opus 5 on July 24, and the headline number is the one that matters to anyone running a P&L: near-frontier intelligence at roughly half the cost of its top-tier sibling, Claude Fable 5. Same $5 input / $25 output pricing as the model it replaces, but a meaningfully different ceiling on what that money buys.

Key Points

  • Opus 5 tops Artificial Analysis's AA-Briefcase agentic knowledge-work benchmark with a 1,720 Elo score, at $17.79 per task, a 20% cost reduction from Fable 5's $22.30
  • It scored 30.2% on ARC-AGI-3, a benchmark built specifically to resist memorization, nearly quadrupling the prior record of 7.8% set by GPT-5.6 Sol
  • Paired with Anthropic's Auto Mode, browser-agent prompt injection attacks dropped to a 0% success rate across 129 test scenarios, down from 3.7% without those safeguards
  • Anthropic and Andon Labs used Opus 5 to advance Drone-Bench, a joint research effort testing whether AI models can autonomously fly a drone to locate and follow a person indoors
  • The model is now the default on Claude Max and the strongest available on Claude Pro, with a 1-million-token context window and a Fast Mode running at 2.5x speed

The Cost Curve Just Bent in the Buyer's Favor

For two years, "frontier intelligence" and "affordable at scale" have lived in different rooms. Opus 5 is Anthropic's attempt to put them in the same one. According to Artificial Analysis, which independently benchmarked the model ahead of release, Opus 5 at max effort scores 1,720 Elo on AA-Briefcase, its proprietary test for agentic knowledge work, 146 points ahead of Fable 5, while costing 20% less per task. That is not a marginal efficiency gain. That is a lab deciding the ceiling and the price tag no longer have to move together.

A Reasoning Leap That Resists Cheating

The more interesting number, for anyone tired of benchmark theater, is ARC-AGI-3. It is designed so models cannot lean on anything memorized during training, which makes most scores on it stubbornly low. Opus 5 scored 30.2%, nearly quadrupling the previous record of 7.8%. ARC Prize's administrators noted the model independently formulated an algebraic reflection equation mid-task, something no frontier model had done before. That is not pattern-matching. That is a model working out a rule it was never shown.

The Security Fix Nobody Asked For and Everyone Needed

Marketers running AI agents against live customer data have quietly been living with a known risk: indirect prompt injection, where a poisoned webpage or document hijacks an agent mid-task. Opus 5 combined with Auto Mode brought browser-agent attack success down to zero across 129 tested scenarios, from 3.7% without those layered defenses. That is the difference between an agent you supervise and an agent you trust.

Reasoning That Extends Past the Screen

Anthropic's continued work with Andon Labs on Drone-Bench, testing whether models can autonomously fly a drone to locate and follow a person, is a signal worth sitting with. The gap between "writes good ad copy" and "operates physical systems" is closing faster than most growth teams have planned for. That is exactly the kind of shift a solid growth strategy needs to account for now, not next year.

For marketing and growth leaders, the practical takeaway is simpler than the benchmarks suggest: the intelligence you were rationing because of cost just got cheaper, and the agents you were nervous about handing real access to just got safer. Neither excuse holds much longer.

If you're trying to figure out where Opus 5 actually changes your stack, our AI marketing services team can walk through it with you.

Anthropic's Persona Vectors Breakthrough

Anthropic's Persona Vectors Breakthrough

Remember when Microsoft's Bing chatbot went rogue and started calling itself "Sydney," declaring love for users and threatening blackmail? Or when...

Read More
Claude Gets Memory—And Anthropic Just Leapfrogged OpenAI on Transparency

Claude Gets Memory—And Anthropic Just Leapfrogged OpenAI on Transparency

Anthropic announced this week that Claude is getting persistent memory across conversations. Starting today for Max subscribers (rolling out to Pro...

Read More
Anthropic Built an AI to Interview 1,250 People About AI: Here's What We Learned

Anthropic Built an AI to Interview 1,250 People About AI: Here's What We Learned

Anthropic just released research from a tool called Anthropic Interviewer—an AI system that conducted 1,250 interviews with professionals about how...

Read More