Claude Opus 5 release announcement
Official release page with benchmark framing, effort settings, availability, and Anthropic’s performance and cost claims.
Compute College
Read Claude Opus 5 benchmark claims through AI compute economics: task quality, effort settings, token pricing, fast mode, and inference demand.
Claude Opus 5 is Anthropic’s July 24, 2026 Opus release. Anthropic says it improves on Opus 4.8 across coding, knowledge-work, science, and agent evaluations while keeping standard API pricing at $5 per million input tokens and $25 per million output tokens. For compute buyers, the useful question is whether higher task completion changes cost per accepted result, token use, latency, or the amount of premium inference capacity a workload needs.
Memory trick: The price tag is per token; the economic result is per accepted task.
A model can improve the economics of a workload without becoming cheaper per token. If Opus 5 completes more complex tasks with fewer retries, tool calls, or review cycles, its cost per accepted outcome can fall at the same listed price. But if better capability prompts teams to run longer agents, larger contexts, or more autonomous workflows, total inference demand can rise even when the unit rate is unchanged.
At Anthropic’s listed standard rate, a request using 100,000 input tokens and 20,000 output tokens costs $0.50 for input plus $0.50 for output, or $1.00 before caching and platform differences. Fast mode is listed at twice the standard price, so the same token mix would cost $2.00 while aiming for lower latency. The cheaper choice depends on whether faster completion avoids enough waiting, retries, or downstream work.
Example figures are illustrative calculations, not current quoted market prices.
Current example
Anthropic announced Claude Opus 5 on July 24, 2026. Its release page describes higher performance than Opus 4.8 across selected evaluations and says Opus 5 keeps the same standard token price. Anthropic also publishes fast mode at twice the standard price. Those are first-party release and pricing claims, not independent Compute College benchmarks.
Official release page with benchmark framing, effort settings, availability, and Anthropic’s performance and cost claims.
Official pricing reference for standard usage, caching, and fast-mode rates.
Compare the prior release without treating it as the current model example.
Source discipline: evaluate Anthropic benchmark, customer, and safety claims as first-party evidence. Measure your own workload before making a routing, procurement, or capacity decision. Last checked: Sep 16, 2026.
A better benchmark score does not prove a lower production bill. A model may need more effort, longer contexts, different tools, or a different agent harness, and each can change latency, token use, and cost per completed task.
Practical takeaway
Run a recorded comparison on representative production tasks. For each candidate model and effort setting, capture input and output tokens, cache behavior, tool calls, retries, latency, human review, and whether the task was accepted. Use the official price page for the cost calculation, then compare accepted outcomes per dollar and within the required latency budget.
Decision check: does Opus 5 improve the completed outcome your team needs enough to offset its token use, latency, and operating requirements under the chosen effort and speed mode?
Compute College
Follow model releases as AI compute market signals in the ComputeTape Market Brief.
Compute College track
Step 15 of 25: Claude opus 5 benchmark explained