Anthropic legacy frontier model
Claude Opus 5
Anthropic’s first Opus 5 frontier model for complex agentic coding and enterprise work, with a 1M-token context window, 128K output, thinking on by default, and a full effort ladder up to max.
- Context window
- 1M
- tokens
- Max output
- 128K
- tokens
- Input
- $5.00
- per 1M tokens
- Cache read
- $0.50
- per 1M tokens
- Output
- $25.00
- per 1M tokens
01 / Overview
What Claude Opus 5 Is
Claude Opus 5 is Anthropic’s original Opus 5 frontier model for complex agentic coding and enterprise work, released as a step-change upgrade over Claude Opus 4.8.
A New Opus Generation Built Around Sustained Work
Anthropic described Opus 5 as a step-change rather than an incremental update. The largest gains over Opus 4.8 were in deep reasoning, agentic and long-horizon tasks, and the ability to convert additional test-time compute into better results.
That positioning matters because Opus 5 was not designed only for single-turn question answering. It was built for workflows where the model must remain useful across a sequence of decisions: inspect a codebase, delegate to subagents, call tools, verify changes, work through long documents, or maintain a complex enterprise task over an extended session.
The model has a 1,000,000-token context window and supports up to 128,000 output tokens. It accepts text and images and generates text. Thinking is on by default, with high as the default effort level.
Opus 5 has since been superseded by Opus 5.5 as Anthropic’s recommended general frontier model. That makes Opus 5 especially useful today as an operational baseline for teams with existing prompts, evaluations, tool loops, and production behavior built around the first Opus 5 release.
- Model ID:
claude-opus-5. - Released July 24, 2026.
- 1M-token context window and 128K standard output.
- Text and image input with text output.
- Thinking on by default with configurable effort.
- Legacy model with Opus 5.5 as the recommended successor.
- Provider
- Anthropic
- Family
- Claude Opus 5
- Model ID
- claude-opus-5
- Knowledge cutoff
- May 2026
- Input modalities
- Text, Image
- Output modality
- Text
- Lifecycle
- Legacy
02 / Agentic coding
Claude Opus 5 for Agentic Coding
Agentic coding was one of the largest capability jumps in Opus 5, with stronger long-horizon execution, deeper repository work, more reliable verification, and more active delegation to subagents.
Measure Whether the Agent Actually Finishes the Job
Anthropic highlights Opus 5’s ability to stay on task across extended tool-use loops, complete multi-file features, execute larger refactors, and finish end-to-end feature work without leaving placeholders or unfinished steps.
That makes ordinary single-prompt coding benchmarks incomplete. A production coding agent succeeds only when the entire trajectory works: repository inspection, planning, file edits, tests, debugging, tool calls, verification, and final delivery.
Opus 5 also became more willing to delegate in multi-agent systems. Anthropic notes stronger writer-verifier patterns and more effective coordination between subagents. At the same time, the model tends to verify its own work without being explicitly instructed, so prompts inherited from older models can accidentally cause redundant verification.
For EidoStack-style evaluation, the useful metrics are completion rate, number of tool turns, human interventions, test pass rate, false-positive bug reports, output tokens, elapsed time, and total cost per accepted change.
- 01
Multi-file implementation
Carry a feature across repository structure, dependencies, source files, tests, and validation rather than generating one isolated code snippet.
- 02
Long refactors
Maintain the architectural goal through multiple edits and tool calls while reducing unfinished stubs and partial implementations.
- 03
Code review and bug finding
Surface actionable defects while keeping false positives low enough for practical engineering review.
- 04
Multi-agent coordination
Delegate subtasks to parallel agents and combine their work with explicit verification and reduced overlap.
03 / Enterprise work
Enterprise Analysis, Vision, and Document Work
Opus 5 extends its long-horizon behavior beyond coding into complex enterprise workflows involving documents, spreadsheets, slides, charts, diagrams, and visual evidence.
Frontier Capability Is Most Valuable When Several Skills Interact
Enterprise workflows rarely ask for only one capability. A financial analysis may require reading long documents, interpreting tables, checking calculations, maintaining a set of constraints, and producing a polished artifact. A frontend reconstruction task may require visual inspection, code generation, iteration, and comparison against the source image.
Anthropic specifically highlights Opus 5 improvements in vision, office and document tasks, and long-context instruction following. The model can analyze charts, documents, diagrams, and UI visuals and can work across complex multi-sheet spreadsheets and structured slide decks.
The practical evaluation question is not whether the model can describe the input. It is whether it produces a result that survives downstream verification: formulas calculate correctly, slides preserve the intended argument, visual recreations match the reference, and long documents remain internally consistent.
- 01
Document-heavy analysis
Reason across large reports, policies, specifications, contracts, and collections of supporting material.
- 02
Spreadsheet work
Create or edit multi-sheet workbooks with formulas, assumptions, structured data, and analytical logic.
- 03
Presentation workflows
Turn source material into structured slide narratives while preserving evidence, hierarchy, and consistency.
- 04
Visual engineering
Interpret screenshots, charts, diagrams, and UI references as part of coding or knowledge-work tasks.
04 / Pricing
Claude Opus 5 Pricing
Claude Opus 5 costs $5 per million input tokens and $25 per million output tokens, with prompt caching and batch processing available to reduce repeated or offline workload cost.
Price the Completed Workflow, Not One Turn
Opus 5 kept the same base pricing as Opus 4.8 while materially increasing capability. Anthropic positioned this as frontier intelligence at half the base token price of Claude Fable 5.
Prompt caching can be especially important for agentic workloads because large stable prefixes—system instructions, repository context, policies, reference documents, or recurring task definitions—may be reused across many turns.
Anthropic lists 5-minute cache writes at $6.25 per million tokens, 1-hour cache writes at $10 per million tokens, and cache reads at $0.50 per million tokens. Batch API requests receive a 50% discount on standard input and output pricing.
The successor Opus 5.5 lowered the base rates further to $4 input and $20 output per million tokens, so cost itself is now one reason to evaluate migration.
- Standard input: $5.00 per 1M tokens.
- Output: $25.00 per 1M tokens.
- 5-minute cache write: $6.25 per 1M tokens.
- 1-hour cache write: $10.00 per 1M tokens.
- Cache read: $0.50 per 1M tokens.
- Batch API: 50% discount on standard input and output.
1M tokens · USD
- Input
- $5.00
- 5m cache write
- $6.25
- 1h cache write
- $10.00
- Cache read
- $0.50
- Output
- $25.00
Example: 20K input + 4K output
- Input cost
- $0.1000
- Output cost
- $0.1000
- Estimated uncached total
- $0.2000
05 / Context
A 1M Context Window With 128K Standard Output
Claude Opus 5 provides a 1,000,000-token context window as both its default and maximum context size, with up to 128,000 standard output tokens.
Large Context Supports Repository and Enterprise State
A million-token window gives an agent room for much more than a long prompt. It can hold source code, documentation, conversation history, image inputs, tool results, plans, research evidence, enterprise policies, and intermediate work in one active context.
Anthropic emphasizes consistent instruction following, tool calling, and reasoning throughout the long context. That matters because raw capacity is useful only when the model can still identify relevant instructions and evidence deep in the window.
For batch workloads, Anthropic also exposes a beta maximum output of 300,000 tokens using the Message Batches API. Standard interactive output remains 128,000 tokens.
Large context should still be managed deliberately. Prompt caching, retrieval, compaction, and context editing can reduce cost and noise even when the model technically has enough room for everything.
Context window
1,000,000
Standard max output
128,000
Message Batches can use up to 300K output tokens in beta; the standard interactive maximum output is 128K tokens.
06 / Thinking & effort
Thinking Is On by Default in Claude Opus 5
Opus 5 made adaptive thinking the default behavior and introduced a broader practical role for the effort parameter, with levels from low through max.
Effort Became a Primary Model-Selection Control
On Opus 4.8, requests ran without thinking unless adaptive thinking was enabled explicitly. On Opus 5, requests with no thinking field run with thinking on by default. The model decides when and how much to reason on each turn.
Anthropic says Opus 5 converts additional effort into better results more reliably than earlier Opus generations. The available ladder is low, medium, high, xhigh, and max, with high as the default.
This gives teams a useful cost-quality control inside one model. Routine tasks can be tested at lower effort, while architecture, complex debugging, difficult analysis, or capability-critical agent steps can be escalated to xhigh or max when evaluations show a measurable gain.
Opus 5 has a distinctive compatibility rule that separates it from both Opus 4.8 and Opus 5.5: thinking can still be disabled, but only at effort high or below. Sending disabled thinking with xhigh or max returns a 400 error. Opus 5.5 removes the option to disable thinking entirely.
Cost-sensitive step
Test low or medium effort where the task is routine and quality remains above your acceptance threshold.
Capability-critical step
Use xhigh or max only where evaluations show that deeper reasoning materially improves completion quality.
07 / Tools & fast mode
Tool-Oriented Workflows and Fast Mode
Opus 5 expanded the integration surface for long-running agents with mid-conversation tool changes and a research-preview fast mode alongside its core tool-use capabilities.
Change Tools Without Rebuilding the Whole Session
Anthropic introduced mid-conversation tool changes in beta for Opus 5. This allows an application to add or remove tools between conversation turns while preserving the prompt cache rather than treating the tool list as fixed for the entire session.
That is useful for agents whose permissions and available actions evolve during a task. A coding agent might gain a deployment tool only after tests pass. A research workflow might add a specialized data source only when the investigation reaches a specific stage.
Opus 5 also supports a research-preview fast mode on the Claude API. It provides lower-latency inference at separate premium pricing, creating another evaluation dimension: standard Opus 5 quality and economics versus the same model family in a latency-optimized mode.
Other relevant capabilities include text and image input, tool use, prompt caching, large-context processing, code-oriented workflows, and batch generation with the higher-output beta path.
- Supported
Text and image input
Combine written instructions with screenshots, charts, diagrams, documents, and other supported images.
- Supported
Adaptive thinking
Thinking is on by default and controlled primarily through the effort parameter.
- Supported
Tool use
Use application-defined tools inside multi-step agent workflows.
- Supported
Mid-conversation tool changes
Beta support lets applications add or remove tools between turns while preserving prompt-cache continuity.
- Supported
Prompt caching
Reduce repeated processing cost for stable context used across long sessions.
- Supported
Fast mode
Research-preview lower-latency mode on the Claude API with separate premium pricing.
- Supported
300K batch output
Message Batches can request up to 300K output tokens through the documented beta path.
08 / Opus 5 → 5.5
Migrating From Claude Opus 5 to Opus 5.5
Claude Opus 5.5 lowers base pricing, changes default effort, makes adaptive thinking mandatory, and introduces several API compatibility changes that should be tested before production migration.
The Successor Is Cheaper but Not a Drop-In Model-ID Swap
Opus 5.5 costs $4 per million input tokens and $20 per million output tokens, compared with $5 and $25 on Opus 5. Anthropic also reports more than 30% faster output and stronger long-running coding and knowledge-work behavior in the successor.
However, the migration changes API semantics. Opus 5 defaults to high effort; Opus 5.5 defaults to medium. Opus 5 allows thinking to be disabled at high effort or below; Opus 5.5 does not allow thinking to be disabled at all.
Forced tool choice using any or a named tool is also rejected by Opus 5.5, and thinking blocks have new conversation/model compatibility rules. Computer-use integrations on some platforms need the newer toolset.
Therefore, migration evaluation should include more than prompt quality. Replay tool loops, persisted conversations, thinking-block handling, effort settings, UI progress streaming, computer use, latency, cost, and code-review behavior.
- 01
Recalibrate effort
Opus 5 defaults to high; Opus 5.5 defaults to medium. Run a new effort sweep instead of carrying the old setting forward.
- 02
Remove disabled thinking
Opus 5 permits disabled thinking at high or below; Opus 5.5 requires adaptive thinking on every request.
- 03
Retest tool orchestration
Forced tool choice is not supported on Opus 5.5, and computer-use/tool-loop behavior has migration requirements.
- 04
Measure the economics again
Opus 5.5 lowers input/output pricing and can finish many agentic tasks in fewer steps and tokens.
09 / Evaluation
Claude Opus 5 Strengths and Limitations
Claude Opus 5 remains a valuable historical and production baseline for complex agentic coding and enterprise workflows, but Opus 5.5 now offers a stronger current path with lower base cost and updated API semantics.
Strengths
Strong agentic coding baseline
Opus 5 was a major jump in repository-scale implementation, refactoring, code review, long tool loops, and multi-agent coordination.
Full effort ladder
Low through max effort levels allow teams to test cost-quality tradeoffs within the same model.
1M context and 128K output
Large working context and long output capacity suit repository, document, research, and artifact-heavy workflows.
Flexible thinking behavior
Unlike Opus 5.5, Opus 5 can still disable thinking at effort high or below when a legacy integration requires it.
What to consider
Superseded by Opus 5.5
Anthropic recommends migrating to the newer model for improved performance and current frontier behavior.
Higher token pricing
$5 input and $25 output per million tokens are 25% higher than Opus 5.5’s $4/$20 base rates.
Migration semantics differ
Thinking, effort defaults, forced tool choice, thinking blocks, and computer-use integrations change on Opus 5.5.
Moderate latency
The model is designed for frontier-quality work rather than the fastest high-volume request path.
Preserve the baseline before upgrading
Evaluate Claude Opus 5 against its successor on your real workload
Compare repository-scale coding, code review, vision, enterprise documents, long-context behavior, effort settings, cache usage, latency, and total task cost before migrating to Opus 5.5.
Start FreeClaude Opus 5 remains useful as a production baseline for existing integrations, but Anthropic now recommends Claude Opus 5.5 for improved performance and lower token pricing.
Common Questions
What is Claude Opus 5?
Claude Opus 5 is Anthropic’s first Opus 5 frontier model for complex agentic coding and enterprise work. It was released on July 24, 2026 as a step-change upgrade over Opus 4.8.
How much does Claude Opus 5 cost?
Claude Opus 5 costs $5 per 1M input tokens and $25 per 1M output tokens. Five-minute cache writes cost $6.25/MTok, one-hour cache writes cost $10/MTok, and cache reads cost $0.50/MTok.
What is the Claude Opus 5 context window?
Claude Opus 5 has a 1,000,000-token context window as both the default and maximum context size. Standard maximum output is 128,000 tokens.
Can Claude Opus 5 output more than 128K tokens?
Standard requests support up to 128K output tokens. Anthropic also documents a beta path for up to 300K output tokens through the Message Batches API.
What is Claude Opus 5’s knowledge cutoff?
Anthropic lists May 2026 as both the reliable knowledge cutoff and training-data cutoff for Claude Opus 5.
Does Claude Opus 5 support images?
Yes. Claude Opus 5 accepts text and image input and produces text output. Anthropic highlights improved performance on charts, documents, diagrams, and visual UI work.
Does Claude Opus 5 use adaptive thinking?
Yes. Thinking is on by default. The model supports low, medium, high, xhigh, and max effort levels, with high as the default.
Can thinking be disabled on Claude Opus 5?
Yes, but only at effort high or below. Sending thinking disabled with xhigh or max effort returns a 400 error. This differs from Opus 5.5, where thinking cannot be disabled at any effort level.
What is fast mode on Claude Opus 5?
Anthropic offers a research-preview fast mode for Claude Opus 5 on the Claude API. It provides lower-latency inference at separate premium pricing.
How is Claude Opus 5 different from Claude Opus 5.5?
Opus 5.5 is the newer successor. It lowers base pricing from $5/$25 to $4/$20 per million input/output tokens, defaults to medium rather than high effort, keeps adaptive thinking always on, and changes several tool-use and conversation semantics. Anthropic also reports faster output and stronger long-running agentic coding and knowledge work.
Is Claude Opus 5 still available?
Yes. Claude Opus 5 remains available as a legacy model. Anthropic lists its original release date as July 24, 2026 and recommends considering Claude Opus 5.5 for improved performance.
Should I use Claude Opus 5 for a new application?
For a new evaluation, start with the current Opus 5.5 model unless you specifically need Opus 5 compatibility. Opus 5 remains valuable for existing integrations, regression testing, historical comparison, and migration baselines.
Model information
Last updated
Specifications, pricing, capabilities, effort behavior, prompt caching, fast mode, release information, and migration guidance on this page are based on Anthropic’s official Claude Opus 5 model documentation, What’s New guide, and Opus 5.5 migration documentation.