Anthropic legacy frontier model

Claude Opus 5

Anthropic’s first Opus 5 frontier model for complex agentic coding and enterprise work, with a 1M-token context window, 128K output, thinking on by default, and a full effort ladder up to max.

Context window
1M
tokens
Max output
128K
tokens
Input
$5.00
per 1M tokens
Cache read
$0.50
per 1M tokens
Output
$25.00
per 1M tokens

01 / Overview

What Claude Opus 5 Is

Claude Opus 5 is Anthropic’s original Opus 5 frontier model for complex agentic coding and enterprise work, released as a step-change upgrade over Claude Opus 4.8.

A New Opus Generation Built Around Sustained Work

Anthropic described Opus 5 as a step-change rather than an incremental update. The largest gains over Opus 4.8 were in deep reasoning, agentic and long-horizon tasks, and the ability to convert additional test-time compute into better results.

That positioning matters because Opus 5 was not designed only for single-turn question answering. It was built for workflows where the model must remain useful across a sequence of decisions: inspect a codebase, delegate to subagents, call tools, verify changes, work through long documents, or maintain a complex enterprise task over an extended session.

The model has a 1,000,000-token context window and supports up to 128,000 output tokens. It accepts text and images and generates text. Thinking is on by default, with high as the default effort level.

Opus 5 has since been superseded by Opus 5.5 as Anthropic’s recommended general frontier model. That makes Opus 5 especially useful today as an operational baseline for teams with existing prompts, evaluations, tool loops, and production behavior built around the first Opus 5 release.

  • Model ID: claude-opus-5.
  • Released July 24, 2026.
  • 1M-token context window and 128K standard output.
  • Text and image input with text output.
  • Thinking on by default with configurable effort.
  • Legacy model with Opus 5.5 as the recommended successor.
Model profile
Provider
Anthropic
Family
Claude Opus 5
Model ID
claude-opus-5
Knowledge cutoff
May 2026
Input modalities
Text, Image
Output modality
Text
Lifecycle
Legacy

02 / Agentic coding

Claude Opus 5 for Agentic Coding

Agentic coding was one of the largest capability jumps in Opus 5, with stronger long-horizon execution, deeper repository work, more reliable verification, and more active delegation to subagents.

Measure Whether the Agent Actually Finishes the Job

Anthropic highlights Opus 5’s ability to stay on task across extended tool-use loops, complete multi-file features, execute larger refactors, and finish end-to-end feature work without leaving placeholders or unfinished steps.

That makes ordinary single-prompt coding benchmarks incomplete. A production coding agent succeeds only when the entire trajectory works: repository inspection, planning, file edits, tests, debugging, tool calls, verification, and final delivery.

Opus 5 also became more willing to delegate in multi-agent systems. Anthropic notes stronger writer-verifier patterns and more effective coordination between subagents. At the same time, the model tends to verify its own work without being explicitly instructed, so prompts inherited from older models can accidentally cause redundant verification.

For EidoStack-style evaluation, the useful metrics are completion rate, number of tool turns, human interventions, test pass rate, false-positive bug reports, output tokens, elapsed time, and total cost per accepted change.

Agentic coding workflows
  1. 01

    Multi-file implementation

    Carry a feature across repository structure, dependencies, source files, tests, and validation rather than generating one isolated code snippet.

  2. 02

    Long refactors

    Maintain the architectural goal through multiple edits and tool calls while reducing unfinished stubs and partial implementations.

  3. 03

    Code review and bug finding

    Surface actionable defects while keeping false positives low enough for practical engineering review.

  4. 04

    Multi-agent coordination

    Delegate subtasks to parallel agents and combine their work with explicit verification and reduced overlap.

03 / Enterprise work

Enterprise Analysis, Vision, and Document Work

Opus 5 extends its long-horizon behavior beyond coding into complex enterprise workflows involving documents, spreadsheets, slides, charts, diagrams, and visual evidence.

Frontier Capability Is Most Valuable When Several Skills Interact

Enterprise workflows rarely ask for only one capability. A financial analysis may require reading long documents, interpreting tables, checking calculations, maintaining a set of constraints, and producing a polished artifact. A frontend reconstruction task may require visual inspection, code generation, iteration, and comparison against the source image.

Anthropic specifically highlights Opus 5 improvements in vision, office and document tasks, and long-context instruction following. The model can analyze charts, documents, diagrams, and UI visuals and can work across complex multi-sheet spreadsheets and structured slide decks.

The practical evaluation question is not whether the model can describe the input. It is whether it produces a result that survives downstream verification: formulas calculate correctly, slides preserve the intended argument, visual recreations match the reference, and long documents remain internally consistent.

Complex professional workloads
  1. 01

    Document-heavy analysis

    Reason across large reports, policies, specifications, contracts, and collections of supporting material.

  2. 02

    Spreadsheet work

    Create or edit multi-sheet workbooks with formulas, assumptions, structured data, and analytical logic.

  3. 03

    Presentation workflows

    Turn source material into structured slide narratives while preserving evidence, hierarchy, and consistency.

  4. 04

    Visual engineering

    Interpret screenshots, charts, diagrams, and UI references as part of coding or knowledge-work tasks.

04 / Pricing

Claude Opus 5 Pricing

Claude Opus 5 costs $5 per million input tokens and $25 per million output tokens, with prompt caching and batch processing available to reduce repeated or offline workload cost.

Price the Completed Workflow, Not One Turn

Opus 5 kept the same base pricing as Opus 4.8 while materially increasing capability. Anthropic positioned this as frontier intelligence at half the base token price of Claude Fable 5.

Prompt caching can be especially important for agentic workloads because large stable prefixes—system instructions, repository context, policies, reference documents, or recurring task definitions—may be reused across many turns.

Anthropic lists 5-minute cache writes at $6.25 per million tokens, 1-hour cache writes at $10 per million tokens, and cache reads at $0.50 per million tokens. Batch API requests receive a 50% discount on standard input and output pricing.

The successor Opus 5.5 lowered the base rates further to $4 input and $20 output per million tokens, so cost itself is now one reason to evaluate migration.

  • Standard input: $5.00 per 1M tokens.
  • Output: $25.00 per 1M tokens.
  • 5-minute cache write: $6.25 per 1M tokens.
  • 1-hour cache write: $10.00 per 1M tokens.
  • Cache read: $0.50 per 1M tokens.
  • Batch API: 50% discount on standard input and output.
Token and cache pricing

1M tokens · USD

Input
$5.00
5m cache write
$6.25
1h cache write
$10.00
Cache read
$0.50
Output
$25.00

Example: 20K input + 4K output

Input cost
$0.1000
Output cost
$0.1000
Estimated uncached total
$0.2000

05 / Context

A 1M Context Window With 128K Standard Output

Claude Opus 5 provides a 1,000,000-token context window as both its default and maximum context size, with up to 128,000 standard output tokens.

Large Context Supports Repository and Enterprise State

A million-token window gives an agent room for much more than a long prompt. It can hold source code, documentation, conversation history, image inputs, tool results, plans, research evidence, enterprise policies, and intermediate work in one active context.

Anthropic emphasizes consistent instruction following, tool calling, and reasoning throughout the long context. That matters because raw capacity is useful only when the model can still identify relevant instructions and evidence deep in the window.

For batch workloads, Anthropic also exposes a beta maximum output of 300,000 tokens using the Message Batches API. Standard interactive output remains 128,000 tokens.

Large context should still be managed deliberately. Prompt caching, retrieval, compaction, and context editing can reduce cost and noise even when the model technically has enough room for everything.

Context capacity

Context window

1,000,000

Standard max output

128,000

Working contextOutput limit

Message Batches can use up to 300K output tokens in beta; the standard interactive maximum output is 128K tokens.

06 / Thinking & effort

Thinking Is On by Default in Claude Opus 5

Opus 5 made adaptive thinking the default behavior and introduced a broader practical role for the effort parameter, with levels from low through max.

Effort Became a Primary Model-Selection Control

On Opus 4.8, requests ran without thinking unless adaptive thinking was enabled explicitly. On Opus 5, requests with no thinking field run with thinking on by default. The model decides when and how much to reason on each turn.

Anthropic says Opus 5 converts additional effort into better results more reliably than earlier Opus generations. The available ladder is low, medium, high, xhigh, and max, with high as the default.

This gives teams a useful cost-quality control inside one model. Routine tasks can be tested at lower effort, while architecture, complex debugging, difficult analysis, or capability-critical agent steps can be escalated to xhigh or max when evaluations show a measurable gain.

Opus 5 has a distinctive compatibility rule that separates it from both Opus 4.8 and Opus 5.5: thinking can still be disabled, but only at effort high or below. Sending disabled thinking with xhigh or max returns a 400 error. Opus 5.5 removes the option to disable thinking entirely.

Effort ladder
lowmediumhigh · defaultxhighmax

Cost-sensitive step

Test low or medium effort where the task is routine and quality remains above your acceptance threshold.

Capability-critical step

Use xhigh or max only where evaluations show that deeper reasoning materially improves completion quality.

07 / Tools & fast mode

Tool-Oriented Workflows and Fast Mode

Opus 5 expanded the integration surface for long-running agents with mid-conversation tool changes and a research-preview fast mode alongside its core tool-use capabilities.

Change Tools Without Rebuilding the Whole Session

Anthropic introduced mid-conversation tool changes in beta for Opus 5. This allows an application to add or remove tools between conversation turns while preserving the prompt cache rather than treating the tool list as fixed for the entire session.

That is useful for agents whose permissions and available actions evolve during a task. A coding agent might gain a deployment tool only after tests pass. A research workflow might add a specialized data source only when the investigation reaches a specific stage.

Opus 5 also supports a research-preview fast mode on the Claude API. It provides lower-latency inference at separate premium pricing, creating another evaluation dimension: standard Opus 5 quality and economics versus the same model family in a latency-optimized mode.

Other relevant capabilities include text and image input, tool use, prompt caching, large-context processing, code-oriented workflows, and batch generation with the higher-output beta path.

Integration features
  • Text and image input

    Combine written instructions with screenshots, charts, diagrams, documents, and other supported images.

    Supported
  • Adaptive thinking

    Thinking is on by default and controlled primarily through the effort parameter.

    Supported
  • Tool use

    Use application-defined tools inside multi-step agent workflows.

    Supported
  • Mid-conversation tool changes

    Beta support lets applications add or remove tools between turns while preserving prompt-cache continuity.

    Supported
  • Prompt caching

    Reduce repeated processing cost for stable context used across long sessions.

    Supported
  • Fast mode

    Research-preview lower-latency mode on the Claude API with separate premium pricing.

    Supported
  • 300K batch output

    Message Batches can request up to 300K output tokens through the documented beta path.

    Supported

08 / Opus 5 → 5.5

Migrating From Claude Opus 5 to Opus 5.5

Claude Opus 5.5 lowers base pricing, changes default effort, makes adaptive thinking mandatory, and introduces several API compatibility changes that should be tested before production migration.

The Successor Is Cheaper but Not a Drop-In Model-ID Swap

Opus 5.5 costs $4 per million input tokens and $20 per million output tokens, compared with $5 and $25 on Opus 5. Anthropic also reports more than 30% faster output and stronger long-running coding and knowledge-work behavior in the successor.

However, the migration changes API semantics. Opus 5 defaults to high effort; Opus 5.5 defaults to medium. Opus 5 allows thinking to be disabled at high effort or below; Opus 5.5 does not allow thinking to be disabled at all.

Forced tool choice using any or a named tool is also rejected by Opus 5.5, and thinking blocks have new conversation/model compatibility rules. Computer-use integrations on some platforms need the newer toolset.

Therefore, migration evaluation should include more than prompt quality. Replay tool loops, persisted conversations, thinking-block handling, effort settings, UI progress streaming, computer use, latency, cost, and code-review behavior.

Migration differences
  1. 01

    Recalibrate effort

    Opus 5 defaults to high; Opus 5.5 defaults to medium. Run a new effort sweep instead of carrying the old setting forward.

  2. 02

    Remove disabled thinking

    Opus 5 permits disabled thinking at high or below; Opus 5.5 requires adaptive thinking on every request.

  3. 03

    Retest tool orchestration

    Forced tool choice is not supported on Opus 5.5, and computer-use/tool-loop behavior has migration requirements.

  4. 04

    Measure the economics again

    Opus 5.5 lowers input/output pricing and can finish many agentic tasks in fewer steps and tokens.

09 / Evaluation

Claude Opus 5 Strengths and Limitations

Claude Opus 5 remains a valuable historical and production baseline for complex agentic coding and enterprise workflows, but Opus 5.5 now offers a stronger current path with lower base cost and updated API semantics.

Strengths

  • Strong agentic coding baseline

    Opus 5 was a major jump in repository-scale implementation, refactoring, code review, long tool loops, and multi-agent coordination.

  • Full effort ladder

    Low through max effort levels allow teams to test cost-quality tradeoffs within the same model.

  • 1M context and 128K output

    Large working context and long output capacity suit repository, document, research, and artifact-heavy workflows.

  • Flexible thinking behavior

    Unlike Opus 5.5, Opus 5 can still disable thinking at effort high or below when a legacy integration requires it.

What to consider

  • Superseded by Opus 5.5

    Anthropic recommends migrating to the newer model for improved performance and current frontier behavior.

  • Higher token pricing

    $5 input and $25 output per million tokens are 25% higher than Opus 5.5’s $4/$20 base rates.

  • Migration semantics differ

    Thinking, effort defaults, forced tool choice, thinking blocks, and computer-use integrations change on Opus 5.5.

  • Moderate latency

    The model is designed for frontier-quality work rather than the fastest high-volume request path.

Preserve the baseline before upgrading

Evaluate Claude Opus 5 against its successor on your real workload

Compare repository-scale coding, code review, vision, enterprise documents, long-context behavior, effort settings, cache usage, latency, and total task cost before migrating to Opus 5.5.

Start Free

Claude Opus 5 remains useful as a production baseline for existing integrations, but Anthropic now recommends Claude Opus 5.5 for improved performance and lower token pricing.

Common Questions

What is Claude Opus 5?

Claude Opus 5 is Anthropic’s first Opus 5 frontier model for complex agentic coding and enterprise work. It was released on July 24, 2026 as a step-change upgrade over Opus 4.8.

How much does Claude Opus 5 cost?

Claude Opus 5 costs $5 per 1M input tokens and $25 per 1M output tokens. Five-minute cache writes cost $6.25/MTok, one-hour cache writes cost $10/MTok, and cache reads cost $0.50/MTok.

What is the Claude Opus 5 context window?

Claude Opus 5 has a 1,000,000-token context window as both the default and maximum context size. Standard maximum output is 128,000 tokens.

Can Claude Opus 5 output more than 128K tokens?

Standard requests support up to 128K output tokens. Anthropic also documents a beta path for up to 300K output tokens through the Message Batches API.

What is Claude Opus 5’s knowledge cutoff?

Anthropic lists May 2026 as both the reliable knowledge cutoff and training-data cutoff for Claude Opus 5.

Does Claude Opus 5 support images?

Yes. Claude Opus 5 accepts text and image input and produces text output. Anthropic highlights improved performance on charts, documents, diagrams, and visual UI work.

Does Claude Opus 5 use adaptive thinking?

Yes. Thinking is on by default. The model supports low, medium, high, xhigh, and max effort levels, with high as the default.

Can thinking be disabled on Claude Opus 5?

Yes, but only at effort high or below. Sending thinking disabled with xhigh or max effort returns a 400 error. This differs from Opus 5.5, where thinking cannot be disabled at any effort level.

What is fast mode on Claude Opus 5?

Anthropic offers a research-preview fast mode for Claude Opus 5 on the Claude API. It provides lower-latency inference at separate premium pricing.

How is Claude Opus 5 different from Claude Opus 5.5?

Opus 5.5 is the newer successor. It lowers base pricing from $5/$25 to $4/$20 per million input/output tokens, defaults to medium rather than high effort, keeps adaptive thinking always on, and changes several tool-use and conversation semantics. Anthropic also reports faster output and stronger long-running agentic coding and knowledge work.

Is Claude Opus 5 still available?

Yes. Claude Opus 5 remains available as a legacy model. Anthropic lists its original release date as July 24, 2026 and recommends considering Claude Opus 5.5 for improved performance.

Should I use Claude Opus 5 for a new application?

For a new evaluation, start with the current Opus 5.5 model unless you specifically need Opus 5 compatibility. Opus 5 remains valuable for existing integrations, regression testing, historical comparison, and migration baselines.

Model information

Last updated

Specifications, pricing, capabilities, effort behavior, prompt caching, fast mode, release information, and migration guidance on this page are based on Anthropic’s official Claude Opus 5 model documentation, What’s New guide, and Opus 5.5 migration documentation.

Claude Opus 5 — Pricing, 1M Context, Agentic Coding & Effort | EidoStack