Anthropic legacy frontier model
Claude Opus 4.5
A legacy Opus model from the manual extended-thinking era, combining a 200K-token context window, 64K output, text and image input, tool use, prompt caching, and the unusual pairing of effort controls with explicit budget_tokens.
- Context window
- 200K
- tokens
- Max output
- 64K
- tokens
- Input
- $5.00
- per 1M tokens
- Cache read
- $0.50
- per 1M tokens
- Output
- $25.00
- per 1M tokens
01 / Overview
What Claude Opus 4.5 Is
Claude Opus 4.5 is a legacy Anthropic frontier model from the generation immediately before adaptive thinking became the standard Opus reasoning interface.
A Distinctive Pre-Adaptive-Thinking Opus
Anthropic released Claude Opus 4.5 on November 24, 2025. The pinned Claude API model ID is claude-opus-4-5-20251101, while the provider also exposes the alias claude-opus-4-5.
The model provides a 200,000-token context window and supports up to 64,000 output tokens. It accepts text and images and produces text.
Its most distinctive technical characteristic is the reasoning interface. Opus 4.5 uses manual extended thinking rather than adaptive thinking. Developers explicitly enable thinking and set a budget_tokens value that controls how much reasoning capacity is available.
At the same time, Opus 4.5 supports Anthropic’s effort control. This makes it unusual in the Claude lineup: Anthropic describes it as the only model that is extended-thinking-only while also supporting effort.
Anthropic lists Opus 4.5 as Active (legacy) and recommends migrating to Claude Opus 5.5.
- Model ID:
claude-opus-4-5-20251101. - Claude API alias:
claude-opus-4-5. - Released November 24, 2025.
- 200K context window.
- 64K maximum output.
- Text and image input with text output.
- Extended thinking with manual
budget_tokens. - Active legacy status.
- Provider
- Anthropic
- Family
- Claude Opus 4.5
- Model ID
- claude-opus-4-5-20251101
- Knowledge cutoff
- May 2025
- Training cutoff
- Aug 2025
- Input modalities
- Text, Image
- Lifecycle
- Active · legacy
02 / Extended thinking
Manual Extended Thinking With `budget_tokens`
Claude Opus 4.5 belongs to the older reasoning generation where applications explicitly allocate a thinking-token budget instead of letting the model adapt reasoning depth automatically.
The Application Chooses a Reasoning Budget
Extended thinking is enabled with thinking.type: "enabled" and a budget_tokens value.
Anthropic requires the thinking budget to be at least 1,024 tokens and smaller than max_tokens in the ordinary configuration. Thinking tokens count toward the output ceiling, so a large reasoning budget leaves less room for the visible final response.
The budget is a target rather than a guarantee that every token will be consumed. Claude can stop reasoning earlier when the task does not require the entire allocation.
This architecture creates a direct deployment question that does not exist in the same form on newer adaptive-thinking models: how much of the 64K output envelope should be reserved for hidden reasoning versus user-visible output?
For simple tasks, a large budget can waste latency and output capacity. For difficult analysis, coding, or multi-step planning, too little reasoning budget can reduce answer quality.
Short task
Use no extended thinking or a smaller supported budget when deep reasoning does not improve the accepted result.
Difficult task
Allocate a larger budget when complex coding, planning, or analysis benefits from more reasoning depth.
03 / Thinking + effort
The Only Extended-Thinking-Only Claude Model With Effort
Claude Opus 4.5 uniquely combines manual budget_tokens with the higher-level effort control, giving applications two separate levers over model behavior.
Effort and Budget Solve Different Problems
Anthropic’s current documentation distinguishes the two controls.
budget_tokens determines the available reasoning depth for extended thinking. It is a concrete token budget.
effort is broader. It influences how much work Claude puts into the whole response, including reasoning, tool use, and answer depth. Anthropic lists high as the default effort for Opus 4.5.
That means the two settings should not be treated as duplicates. A workload can use the same thinking budget at different effort levels and still produce different behavior.
Opus 4.5 is especially useful as an evaluation baseline because it makes these two dimensions explicit. Measure whether higher effort or a larger thinking budget improves the metric that actually matters: task completion, code correctness, analysis quality, tool efficiency, or reviewer acceptance.
When moving to Opus 4.6 or later, this dual-control pattern changes. Adaptive thinking replaces the need to manually set a reasoning token budget.
- 01
budget_tokens
Sets the manual extended-thinking allowance and directly reserves part of the output budget for reasoning.
- 02
effort
Shapes the overall amount of work Claude performs across reasoning, tool use, and response generation.
- 03
Evaluate separately
Run controlled comparisons so you know whether more effort, more thinking budget, or both improve the accepted result.
- 04
Plan for migration
Newer Opus generations replace manual budget_tokens with adaptive thinking and a different effort contract.
04 / Coding & agents
Coding, Analysis, and Tool-Based Agents
Claude Opus 4.5 was designed for high-capability coding and analysis and supports tool-oriented workflows, but its reasoning and orchestration contract differs from current Opus models.
Agent Evaluation Needs More Than a One-Turn Benchmark
An agentic coding task can involve repository inspection, tool selection, file edits, test execution, failure recovery, and final verification.
Opus 4.5 can participate in these workflows through Claude tool use, including forced tool-choice modes when extended thinking is not active.
A critical compatibility detail is that manual extended thinking restricts tool choice. When thinking.type is enabled, Anthropic supports tool_choice: auto or tool_choice: none; forced any or named-tool selection is incompatible with manual extended thinking.
This creates a design tension unique to the older reasoning API. If an application requires guaranteed tool invocation, it may need to avoid manual thinking on that turn or restructure the workflow.
For agent benchmarks, measure completion rate, tool-call correctness, retries, human interventions, total output, thinking-token usage, and whether forced-tool assumptions are part of the current integration.
- 01
Repository coding
Use tools and code context to implement, debug, or review changes across multiple files.
- 02
Analytical workflows
Apply extended thinking to difficult comparisons, planning, research synthesis, and multi-step reasoning.
- 03
Tool orchestration
Use automatic tool selection during manual extended thinking and preserve thinking blocks across tool-result turns.
- 04
Migration baselines
Record legacy agent completion and tool behavior before moving to adaptive-thinking Opus models.
05 / Pricing
Claude Opus 4.5 Pricing
Claude Opus 4.5 costs $5 per million input tokens and $25 per million output tokens, with prompt caching and a 50% Batch API discount.
Reasoning Tokens Are Part of Output Economics
Anthropic lists 5-minute cache writes at $6.25 per million tokens, 1-hour cache writes at $10 per million tokens, and cache reads at $0.50 per million.
These rates are identical to several later legacy Opus releases, but the workload economics can still differ because Opus 4.5 manually allocates thinking tokens inside the output envelope.
For extended-thinking requests, reasoning tokens contribute to output consumption. A benchmark that compares only the visible final answer can therefore underestimate the cost of a reasoning-heavy request.
Prompt caching remains valuable for stable prefixes such as system instructions, large documents, repository context, schemas, or recurring policies.
Batch API workloads receive a 50% discount on input and output, which is useful for offline evaluations or large non-interactive jobs.
- Input: $5.00 per 1M tokens.
- Output: $25.00 per 1M tokens.
- 5-minute cache write: $6.25 per 1M tokens.
- 1-hour cache write: $10.00 per 1M tokens.
- Cache read: $0.50 per 1M tokens.
- Batch API: 50% discount on input and output.
1M tokens · USD
- Input
- $5.00
- 5m cache write
- $6.25
- 1h cache write
- $10.00
- Cache read
- $0.50
- Output
- $25.00
Example: 20K input + 4K output
- Input cost
- $0.1000
- Output cost
- $0.1000
- Estimated visible-output total
- $0.2000
06 / Context
200K Context and 64K Maximum Output
Claude Opus 4.5 has a 200,000-token context window and supports up to 64,000 output tokens—a substantially smaller working envelope than the 1M/128K profile of newer Opus generations.
Context Limits Shape Legacy Agent Architecture
A 200K context window remains large enough for substantial source code, long documents, conversation history, and tool results, but persistent agents can reach the ceiling much sooner than on current 1M-context Opus models.
This affects architectural choices. Older applications may rely more heavily on retrieval, summarization, prompt caching, conversation trimming, or explicit memory layers to keep requests inside the context limit.
Output planning matters even more when extended thinking is enabled because thinking and visible response tokens share the 64K maximum output envelope.
The context difference is therefore a major migration benefit. Moving from Opus 4.5 to current Opus can provide much more room for repositories, research material, agent state, and large generated artifacts even before comparing reasoning quality.
Context window
200,000
Max output
64,000
Extended-thinking tokens and visible text share the output budget. Newer Opus models provide a much larger 1M context and 128K standard output envelope.
07 / Snapshot ID
Why the Model ID Ends in `20251101`
The EidoStack page uses Anthropic’s pinned Claude API identifier claude-opus-4-5-20251101, which is different from the model’s public release date of November 24, 2025.
Snapshot Identity and Release Date Are Separate Concepts
Anthropic exposes both a pinned snapshot ID and an alias.
The pinned ID is claude-opus-4-5-20251101. The alias is claude-opus-4-5.
A pinned snapshot gives an application an explicit version target rather than relying only on an alias. That is useful for reproducible evaluations, regression testing, and production systems that want to know exactly which model identifier was requested.
The 20251101 suffix should not be interpreted as the public release date. Anthropic’s model page lists November 24, 2025 as the release date.
Cloud-platform identifiers also differ. Amazon Bedrock uses anthropic.claude-opus-4-5-20251101-v1:0, while Google Cloud uses claude-opus-4-5@20251101.
- 01
Pinned Claude API ID
claude-opus-4-5-20251101 identifies the specific snapshot used by this EidoStack model page.
- 02
Claude API alias
claude-opus-4-5 is the shorter provider alias for the Opus 4.5 model family.
- 03
Release date
Anthropic lists November 24, 2025 as the public release date, which is later than the date encoded in the snapshot ID.
- 04
Cross-cloud IDs
Bedrock and Google Cloud use platform-specific identifiers derived from the same versioned model.
08 / Legacy API
Legacy API Behaviors That Matter During Migration
Opus 4.5 predates several breaking API changes introduced in Opus 4.6 and later, so an old integration may depend on behaviors that current Opus models reject.
Prefill, Beta Headers, and Older Output Controls
Anthropic’s migration guide explicitly tells Opus 4.5 users to remove assistant-message prefills when moving to current Opus because Opus 4.6 and later reject that pattern.
Older integrations may also carry beta headers for effort, fine-grained tool streaming, or interleaved thinking. Current adaptive-thinking and effort APIs no longer require those legacy headers.
Structured-output configuration also evolved. Integrations using older output_format patterns should migrate to the current output_config.format path where applicable.
Tool use has its own compatibility detail. Manual extended thinking supports automatic or no-tool selection, but forced any or named-tool modes are incompatible while thinking is enabled.
These differences are easy to miss if a migration test checks only whether the new model returns an answer. Replay the complete request body, headers, tool definitions, history format, and output parser.
- 01
Assistant prefill
Legacy Opus 4.5 integrations may use prefills; Opus 4.6 and later reject assistant-message prefill.
- 02
Legacy beta headers
Remove obsolete effort, fine-grained tool streaming, and interleaved-thinking beta headers when moving to current APIs.
- 03
Output configuration
Migrate older output_format usage to the current output_config.format contract where applicable.
- 04
Tool-choice constraints
Manual extended thinking cannot be combined with forced any-tool or named-tool selection.
09 / Migration
Migrating From Claude Opus 4.5 to Opus 5.5
Moving from Opus 4.5 to current Opus changes reasoning from manual budgeted thinking to always-on adaptive thinking, expands context and output capacity, lowers base pricing, and removes several legacy request patterns.
This Is a Reasoning-Architecture Migration
The biggest change is not the model ID. It is the thinking contract.
Opus 4.5 uses thinking: {"type": "enabled", "budget_tokens": ...}. Opus 5.5 rejects that format. Thinking is always on, and applications control depth with effort instead of setting a manual token budget.
The context window increases from 200K to 1M tokens, and standard maximum output grows from 64K to 128K. Standard pricing falls from $5/$25 to $4/$20 per million input/output tokens.
Several request formats also change. Remove assistant prefills, incompatible sampling parameters, legacy beta headers, and forced tool selection that current Opus does not accept.
Do not mechanically convert a historical budget_tokens value into an effort level. Anthropic recommends testing effort levels on your own evaluations because the controls operate differently.
- 01
Replace manual thinking budgets
Remove thinking enabled plus budget_tokens and use the current adaptive-thinking effort model.
- 02
Expand context assumptions
Re-evaluate retrieval, summarization, max-output planning, and memory architecture against the newer 1M/128K capacity.
- 03
Clean up legacy request fields
Remove assistant prefills, obsolete beta headers, incompatible sampling settings, and unsupported forced-tool patterns.
- 04
Re-run production evals
Measure completion quality, latency, tool behavior, output consumption, cache economics, and accepted-result cost before cutover.
10 / Evaluation
Claude Opus 4.5 Strengths and Limitations
Claude Opus 4.5 remains valuable for compatibility testing and historical reasoning baselines, but its 200K context, 64K output, manual thinking budgets, and legacy API surface make it a very different integration target from current Opus models.
Strengths
Explicit reasoning control
Manual budget_tokens lets legacy applications set a concrete extended-thinking allowance instead of relying only on adaptive behavior.
Unique effort + extended-thinking combination
Opus 4.5 is the only extended-thinking-only Claude model that also supports the effort parameter.
Pinned snapshot ID
The dated model identifier is useful for reproducible historical evaluations and regression baselines.
Mature tool and caching support
Tool use, prompt caching, Batch API, text input, and image input make it capable of substantial production workflows.
What to consider
Much smaller context than current Opus
200K context and 64K output provide significantly less working capacity than the 1M/128K profile of newer Opus models.
Legacy reasoning API
Manual budget_tokens is replaced by adaptive thinking in newer generations and requires migration work.
Legacy request patterns
Assistant prefill and older beta-header conventions can break when moved directly to current Opus models.
Higher price than Opus 5.5
$5/$25 base pricing is higher than the current Opus 5.5 rate of $4/$20.
Preserve the legacy baseline
Compare Claude Opus 4.5 with current Opus models
Replay coding tasks, tool workflows, extended-thinking prompts, cached sessions, image inputs, and long-context requests to measure how newer Opus models change reasoning behavior, capacity, cost, and compatibility.
Start FreeClaude Opus 4.5 is still available, but Anthropic recommends migrating to Claude Opus 5.5 for improved performance.
Common Questions
What is Claude Opus 4.5?
Claude Opus 4.5 is an Anthropic legacy frontier model released on November 24, 2025. It uses manual extended thinking with budget_tokens and has a 200K context window with up to 64K output tokens.
What is the Claude Opus 4.5 model ID?
The pinned Claude API model ID is claude-opus-4-5-20251101. Anthropic also provides the alias claude-opus-4-5.
Why does the Claude Opus 4.5 model ID contain 20251101 if it was released on November 24?
20251101 is part of Anthropic’s pinned snapshot identifier. Anthropic separately lists November 24, 2025 as the public release date, so the snapshot suffix should not be interpreted as the release date.
How much does Claude Opus 4.5 cost?
Claude Opus 4.5 costs $5 per 1M input tokens and $25 per 1M output tokens. Five-minute cache writes cost $6.25/MTok, one-hour cache writes cost $10/MTok, and cache reads cost $0.50/MTok. Batch input and output receive a 50% discount.
What is the Claude Opus 4.5 context window?
Claude Opus 4.5 has a 200,000-token context window and supports up to 64,000 output tokens.
What is Claude Opus 4.5’s knowledge cutoff?
Anthropic lists May 2025 as the reliable knowledge cutoff and August 2025 as the training-data cutoff.
Does Claude Opus 4.5 support images?
Yes. Claude Opus 4.5 accepts text and image input and produces text output.
Does Claude Opus 4.5 support adaptive thinking?
No. Claude Opus 4.5 is an extended-thinking-only model. Adaptive thinking is not available and type adaptive returns an error. Use manual extended thinking with budget_tokens when reasoning is required.
Does Claude Opus 4.5 support effort?
Yes. Anthropic identifies Opus 4.5 as the only extended-thinking-only Claude model that supports effort. Effort works alongside budget_tokens: effort shapes overall behavior while budget_tokens controls reasoning depth.
Can Claude Opus 4.5 force a tool call while extended thinking is enabled?
No. Manual extended thinking supports tool_choice auto or none. Forced any-tool or named-tool selection is incompatible with manual extended thinking.
Does Claude Opus 4.5 support assistant prefilling?
Legacy Opus 4.5 integrations can use assistant-message prefills, but Anthropic’s migration guide instructs developers to remove them because Opus 4.6 and later reject assistant prefill.
How is Claude Opus 4.5 different from Claude Opus 4.6?
Opus 4.5 has 200K context, 64K output, and only manual extended thinking. Opus 4.6 expands to 1M context and 128K output, adds adaptive thinking while temporarily retaining the deprecated budget_tokens path, and changes several legacy API behaviors.
How is Claude Opus 4.5 different from Claude Opus 5.5?
Opus 5.5 uses always-on adaptive thinking, has 1M context and 128K output, and costs $4/$20 per million input/output tokens. Opus 4.5 uses manual extended thinking, has 200K context and 64K output, and costs $5/$25.
Is Claude Opus 4.5 still available?
Yes. Anthropic lists Claude Opus 4.5 as Active (legacy). It was released November 24, 2025, with retirement not scheduled sooner than November 24, 2026.
Should I use Claude Opus 4.5 for a new application?
For new development, evaluate the current Opus generation first. Claude Opus 4.5 is most useful for existing integrations, reproducible snapshot testing, legacy extended-thinking compatibility, and migration baselines.
Model information
Last updated
Specifications, pricing, lifecycle, model IDs, extended-thinking behavior, effort controls, tool-use constraints, and migration guidance on this page are based on Anthropic’s official Claude Opus 4.5 overview, extended-thinking documentation, effort documentation, tool-use documentation, and Opus 5.5 migration guide.