Anthropic deprecated Sonnet model
Claude Sonnet 4.5
A dated Sonnet snapshot from the manual extended-thinking era, built for complex agents and coding with a 200K-token context window, 64K output, image input, tool use, prompt caching, and explicit reasoning budgets.
- Context window
- 200K
- tokens
- Max output
- 64K
- tokens
- Input
- $3.00
- per 1M tokens
- Cache read
- $0.30
- per 1M tokens
- Output
- $15.00
- per 1M tokens
01 / Overview
What Claude Sonnet 4.5 Is
Claude Sonnet 4.5 is Anthropic’s September 2025 Sonnet snapshot for complex agents, coding, analysis, and other high-capability production workloads from the generation before adaptive thinking became the standard Claude reasoning interface.
A Dated Snapshot From the Manual-Reasoning Era
Anthropic launched Claude Sonnet 4.5 on September 29, 2025 and described it at launch as its best model for complex agents and coding.
The pinned Claude API model ID is claude-sonnet-4-5-20250929, while claude-sonnet-4-5 is the provider alias.
The model provides a 200,000-token context window and supports up to 64,000 output tokens. It accepts text and images and produces text.
Its reasoning system is fundamentally different from Sonnet 4.6 and newer generations. Sonnet 4.5 supports manual extended thinking only. Applications explicitly enable thinking and allocate a budget_tokens value rather than using adaptive thinking and effort.
Anthropic deprecated Sonnet 4.5 on September 30, 2026. It remains functional during the deprecation window, but Claude API retirement is scheduled for November 30, 2026. Anthropic recommends Claude Sonnet 5.5 as the replacement.
- Model ID:
claude-sonnet-4-5-20250929. - Released September 29, 2025.
- Deprecated September 30, 2026.
- Claude API retirement scheduled for November 30, 2026.
- 200K context window.
- 64K maximum output.
- Text and image input with text output.
- Manual extended thinking only.
- Provider
- Anthropic
- Family
- Claude Sonnet 4.5
- Model ID
- claude-sonnet-4-5-20250929
- Reliable knowledge cutoff
- Jan 2025
- Input modalities
- Text, Image
- Output modality
- Text
- Lifecycle
- Deprecated
02 / Lifecycle
Claude Sonnet 4.5 Is Deprecated
Claude Sonnet 4.5 is no longer recommended for new development, and Anthropic has scheduled its Claude API retirement for November 30, 2026.
The Retirement Date Changes the Evaluation Goal
For an active model, evaluation asks whether the model is good enough for a new workload.
For a deprecated model, the question changes: what must be preserved when the dependency is removed?
A migration suite should capture representative production behavior before cutover. This includes difficult prompts, tool trajectories, extended-thinking configurations, cached conversations, image inputs, structured responses, latency, and failure cases.
The dated snapshot ID makes Sonnet 4.5 useful as a reproducible historical baseline. You can compare the same fixed model version against Sonnet 5.5 rather than testing two moving aliases.
However, reproducibility should not become a reason to delay migration. After Claude API retirement, applications that still depend on claude-sonnet-4-5-20250929 need another supported model path.
- 01
September 29, 2025
Anthropic launched Claude Sonnet 4.5 for complex agents and coding.
- 02
September 30, 2026
Anthropic announced the model’s deprecation.
- 03
November 30, 2026
Scheduled Claude API retirement date for claude-sonnet-4-5-20250929.
- 04
Recommended replacement
Anthropic recommends migrating production workloads to Claude Sonnet 5.5.
03 / Extended thinking
Manual Extended Thinking With `budget_tokens`
Claude Sonnet 4.5 uses Anthropic’s legacy manual extended-thinking system: the application explicitly decides whether the model should think and how large the reasoning budget should be.
Reasoning Depth Is an Explicit Token Budget
Thinking is enabled with thinking: {"type": "enabled", "budget_tokens": N}.
The minimum budget_tokens value is 1,024 tokens.
Outside interleaved-thinking mode, the thinking budget must be smaller than max_tokens, because reasoning tokens and the visible final response share the same output ceiling.
The budget is a target rather than a promise that Claude will consume every allocated token. The model can finish reasoning early when the problem does not require the full budget.
This gives a legacy application a direct control that newer adaptive-thinking models intentionally remove: a numerical reasoning allowance.
That can be useful when a workload was tuned around predictable maximum reasoning spend or latency. It also creates migration work, because Sonnet 5.5 does not accept manual budget_tokens.
Sonnet 4.5 does not support adaptive thinking. Sending thinking.type: "adaptive" returns an error.
Simple request
Omit thinking when the workload does not benefit from explicit reasoning.
Difficult request
Enable extended thinking and allocate a measured budget based on quality, latency, and cost evaluations.
04 / Coding & agents
A Historical Baseline for Complex Agents and Coding
Anthropic launched Sonnet 4.5 as its strongest model at the time for complex agents and coding, making it a useful historical benchmark for long-running engineering and tool workflows.
Evaluate the Trajectory, Not Just the Final Answer
A coding agent can read repository files, search symbols, call tools, edit code, run tests, inspect failures, revise its plan, and repeat several times before the user receives the final result.
Sonnet 4.5’s launch positioning makes these trajectories particularly relevant when migrating.
A useful regression suite should measure task completion rate, code correctness, validation success, tool-call count, repeated work, reasoning-token usage, human intervention, latency to accepted result, and total cost per completed task.
This is especially important because Sonnet 5.5 changes the reasoning and tool contract. A successor can be more capable overall while still requiring application changes to reproduce the old workflow.
- 01
Repository coding
Work across multiple files, dependencies, tests, and tool calls instead of generating isolated snippets.
- 02
Long debugging loops
Form hypotheses, inspect evidence, make changes, and repeat validation until the failure is resolved.
- 03
Tool-based research
Use tools and retrieved evidence across multiple steps while preserving the original objective.
- 04
Migration regression sets
Capture successful Sonnet 4.5 trajectories and replay them against Sonnet 5.5 before production cutover.
05 / Pricing
Claude Sonnet 4.5 Pricing
Claude Sonnet 4.5 costs $3 per million input tokens and $15 per million output tokens, with separate prompt-caching rates and a 50% Batch API discount.
Legacy Economics Before the $2/$10 Sonnet Generation
Anthropic lists 5-minute cache writes at $3.75 per million tokens, 1-hour cache writes at $6 per million, and cache reads at $0.30 per million.
Prompt caching can reduce repeated input cost when the application reuses system instructions, repository content, policies, documents, tool definitions, or other stable prefixes.
Batch API processing receives a 50% discount on standard input and output pricing.
Current Sonnet 5.5 is cheaper at $2 input and $10 output per million tokens. The nominal rate reduction is significant, but migration economics are not a direct price-ratio calculation because newer Claude generations also use different tokenization.
The right comparison is cost per accepted result across the same workload, including thinking, tools, retries, cache operations, and any fallback calls.
- Input: $3.00 per 1M tokens.
- Output: $15.00 per 1M tokens.
- 5-minute cache write: $3.75 per 1M tokens.
- 1-hour cache write: $6.00 per 1M tokens.
- Cache read: $0.30 per 1M tokens.
- Batch API: 50% discount on input and output.
1M tokens · USD
- Input
- $3.00
- 5m cache write
- $3.75
- 1h cache write
- $6.00
- Cache read
- $0.30
- Output
- $15.00
Example: 20K input + 4K output
- Input cost
- $0.0600
- Output cost
- $0.0600
- Estimated total before extra thinking output
- $0.1200
06 / Context
200K Context and 64K Maximum Output
Claude Sonnet 4.5’s standard model profile is a 200,000-token context window with up to 64,000 output tokens.
A Smaller Working Envelope Than Modern Sonnet
The 200K context window can still hold substantial code, documentation, conversation history, images, retrieved evidence, and tool results.
But it is much smaller than the 1M-token context window of Sonnet 4.6, Sonnet 5, and Sonnet 5.5.
The output ceiling is also half the 128K standard output limit of those newer generations.
These limits shape agent architecture. A Sonnet 4.5 deployment may rely more heavily on retrieval, summarization, trimming, or external memory to keep a long task inside the available window.
Manual extended thinking further affects output planning because thinking tokens consume output capacity together with the visible response.
When migrating, do not assume old context-management logic is still optimal. A 1M-context successor can support a different balance between retrieval, compression, retained history, and working state.
Context window
200,000
Max output
64,000
Manual thinking tokens and visible response text share the output envelope. Current Sonnet generations provide substantially larger 1M context and 128K standard output capacity.
07 / Tools & thinking
Interleaved Thinking Requires the Legacy Beta Header
Claude Sonnet 4.5 can reason between tool calls, but manual interleaved thinking requires the interleaved-thinking-2025-05-14 beta header.
Tool Use Has Manual-Thinking Constraints
Interleaved thinking lets Claude examine a tool result and reason again before choosing the next action.
On Sonnet 4.5, this behavior is not automatic. The application enables it through the legacy beta header.
Manual extended thinking also restricts tool_choice. While thinking is enabled, Anthropic supports automatic tool selection or no tool. Forced tool_choice: any and forced named-tool selection are incompatible with manual extended thinking.
Response prefilling follows a similar pattern: older models can use prefilled assistant responses when thinking is off, but prefilling cannot be combined with active thinking.
Thinking-block persistence is another generational difference. Sonnet 4.5 strips earlier thinking blocks from context when later non-tool-result user content is added. Sonnet 4.6 and newer preservation behavior is different.
That can affect prompt caching and long conversation replay, so migrate complete multi-turn tool traces rather than evaluating only isolated requests.
- 01
Beta interleaving
Use the interleaved-thinking-2025-05-14 beta header to let manual extended thinking resume between tool calls.
- 02
No forced tools during thinking
Manual thinking supports automatic or no-tool selection, not forced any-tool or named-tool modes.
- 03
Prefill only without thinking
Assistant-response prefilling cannot be combined with an active manual thinking request.
- 04
Older thinking-block retention
Prior Sonnet 4.5 thinking blocks are stripped when later non-tool-result user content enters the conversation, unlike newer preservation behavior.
08 / Tokenizer
Sonnet 4.5 Uses the Earlier Claude Tokenizer Generation
Claude Sonnet 4.5 belongs to the older tokenizer generation, making it a useful baseline when comparing token counts and context economics with newer Claude models.
Equivalent Text Can Consume More Tokens After Migration
Anthropic documents that Claude 4.7 and later models use a newer tokenizer that can produce roughly 30% more tokens for the same text, depending on workload content.
Sonnet 4.5 therefore sits on the earlier side of that boundary.
A migration should re-measure token counts for real prompts, context-window headroom, max_tokens, monthly cost projections, rate-limit assumptions, prompt-caching economics, and trimming thresholds.
This matters because Sonnet 5.5 has a larger context window and lower nominal token price, yet the same source text can occupy a different number of tokens.
Never convert a Sonnet 4.5 cost forecast to Sonnet 5.5 using price alone. Recount the actual payload.
- 01
Recount production prompts
Use real request payloads rather than carrying Sonnet 4.5 token estimates into a newer model.
- 02
Recheck context thresholds
Equivalent text can occupy a different amount of the context window after migration.
- 03
Recalculate output headroom
Revisit max_tokens and visible-output assumptions under the successor’s tokenizer and reasoning behavior.
- 04
Compare measured economics
Combine real token counts with the successor’s lower pricing instead of assuming the headline price reduction equals the workload saving.
09 / Migration
Migrating From Claude Sonnet 4.5 to Sonnet 5.5
The recommended replacement changes almost every important integration dimension: thinking, context capacity, output size, pricing, tokenization, tool orchestration, conversation state, and several legacy request patterns.
This Is Not a Model-ID-Only Upgrade
Sonnet 4.5 runs without thinking unless manual extended thinking is explicitly enabled.
Sonnet 5.5 runs with adaptive thinking when the thinking field is omitted. It does not accept thinking.type: "enabled" or budget_tokens.
That means the old reasoning configuration must be removed rather than copied.
Capacity changes from 200K context / 64K output to 1M context / 128K standard output.
Standard pricing falls from $3/$15 to $2/$10 per million input/output tokens.
Tool behavior also changes. Manual extended thinking’s legacy orchestration patterns do not map directly to the current model, and Sonnet 5.5 itself does not support forced any-tool or named-tool choice.
Assistant prefill should also be removed. Claude 4.6 and later models reject prefilled assistant messages entirely.
Finally, re-evaluate token counts and retained conversation state. The tokenizer generation changes, and newer models preserve thinking state differently.
- 01
Replace budget_tokens
Remove manual extended thinking and migrate reasoning evaluation to Sonnet 5.5 adaptive thinking and effort.
- 02
Remove legacy request patterns
Retest assistant prefill, forced tool workflows, interleaved-thinking headers, sampling controls, and response parsing.
- 03
Rebuild context assumptions
Take advantage of the larger 1M/128K envelope while re-measuring token counts and conversation-state behavior.
- 04
Finish before retirement
Validate production equivalence and deploy the replacement before the November 30, 2026 Claude API retirement date.
10 / Capabilities
Claude Sonnet 4.5 Capabilities
Sonnet 4.5 combines text and image input, manual extended thinking, tool use, prompt caching, Batch API processing, and a dated snapshot identity in a deprecated pre-adaptive-thinking Sonnet model.
A Mature Legacy Feature Set
The model accepts text and images and produces text.
Manual extended thinking provides explicit reasoning budgets. Tool use supports agent workflows, with interleaved thinking available through a beta header and restrictions on forced tool choice while thinking is active.
Prompt caching reduces repeated processing cost for stable context. Batch API execution receives a 50% input/output discount.
The pinned snapshot ID is useful for reproducible regression tests because it identifies a specific model release.
The primary constraint is now lifecycle rather than raw feature coverage: the model is deprecated and has a scheduled retirement date.
- Supported
Text
Accept text input and produce text output.
- Supported
Image input
Analyze supported images together with text context.
- Supported
Manual extended thinking
Use thinking enabled plus budget_tokens for explicit reasoning allocation.
- Not listed
Adaptive thinking
Sonnet 4.5 does not support thinking type adaptive.
- Supported
Tool use
Supports tool workflows with legacy manual-thinking constraints and optional interleaved thinking.
- Supported
Prompt caching
Reduce repeated input cost for stable prompt prefixes.
- Supported
Batch API
Batch input and output receive a 50% discount.
11 / Evaluation
Claude Sonnet 4.5 Strengths and Limitations
Claude Sonnet 4.5 remains valuable as a reproducible coding and agent benchmark, but its deprecated lifecycle, 200K context, 64K output, manual thinking API, and legacy conversation semantics make migration the primary production priority.
Strengths
Reproducible dated snapshot
The pinned 20250929 ID provides a stable historical target for regression testing and migration comparisons.
Explicit reasoning budget
Manual budget_tokens gives legacy applications direct control over maximum reasoning allocation.
Strong historical coding baseline
Anthropic launched Sonnet 4.5 specifically as its strongest model at the time for complex agents and coding.
Mature legacy tooling
Image input, tools, prompt caching, interleaved thinking, and Batch API support cover substantial production workflows.
What to consider
Scheduled retirement
Claude API retirement is scheduled for November 30, 2026, so new production dependencies should not be built on Sonnet 4.5.
Smaller context and output
200K context and 64K output are substantially below current Sonnet’s 1M/128K capacity.
Legacy reasoning contract
Adaptive thinking is unavailable and budget_tokens must be replaced when moving to current Sonnet models.
Older tokenizer and state semantics
Token counts and thinking-block preservation differ from newer generations, so migration needs real conversation and cost testing.
Retire the legacy dependency with evidence
Compare Claude Sonnet 4.5 with Sonnet 5.5
Replay coding, tool-use, reasoning, cached-context, image, and long-session workloads before retirement to measure compatibility, quality, token usage, latency, and cost under the current Sonnet generation.
Start FreeAnthropic deprecated Claude Sonnet 4.5 on September 30, 2026 and recommends migrating to Claude Sonnet 5.5 before Claude API retirement on November 30, 2026.
Common Questions
What is Claude Sonnet 4.5?
Claude Sonnet 4.5 is an Anthropic model released on September 29, 2025 for complex agents, coding, analysis, and other high-capability workloads. Its pinned Claude API ID is claude-sonnet-4-5-20250929.
Is Claude Sonnet 4.5 deprecated?
Yes. Anthropic announced deprecation on September 30, 2026 and scheduled Claude API retirement for November 30, 2026. Anthropic recommends migrating to Claude Sonnet 5.5.
How much does Claude Sonnet 4.5 cost?
Claude Sonnet 4.5 costs $3 per 1M input tokens and $15 per 1M output tokens. Five-minute cache writes cost $3.75/MTok, one-hour cache writes cost $6/MTok, and cache reads cost $0.30/MTok. Batch input and output receive a 50% discount.
What is the Claude Sonnet 4.5 context window?
Claude Sonnet 4.5 has a 200,000-token standard context window and supports up to 64,000 output tokens.
What is Claude Sonnet 4.5’s knowledge cutoff?
Anthropic lists January 2025 as the reliable knowledge cutoff for Claude Sonnet 4.5.
Does Claude Sonnet 4.5 support images?
Yes. Claude Sonnet 4.5 accepts text and image input and generates text output.
Does Claude Sonnet 4.5 support adaptive thinking?
No. Claude Sonnet 4.5 is a manual extended-thinking model. Use thinking type enabled with budget_tokens when reasoning is required; adaptive thinking is available only on newer generations.
What is the minimum budget_tokens value on Claude Sonnet 4.5?
Anthropic requires at least 1,024 thinking tokens. Outside interleaved-thinking mode, budget_tokens must also be smaller than max_tokens so the model has output capacity left for the final answer.
Does Claude Sonnet 4.5 support effort?
No. Anthropic lists no default effort for Sonnet 4.5. Reasoning depth is controlled by budget_tokens rather than the adaptive-thinking effort system.
Does Claude Sonnet 4.5 support interleaved thinking?
Yes. For manual extended thinking, add the interleaved-thinking-2025-05-14 beta header so Claude can reason between tool calls.
Can Claude Sonnet 4.5 force a tool call while thinking is enabled?
No. Manual extended thinking supports automatic tool selection or no tool. Forced any-tool or named-tool choice is incompatible with active manual thinking.
How is Claude Sonnet 4.5 different from Claude Sonnet 4.6?
Sonnet 4.5 has 200K context, 64K output, and manual extended thinking only. Sonnet 4.6 expands to 1M context and 128K output, adds adaptive thinking and effort while temporarily retaining deprecated budget_tokens compatibility, and changes thinking-block preservation.
How is Claude Sonnet 4.5 different from Claude Sonnet 5.5?
Sonnet 5.5 provides 1M context, 128K standard output, adaptive thinking, newer agent controls, and $2/$10 pricing. Sonnet 4.5 has 200K context, 64K output, manual budget_tokens thinking, older tokenizer behavior, and $3/$15 pricing.
Why is the model ID claude-sonnet-4-5-20250929?
The ID is Anthropic’s pinned dated snapshot identifier. The 20250929 suffix corresponds to the model generation launched on September 29, 2025 and is useful for reproducible version targeting.
Should I use Claude Sonnet 4.5 for a new application?
No new long-lived production dependency should be designed around a model scheduled for retirement. Use Sonnet 4.5 for existing integrations, regression testing, and migration verification, and evaluate Claude Sonnet 5.5 for new development.
Model information
Last updated
Specifications, pricing, deprecation status, retirement date, manual extended-thinking behavior, interleaved-thinking requirements, tokenizer generation, cache behavior, and migration guidance on this page are based on Anthropic’s official Claude Sonnet 4.5 model documentation, release notes, extended-thinking documentation, prompt-caching documentation, and Sonnet 5.5 migration guide.