Anthropic deprecated Sonnet model

Claude Sonnet 4.5

A dated Sonnet snapshot from the manual extended-thinking era, built for complex agents and coding with a 200K-token context window, 64K output, image input, tool use, prompt caching, and explicit reasoning budgets.

Context window
200K
tokens
Max output
64K
tokens
Input
$3.00
per 1M tokens
Cache read
$0.30
per 1M tokens
Output
$15.00
per 1M tokens

01 / Overview

What Claude Sonnet 4.5 Is

Claude Sonnet 4.5 is Anthropic’s September 2025 Sonnet snapshot for complex agents, coding, analysis, and other high-capability production workloads from the generation before adaptive thinking became the standard Claude reasoning interface.

A Dated Snapshot From the Manual-Reasoning Era

Anthropic launched Claude Sonnet 4.5 on September 29, 2025 and described it at launch as its best model for complex agents and coding.

The pinned Claude API model ID is claude-sonnet-4-5-20250929, while claude-sonnet-4-5 is the provider alias.

The model provides a 200,000-token context window and supports up to 64,000 output tokens. It accepts text and images and produces text.

Its reasoning system is fundamentally different from Sonnet 4.6 and newer generations. Sonnet 4.5 supports manual extended thinking only. Applications explicitly enable thinking and allocate a budget_tokens value rather than using adaptive thinking and effort.

Anthropic deprecated Sonnet 4.5 on September 30, 2026. It remains functional during the deprecation window, but Claude API retirement is scheduled for November 30, 2026. Anthropic recommends Claude Sonnet 5.5 as the replacement.

  • Model ID: claude-sonnet-4-5-20250929.
  • Released September 29, 2025.
  • Deprecated September 30, 2026.
  • Claude API retirement scheduled for November 30, 2026.
  • 200K context window.
  • 64K maximum output.
  • Text and image input with text output.
  • Manual extended thinking only.
Model profile
Provider
Anthropic
Family
Claude Sonnet 4.5
Model ID
claude-sonnet-4-5-20250929
Reliable knowledge cutoff
Jan 2025
Input modalities
Text, Image
Output modality
Text
Lifecycle
Deprecated

02 / Lifecycle

Claude Sonnet 4.5 Is Deprecated

Claude Sonnet 4.5 is no longer recommended for new development, and Anthropic has scheduled its Claude API retirement for November 30, 2026.

The Retirement Date Changes the Evaluation Goal

For an active model, evaluation asks whether the model is good enough for a new workload.

For a deprecated model, the question changes: what must be preserved when the dependency is removed?

A migration suite should capture representative production behavior before cutover. This includes difficult prompts, tool trajectories, extended-thinking configurations, cached conversations, image inputs, structured responses, latency, and failure cases.

The dated snapshot ID makes Sonnet 4.5 useful as a reproducible historical baseline. You can compare the same fixed model version against Sonnet 5.5 rather than testing two moving aliases.

However, reproducibility should not become a reason to delay migration. After Claude API retirement, applications that still depend on claude-sonnet-4-5-20250929 need another supported model path.

Snapshot and retirement
  1. 01

    September 29, 2025

    Anthropic launched Claude Sonnet 4.5 for complex agents and coding.

  2. 02

    September 30, 2026

    Anthropic announced the model’s deprecation.

  3. 03

    November 30, 2026

    Scheduled Claude API retirement date for claude-sonnet-4-5-20250929.

  4. 04

    Recommended replacement

    Anthropic recommends migrating production workloads to Claude Sonnet 5.5.

03 / Extended thinking

Manual Extended Thinking With `budget_tokens`

Claude Sonnet 4.5 uses Anthropic’s legacy manual extended-thinking system: the application explicitly decides whether the model should think and how large the reasoning budget should be.

Reasoning Depth Is an Explicit Token Budget

Thinking is enabled with thinking: {"type": "enabled", "budget_tokens": N}.

The minimum budget_tokens value is 1,024 tokens.

Outside interleaved-thinking mode, the thinking budget must be smaller than max_tokens, because reasoning tokens and the visible final response share the same output ceiling.

The budget is a target rather than a promise that Claude will consume every allocated token. The model can finish reasoning early when the problem does not require the full budget.

This gives a legacy application a direct control that newer adaptive-thinking models intentionally remove: a numerical reasoning allowance.

That can be useful when a workload was tuned around predictable maximum reasoning spend or latency. It also creates migration work, because Sonnet 5.5 does not accept manual budget_tokens.

Sonnet 4.5 does not support adaptive thinking. Sending thinking.type: "adaptive" returns an error.

Manual reasoning budget
off · default if omittedminimum budgetbudget_tokens

Simple request

Omit thinking when the workload does not benefit from explicit reasoning.

Difficult request

Enable extended thinking and allocate a measured budget based on quality, latency, and cost evaluations.

04 / Coding & agents

A Historical Baseline for Complex Agents and Coding

Anthropic launched Sonnet 4.5 as its strongest model at the time for complex agents and coding, making it a useful historical benchmark for long-running engineering and tool workflows.

Evaluate the Trajectory, Not Just the Final Answer

A coding agent can read repository files, search symbols, call tools, edit code, run tests, inspect failures, revise its plan, and repeat several times before the user receives the final result.

Sonnet 4.5’s launch positioning makes these trajectories particularly relevant when migrating.

A useful regression suite should measure task completion rate, code correctness, validation success, tool-call count, repeated work, reasoning-token usage, human intervention, latency to accepted result, and total cost per completed task.

This is especially important because Sonnet 5.5 changes the reasoning and tool contract. A successor can be more capable overall while still requiring application changes to reproduce the old workflow.

Complex agent workloads
  1. 01

    Repository coding

    Work across multiple files, dependencies, tests, and tool calls instead of generating isolated snippets.

  2. 02

    Long debugging loops

    Form hypotheses, inspect evidence, make changes, and repeat validation until the failure is resolved.

  3. 03

    Tool-based research

    Use tools and retrieved evidence across multiple steps while preserving the original objective.

  4. 04

    Migration regression sets

    Capture successful Sonnet 4.5 trajectories and replay them against Sonnet 5.5 before production cutover.

05 / Pricing

Claude Sonnet 4.5 Pricing

Claude Sonnet 4.5 costs $3 per million input tokens and $15 per million output tokens, with separate prompt-caching rates and a 50% Batch API discount.

Legacy Economics Before the $2/$10 Sonnet Generation

Anthropic lists 5-minute cache writes at $3.75 per million tokens, 1-hour cache writes at $6 per million, and cache reads at $0.30 per million.

Prompt caching can reduce repeated input cost when the application reuses system instructions, repository content, policies, documents, tool definitions, or other stable prefixes.

Batch API processing receives a 50% discount on standard input and output pricing.

Current Sonnet 5.5 is cheaper at $2 input and $10 output per million tokens. The nominal rate reduction is significant, but migration economics are not a direct price-ratio calculation because newer Claude generations also use different tokenization.

The right comparison is cost per accepted result across the same workload, including thinking, tools, retries, cache operations, and any fallback calls.

  • Input: $3.00 per 1M tokens.
  • Output: $15.00 per 1M tokens.
  • 5-minute cache write: $3.75 per 1M tokens.
  • 1-hour cache write: $6.00 per 1M tokens.
  • Cache read: $0.30 per 1M tokens.
  • Batch API: 50% discount on input and output.
Token and cache pricing

1M tokens · USD

Input
$3.00
5m cache write
$3.75
1h cache write
$6.00
Cache read
$0.30
Output
$15.00

Example: 20K input + 4K output

Input cost
$0.0600
Output cost
$0.0600
Estimated total before extra thinking output
$0.1200

06 / Context

200K Context and 64K Maximum Output

Claude Sonnet 4.5’s standard model profile is a 200,000-token context window with up to 64,000 output tokens.

A Smaller Working Envelope Than Modern Sonnet

The 200K context window can still hold substantial code, documentation, conversation history, images, retrieved evidence, and tool results.

But it is much smaller than the 1M-token context window of Sonnet 4.6, Sonnet 5, and Sonnet 5.5.

The output ceiling is also half the 128K standard output limit of those newer generations.

These limits shape agent architecture. A Sonnet 4.5 deployment may rely more heavily on retrieval, summarization, trimming, or external memory to keep a long task inside the available window.

Manual extended thinking further affects output planning because thinking tokens consume output capacity together with the visible response.

When migrating, do not assume old context-management logic is still optimal. A 1M-context successor can support a different balance between retrieval, compression, retained history, and working state.

Context capacity

Context window

200,000

Max output

64,000

Input + working contextOutput limit

Manual thinking tokens and visible response text share the output envelope. Current Sonnet generations provide substantially larger 1M context and 128K standard output capacity.

07 / Tools & thinking

Interleaved Thinking Requires the Legacy Beta Header

Claude Sonnet 4.5 can reason between tool calls, but manual interleaved thinking requires the interleaved-thinking-2025-05-14 beta header.

Tool Use Has Manual-Thinking Constraints

Interleaved thinking lets Claude examine a tool result and reason again before choosing the next action.

On Sonnet 4.5, this behavior is not automatic. The application enables it through the legacy beta header.

Manual extended thinking also restricts tool_choice. While thinking is enabled, Anthropic supports automatic tool selection or no tool. Forced tool_choice: any and forced named-tool selection are incompatible with manual extended thinking.

Response prefilling follows a similar pattern: older models can use prefilled assistant responses when thinking is off, but prefilling cannot be combined with active thinking.

Thinking-block persistence is another generational difference. Sonnet 4.5 strips earlier thinking blocks from context when later non-tool-result user content is added. Sonnet 4.6 and newer preservation behavior is different.

That can affect prompt caching and long conversation replay, so migrate complete multi-turn tool traces rather than evaluating only isolated requests.

Legacy thinking orchestration
  1. 01

    Beta interleaving

    Use the interleaved-thinking-2025-05-14 beta header to let manual extended thinking resume between tool calls.

  2. 02

    No forced tools during thinking

    Manual thinking supports automatic or no-tool selection, not forced any-tool or named-tool modes.

  3. 03

    Prefill only without thinking

    Assistant-response prefilling cannot be combined with an active manual thinking request.

  4. 04

    Older thinking-block retention

    Prior Sonnet 4.5 thinking blocks are stripped when later non-tool-result user content enters the conversation, unlike newer preservation behavior.

08 / Tokenizer

Sonnet 4.5 Uses the Earlier Claude Tokenizer Generation

Claude Sonnet 4.5 belongs to the older tokenizer generation, making it a useful baseline when comparing token counts and context economics with newer Claude models.

Equivalent Text Can Consume More Tokens After Migration

Anthropic documents that Claude 4.7 and later models use a newer tokenizer that can produce roughly 30% more tokens for the same text, depending on workload content.

Sonnet 4.5 therefore sits on the earlier side of that boundary.

A migration should re-measure token counts for real prompts, context-window headroom, max_tokens, monthly cost projections, rate-limit assumptions, prompt-caching economics, and trimming thresholds.

This matters because Sonnet 5.5 has a larger context window and lower nominal token price, yet the same source text can occupy a different number of tokens.

Never convert a Sonnet 4.5 cost forecast to Sonnet 5.5 using price alone. Recount the actual payload.

Pre-new-tokenizer baseline
  1. 01

    Recount production prompts

    Use real request payloads rather than carrying Sonnet 4.5 token estimates into a newer model.

  2. 02

    Recheck context thresholds

    Equivalent text can occupy a different amount of the context window after migration.

  3. 03

    Recalculate output headroom

    Revisit max_tokens and visible-output assumptions under the successor’s tokenizer and reasoning behavior.

  4. 04

    Compare measured economics

    Combine real token counts with the successor’s lower pricing instead of assuming the headline price reduction equals the workload saving.

09 / Migration

Migrating From Claude Sonnet 4.5 to Sonnet 5.5

The recommended replacement changes almost every important integration dimension: thinking, context capacity, output size, pricing, tokenization, tool orchestration, conversation state, and several legacy request patterns.

This Is Not a Model-ID-Only Upgrade

Sonnet 4.5 runs without thinking unless manual extended thinking is explicitly enabled.

Sonnet 5.5 runs with adaptive thinking when the thinking field is omitted. It does not accept thinking.type: "enabled" or budget_tokens.

That means the old reasoning configuration must be removed rather than copied.

Capacity changes from 200K context / 64K output to 1M context / 128K standard output.

Standard pricing falls from $3/$15 to $2/$10 per million input/output tokens.

Tool behavior also changes. Manual extended thinking’s legacy orchestration patterns do not map directly to the current model, and Sonnet 5.5 itself does not support forced any-tool or named-tool choice.

Assistant prefill should also be removed. Claude 4.6 and later models reject prefilled assistant messages entirely.

Finally, re-evaluate token counts and retained conversation state. The tokenizer generation changes, and newer models preserve thinking state differently.

Sonnet 4.5 → Sonnet 5.5
  1. 01

    Replace budget_tokens

    Remove manual extended thinking and migrate reasoning evaluation to Sonnet 5.5 adaptive thinking and effort.

  2. 02

    Remove legacy request patterns

    Retest assistant prefill, forced tool workflows, interleaved-thinking headers, sampling controls, and response parsing.

  3. 03

    Rebuild context assumptions

    Take advantage of the larger 1M/128K envelope while re-measuring token counts and conversation-state behavior.

  4. 04

    Finish before retirement

    Validate production equivalence and deploy the replacement before the November 30, 2026 Claude API retirement date.

10 / Capabilities

Claude Sonnet 4.5 Capabilities

Sonnet 4.5 combines text and image input, manual extended thinking, tool use, prompt caching, Batch API processing, and a dated snapshot identity in a deprecated pre-adaptive-thinking Sonnet model.

A Mature Legacy Feature Set

The model accepts text and images and produces text.

Manual extended thinking provides explicit reasoning budgets. Tool use supports agent workflows, with interleaved thinking available through a beta header and restrictions on forced tool choice while thinking is active.

Prompt caching reduces repeated processing cost for stable context. Batch API execution receives a 50% input/output discount.

The pinned snapshot ID is useful for reproducible regression tests because it identifies a specific model release.

The primary constraint is now lifecycle rather than raw feature coverage: the model is deprecated and has a scheduled retirement date.

Supported capabilities
  • Text

    Accept text input and produce text output.

    Supported
  • Image input

    Analyze supported images together with text context.

    Supported
  • Manual extended thinking

    Use thinking enabled plus budget_tokens for explicit reasoning allocation.

    Supported
  • Adaptive thinking

    Sonnet 4.5 does not support thinking type adaptive.

    Not listed
  • Tool use

    Supports tool workflows with legacy manual-thinking constraints and optional interleaved thinking.

    Supported
  • Prompt caching

    Reduce repeated input cost for stable prompt prefixes.

    Supported
  • Batch API

    Batch input and output receive a 50% discount.

    Supported

11 / Evaluation

Claude Sonnet 4.5 Strengths and Limitations

Claude Sonnet 4.5 remains valuable as a reproducible coding and agent benchmark, but its deprecated lifecycle, 200K context, 64K output, manual thinking API, and legacy conversation semantics make migration the primary production priority.

Strengths

  • Reproducible dated snapshot

    The pinned 20250929 ID provides a stable historical target for regression testing and migration comparisons.

  • Explicit reasoning budget

    Manual budget_tokens gives legacy applications direct control over maximum reasoning allocation.

  • Strong historical coding baseline

    Anthropic launched Sonnet 4.5 specifically as its strongest model at the time for complex agents and coding.

  • Mature legacy tooling

    Image input, tools, prompt caching, interleaved thinking, and Batch API support cover substantial production workflows.

What to consider

  • Scheduled retirement

    Claude API retirement is scheduled for November 30, 2026, so new production dependencies should not be built on Sonnet 4.5.

  • Smaller context and output

    200K context and 64K output are substantially below current Sonnet’s 1M/128K capacity.

  • Legacy reasoning contract

    Adaptive thinking is unavailable and budget_tokens must be replaced when moving to current Sonnet models.

  • Older tokenizer and state semantics

    Token counts and thinking-block preservation differ from newer generations, so migration needs real conversation and cost testing.

Retire the legacy dependency with evidence

Compare Claude Sonnet 4.5 with Sonnet 5.5

Replay coding, tool-use, reasoning, cached-context, image, and long-session workloads before retirement to measure compatibility, quality, token usage, latency, and cost under the current Sonnet generation.

Start Free

Anthropic deprecated Claude Sonnet 4.5 on September 30, 2026 and recommends migrating to Claude Sonnet 5.5 before Claude API retirement on November 30, 2026.

Common Questions

What is Claude Sonnet 4.5?

Claude Sonnet 4.5 is an Anthropic model released on September 29, 2025 for complex agents, coding, analysis, and other high-capability workloads. Its pinned Claude API ID is claude-sonnet-4-5-20250929.

Is Claude Sonnet 4.5 deprecated?

Yes. Anthropic announced deprecation on September 30, 2026 and scheduled Claude API retirement for November 30, 2026. Anthropic recommends migrating to Claude Sonnet 5.5.

How much does Claude Sonnet 4.5 cost?

Claude Sonnet 4.5 costs $3 per 1M input tokens and $15 per 1M output tokens. Five-minute cache writes cost $3.75/MTok, one-hour cache writes cost $6/MTok, and cache reads cost $0.30/MTok. Batch input and output receive a 50% discount.

What is the Claude Sonnet 4.5 context window?

Claude Sonnet 4.5 has a 200,000-token standard context window and supports up to 64,000 output tokens.

What is Claude Sonnet 4.5’s knowledge cutoff?

Anthropic lists January 2025 as the reliable knowledge cutoff for Claude Sonnet 4.5.

Does Claude Sonnet 4.5 support images?

Yes. Claude Sonnet 4.5 accepts text and image input and generates text output.

Does Claude Sonnet 4.5 support adaptive thinking?

No. Claude Sonnet 4.5 is a manual extended-thinking model. Use thinking type enabled with budget_tokens when reasoning is required; adaptive thinking is available only on newer generations.

What is the minimum budget_tokens value on Claude Sonnet 4.5?

Anthropic requires at least 1,024 thinking tokens. Outside interleaved-thinking mode, budget_tokens must also be smaller than max_tokens so the model has output capacity left for the final answer.

Does Claude Sonnet 4.5 support effort?

No. Anthropic lists no default effort for Sonnet 4.5. Reasoning depth is controlled by budget_tokens rather than the adaptive-thinking effort system.

Does Claude Sonnet 4.5 support interleaved thinking?

Yes. For manual extended thinking, add the interleaved-thinking-2025-05-14 beta header so Claude can reason between tool calls.

Can Claude Sonnet 4.5 force a tool call while thinking is enabled?

No. Manual extended thinking supports automatic tool selection or no tool. Forced any-tool or named-tool choice is incompatible with active manual thinking.

How is Claude Sonnet 4.5 different from Claude Sonnet 4.6?

Sonnet 4.5 has 200K context, 64K output, and manual extended thinking only. Sonnet 4.6 expands to 1M context and 128K output, adds adaptive thinking and effort while temporarily retaining deprecated budget_tokens compatibility, and changes thinking-block preservation.

How is Claude Sonnet 4.5 different from Claude Sonnet 5.5?

Sonnet 5.5 provides 1M context, 128K standard output, adaptive thinking, newer agent controls, and $2/$10 pricing. Sonnet 4.5 has 200K context, 64K output, manual budget_tokens thinking, older tokenizer behavior, and $3/$15 pricing.

Why is the model ID claude-sonnet-4-5-20250929?

The ID is Anthropic’s pinned dated snapshot identifier. The 20250929 suffix corresponds to the model generation launched on September 29, 2025 and is useful for reproducible version targeting.

Should I use Claude Sonnet 4.5 for a new application?

No new long-lived production dependency should be designed around a model scheduled for retirement. Use Sonnet 4.5 for existing integrations, regression testing, and migration verification, and evaluate Claude Sonnet 5.5 for new development.

Model information

Last updated

Specifications, pricing, deprecation status, retirement date, manual extended-thinking behavior, interleaved-thinking requirements, tokenizer generation, cache behavior, and migration guidance on this page are based on Anthropic’s official Claude Sonnet 4.5 model documentation, release notes, extended-thinking documentation, prompt-caching documentation, and Sonnet 5.5 migration guide.

Claude Sonnet 4.5 — Pricing, 200K Context, Extended Thinking & Retirement | EidoStack