Anthropic legacy Sonnet model
Claude Sonnet 4.6
A mature pre-Claude-5 Sonnet model that combined 1M context, 128K output, adaptive thinking, legacy extended-thinking compatibility, multimodal input, and strong agentic search at the classic $3/$15 Sonnet price.
- Context window
- 1M
- tokens
- Max output
- 128K
- tokens
- Input
- $3.00
- per 1M tokens
- Cache read
- $0.30
- per 1M tokens
- Output
- $15.00
- per 1M tokens
01 / Overview
What Claude Sonnet 4.6 Is
Claude Sonnet 4.6 is a legacy Anthropic frontier model released on February 17, 2026 as a balanced Sonnet generation for coding, agentic search, tool-heavy workflows, analysis, and everyday professional work.
The Last Major Sonnet Baseline Before Claude 5
Sonnet 4.6 sits at an important architectural boundary.
It expanded the Sonnet tier to a 1,000,000-token context window and 128,000-token standard output limit while introducing adaptive thinking and general-purpose effort controls. At the same time, it preserved backward compatibility with the older manual budget_tokens extended-thinking API.
That makes Sonnet 4.6 different from both sides of the transition. Earlier Sonnet models rely on manual extended thinking. Sonnet 5 removes manual thinking budgets, changes tokenization, and turns adaptive thinking on by default. Sonnet 4.6 can use either adaptive thinking or deprecated extended thinking, and a request that omits thinking still runs without reasoning.
Anthropic launched the model with an emphasis on balanced speed and intelligence and specifically highlighted improved agentic search performance with fewer consumed tokens.
The model remains available as Active (legacy), but Anthropic recommends Claude Sonnet 5.5 for new deployments.
- Model ID:
claude-sonnet-4-6. - Released February 17, 2026.
- Retirement not sooner than February 17, 2027.
- 1M context window and 128K standard output.
- Text and image input with text output.
- Adaptive thinking supported; extended thinking deprecated but still accepted.
- Provider
- Anthropic
- Family
- Claude Sonnet 4.6
- Model ID
- claude-sonnet-4-6
- Reliable knowledge cutoff
- Aug 2025
- Training cutoff
- Jan 2026
- Input modalities
- Text, Image
- Lifecycle
- Active · legacy
02 / Thinking transition
Adaptive Thinking Plus Legacy Extended Thinking
Claude Sonnet 4.6 is one of the transition models that supports Anthropic’s newer adaptive-thinking system while still accepting the older manual extended-thinking format.
Reasoning Is Available, but It Is Not On by Default
A Sonnet 4.6 request with no thinking field runs without thinking.
When reasoning is useful, Anthropic recommends adaptive thinking with thinking: {"type": "adaptive"}. The model then decides when and how much to reason based on task complexity and the configured effort level.
Sonnet 4.6 also still accepts the older extended-thinking format based on thinking.type: "enabled" plus budget_tokens. Anthropic marks that interface as deprecated.
This dual compatibility is useful for legacy systems because an application can migrate reasoning architecture before migrating model generation. It is also the main reason Sonnet 4.6 deserves its own evaluation page rather than being treated as a smaller Sonnet 5.
Anthropic lists high as the API default effort. Current guidance recommends medium as a practical starting point for many applications, low for high-volume or latency-sensitive work, high for difficult reasoning, and max when maximum capability matters more than token consumption.
Sonnet 4.6 does not support the later xhigh effort level.
Adaptive thinking
Use adaptive thinking plus effort for current-style reasoning without manually predicting a token budget.
Legacy compatibility
thinking enabled plus budget_tokens still works on Sonnet 4.6, but Anthropic marks it deprecated and newer Sonnet models reject it.
03 / Agentic search
Agentic Search With Lower Token Consumption
Anthropic introduced Sonnet 4.6 with improved agentic search performance and specifically highlighted lower token consumption for these workflows.
Search Quality Is a Trajectory Problem
Agentic search is different from answering from static context. The model may formulate a query, inspect results, decide whether evidence is sufficient, search again, fetch a document, filter irrelevant material, and repeat until it can synthesize an answer.
For this kind of workload, useful evaluation metrics include search count, evidence coverage, redundant retrieval, total token consumption, stopping behavior, source-grounded answer quality, and latency to a supported conclusion.
Sonnet 4.6 launched at the same time Anthropic made web search and programmatic tool calling generally available without beta headers, and added dynamic filtering for web search and web fetch.
That surrounding platform generation makes Sonnet 4.6 a useful historical baseline for search-heavy agents: it represents the mature pre-Claude-5 stack where tool orchestration, code execution, search, and large context were already central to production workflows.
- 01
Multi-step research
Search, inspect, filter, and retrieve additional evidence until the model has enough information to synthesize an answer.
- 02
Tool-heavy coding
Combine repository context with code execution, search, file inspection, and iterative validation.
- 03
Evidence-aware agents
Measure whether the model finds sufficient support without wasting calls on repetitive retrieval.
- 04
Historical regression baseline
Compare newer Sonnet models against a mature pre-Claude-5 search and tool-use workflow.
04 / Pricing
Claude Sonnet 4.6 Pricing
Claude Sonnet 4.6 costs $3 per million input tokens and $15 per million output tokens, with prompt-caching rates and a 50% Batch API discount.
The Classic Sonnet Price Point
Anthropic lists 5-minute cache writes at $3.75 per million tokens, 1-hour cache writes at $6 per million, and cache reads at $0.30 per million.
These rates make caching particularly relevant for workflows that repeatedly reuse large stable prefixes such as system instructions, repositories, reference documentation, policies, or long research context.
The Batch API applies a 50% discount to standard input and output rates, making it useful for offline evaluations and bulk processing.
Current Sonnet 5.5 lowers the base rate to $2 input and $10 output per million tokens. A migration therefore changes both model behavior and nominal economics.
However, Sonnet 5 and 5.5 also use a different tokenizer than Sonnet 4.6. Anthropic warns that the same text can produce about 30% more tokens on newer Sonnet models. Effective savings should therefore be measured with real workload token counts rather than estimated from headline rates alone.
1M tokens · USD
- Input
- $3.00
- 5m cache write
- $3.75
- 1h cache write
- $6.00
- Cache read
- $0.30
- Output
- $15.00
Example: 20K input + 4K output
- Input cost
- $0.0600
- Output cost
- $0.0600
- Estimated uncached total
- $0.1200
05 / Context
1M Context and 128K Standard Output
Claude Sonnet 4.6 supports a 1,000,000-token context window and up to 128,000 output tokens in standard requests.
A Large-Context Sonnet Before Claude 5
When Sonnet 4.6 launched, Anthropic described its 1M context window as beta. The model later moved to the full 1M context profile documented today.
That capacity makes it suitable for large repositories, long document collections, extensive conversation histories, tool results, retrieved evidence, and persistent agent state.
Standard output is capped at 128K tokens. Anthropic also documents up to 300K output tokens through Message Batches in beta.
The context window is large enough that poor context management can become an economic problem before it becomes a technical limit. Sending irrelevant state repeatedly costs money and can reduce the prominence of the instructions or evidence that matter most.
Prompt caching, retrieval, summarization, and selective history retention remain useful even when the model can technically accept a million tokens.
Context window
1,000,000
Standard max output
128,000
Message Batches can expose up to 300K output tokens in beta. Standard requests use the 128K maximum output limit.
06 / Vision
Vision Before the High-Resolution Claude Generation
Claude Sonnet 4.6 accepts image input, but it belongs to the pre-high-resolution Claude vision generation and uses a lower image-resolution and visual-token ceiling than newer Sonnet models.
Image-Heavy Workloads Need Re-Baselining During Migration
Anthropic’s Sonnet 5.5 migration guide documents the difference directly.
Sonnet 4.6 processes images up to 1,568 pixels on the long edge and up to roughly 1,568 visual tokens per image.
Newer Sonnet 5.5 uses the high-resolution tier, supporting up to 2,576 pixels on the long edge and up to 4,784 visual tokens per image.
That is not merely a quality change. It can materially change cost. Anthropic gives an example where a 2000×1500 image can cost about 2.5 times as many visual tokens on Sonnet 5.5.
For UI analysis, documents, diagrams, and computer-use workflows, evaluate whether the additional visual fidelity improves task success enough to justify the token increase. Downsample images when the extra detail is unnecessary.
- 01
1,568 px long edge
Sonnet 4.6 belongs to the lower-resolution image-processing generation documented by Anthropic.
- 02
Lower visual-token ceiling
Images consume fewer visual tokens than on Sonnet 5.5, which changes the economics of screenshot- and document-heavy workloads.
- 03
Migration can increase image cost
Higher-resolution successor models may spend more tokens even when the source image is unchanged.
- 04
Downsample deliberately
Use the resolution required by the task rather than automatically sending every image at maximum fidelity.
07 / API behavior
API Behaviors That Distinguish Sonnet 4.6
Sonnet 4.6 keeps several pre-Claude-5 behaviors that later Sonnet generations change, making it an important compatibility baseline for older production integrations.
Stable Dateless ID and Legacy Reasoning Compatibility
The model ID claude-sonnet-4-6 is a pinned snapshot, not an evergreen alias.
Anthropic changed its versioning model starting with the 4.6 generation: a dateless ID identifies a fixed release rather than silently moving to newer weights.
Sonnet 4.6 also rejects assistant-message prefilling. This matters when migrating from Sonnet 4.5 or older applications that shaped output by pre-populating the final assistant turn.
Thinking-related sampling rules are less restrictive than on Claude 5, but thinking itself imposes constraints. When thinking is active, temperature must remain at its default value.
Sonnet 5.5 migration guidance also requires removing non-default temperature, top_p, and top_k values from older Sonnet integrations.
Finally, Sonnet 4.6 preserves older tool-orchestration patterns that later Sonnet models change, so migration testing should include the complete request body rather than just changing the model string.
- 01
Pinned dateless model ID
claude-sonnet-4-6 identifies a fixed model snapshot even though its ID contains no date.
- 02
No assistant prefill
A prefilled final assistant message is not supported on Sonnet 4.6.
- 03
Legacy thinking mode still accepted
Manual enabled thinking with budget_tokens remains available only for backward compatibility.
- 04
Migration changes sampling and tools
Newer Sonnet models require retesting non-default sampling parameters, tool forcing, and conversation-state behavior.
08 / Migration
Migrating From Claude Sonnet 4.6 to Sonnet 5.5
The move from Sonnet 4.6 to Sonnet 5.5 changes thinking defaults, tokenization, effort calibration, visual-token economics, tool orchestration, prompt caching, and pricing.
Treat the Upgrade as a Compatibility and Evaluation Project
Sonnet 4.6 runs without thinking when the field is omitted. Sonnet 5.5 runs with adaptive thinking by default. The old manual thinking: enabled plus budget_tokens format is rejected by Sonnet 5.5.
Effort levels also need a fresh evaluation. Sonnet 4.6 supports low, medium, high, and max; Sonnet 5.5 adds xhigh, and Anthropic notes that effort levels are recalibrated.
Tokenization changes significantly. Anthropic states that identical text produces about 30% more tokens on newer Sonnet models compared with Sonnet 4.6, depending on content.
Image processing moves to a higher-resolution tier, which can substantially increase visual token usage.
The newer model also lowers the prompt-cache minimum and changes forced-tool and conversation-state behavior. At the same time, base token pricing falls from $3/$15 on Sonnet 4.6 to $2/$10 on Sonnet 5.5.
- 01
Replace thinking budgets
Move deprecated budget_tokens requests to adaptive thinking and effort, and account for the fact that thinking is on by default on Sonnet 5.5.
- 02
Recount text and images
Recalculate token usage, max_tokens, context thresholds, image cost, and monthly estimates under the newer tokenizer and vision tier.
- 03
Retest tool orchestration
Review forced tool selection, conversation history, thinking blocks, computer use, and newer dynamic agent controls.
- 04
Re-baseline economics
Compare completed-task cost using real token counts, cache behavior, retries, latency, and acceptance rate rather than headline prices alone.
09 / Capabilities
Claude Sonnet 4.6 Capabilities
Claude Sonnet 4.6 combines text and image input, a 1M context window, adaptive reasoning, legacy extended thinking, tool workflows, prompt caching, and large Batch output in a mature pre-Claude-5 model.
A Broad Production Feature Set Before the Claude 5 Transition
The model accepts text and images and returns text. Adaptive thinking is the recommended reasoning path, while deprecated manual extended thinking remains available for legacy code. Effort can tune the overall amount of work the model performs even when thinking is not active.
Prompt caching lowers repeated-context cost. Batch processing cuts standard input/output rates by 50% and can expose up to 300K output tokens in beta.
Sonnet 4.6 is available across the Claude API, Amazon Bedrock, Google Cloud, Microsoft Foundry, and Claude Platform on AWS.
Its primary limitation today is generation age: newer Sonnet models improve performance, lower token pricing, change reasoning defaults, and add more advanced orchestration controls.
- Supported
Text
Accept text input and produce text output.
- Supported
Image input
Analyze supported images using the pre-high-resolution Claude vision tier.
- Supported
Adaptive thinking
Recommended reasoning mode controlled with effort.
- Supported
Legacy extended thinking
Manual budget_tokens still works but is deprecated.
- Supported
Tool workflows
Supports tool-oriented agents, search, code execution, and broader Claude platform workflows.
- Supported
Prompt caching
Reduce repeated processing cost for stable prompt prefixes.
- Supported
Batch API
Provides a 50% discount and a beta path to 300K output tokens.
10 / Evaluation
Claude Sonnet 4.6 Strengths and Limitations
Sonnet 4.6 remains a useful compatibility and regression baseline because it combines modern-scale context with a distinctly pre-Claude-5 reasoning, tokenizer, vision, and tool contract.
Strengths
1M context and 128K output
Large working capacity supports repositories, long documents, research state, tool history, and substantial generated artifacts.
Adaptive-thinking transition point
Applications can use modern adaptive thinking while still retaining deprecated budget_tokens compatibility during migration.
Balanced agentic baseline
Anthropic launched Sonnet 4.6 around balanced speed and intelligence and highlighted improved agentic search with lower token consumption.
Mature production ecosystem
Caching, Batch API, multimodal input, tools, search, and broad cloud availability make it representative of mature pre-Claude-5 deployments.
What to consider
Legacy status
Anthropic recommends migrating to Claude Sonnet 5.5 for improved performance.
Higher base price than current Sonnet
$3/$15 standard pricing is higher than Sonnet 5.5 at $2/$10.
Older vision tier
Image resolution and visual-token limits are lower than on Sonnet 5.5, so visual quality and migration economics differ.
Older thinking and tokenizer contract
Thinking is off by default, manual budgets are still accepted, and newer Sonnet models use different tokenization and orchestration behavior.
Measure the migration, not just the model name
Compare Claude Sonnet 4.6 with current Sonnet models
Replay coding, search, tool-use, vision, cached-context, and reasoning-heavy workloads to quantify how newer Sonnet models change task quality, tokenization, latency, thinking behavior, and total cost.
Start FreeAnthropic lists Sonnet 4.6 as Active (legacy) and recommends migrating to Claude Sonnet 5.5 for improved performance.
Common Questions
What is Claude Sonnet 4.6?
Claude Sonnet 4.6 is an Anthropic legacy Sonnet model released on February 17, 2026. It combines a 1M context window, 128K standard output, adaptive thinking, deprecated extended-thinking compatibility, image input, tool use, and prompt caching.
How much does Claude Sonnet 4.6 cost?
Claude Sonnet 4.6 costs $3 per 1M input tokens and $15 per 1M output tokens. Five-minute cache writes cost $3.75/MTok, one-hour cache writes cost $6/MTok, and cache reads cost $0.30/MTok. Batch input and output receive a 50% discount.
What is the Claude Sonnet 4.6 context window?
Claude Sonnet 4.6 has a 1,000,000-token context window and supports up to 128,000 output tokens in standard requests.
Can Claude Sonnet 4.6 output more than 128K tokens?
Anthropic documents a beta maximum of up to 300,000 output tokens through the Message Batches API. Standard requests use the 128K output limit.
What is Claude Sonnet 4.6’s knowledge cutoff?
Anthropic lists August 2025 as the reliable knowledge cutoff and January 2026 as the training-data cutoff.
Does Claude Sonnet 4.6 support images?
Yes. Sonnet 4.6 accepts image input and text output. Anthropic’s migration guide places it in the older image tier with up to 1,568 pixels on the long edge, below the higher-resolution tier used by Sonnet 5.5.
Does Claude Sonnet 4.6 support adaptive thinking?
Yes. Adaptive thinking is the recommended reasoning mode. Sonnet 4.6 also still accepts the older thinking enabled plus budget_tokens format, but Anthropic marks that mode deprecated.
Is thinking on by default in Claude Sonnet 4.6?
No. A Sonnet 4.6 request that omits the thinking field runs without thinking. This differs from Sonnet 5 and Sonnet 5.5, where adaptive thinking is on by default.
What effort levels does Claude Sonnet 4.6 support?
Anthropic documents low, medium, high, and max effort for Sonnet 4.6. High is the API default, while current guidance recommends medium as a practical starting point for many agentic coding and tool-heavy applications. Sonnet 4.6 does not support xhigh.
Is claude-sonnet-4-6 an alias?
No. Anthropic’s versioning documentation states that dateless model IDs from the 4.6 generation onward are pinned model snapshots, not evergreen aliases.
How is Claude Sonnet 4.6 different from Claude Sonnet 5?
Sonnet 4.6 runs without thinking by default, still accepts deprecated manual budget_tokens thinking, uses the older tokenizer, and costs $3/$15. Sonnet 5 turns adaptive thinking on by default, removes manual thinking budgets, introduces a tokenizer that uses roughly 30% more tokens for the same text, and lowers standard pricing to $2/$10.
How is Claude Sonnet 4.6 different from Claude Sonnet 5.5?
Sonnet 5.5 keeps the 1M/128K capacity but uses a newer tokenizer and high-resolution vision tier, turns adaptive thinking on by default, adds xhigh effort and newer agent controls, changes forced-tool behavior, and costs $2/$10 instead of Sonnet 4.6’s $3/$15.
Is Claude Sonnet 4.6 still available?
Yes. Anthropic lists Claude Sonnet 4.6 as Active (legacy). It was released February 17, 2026 and is not scheduled for retirement sooner than February 17, 2027.
Should I use Claude Sonnet 4.6 for a new application?
For new development, Anthropic recommends evaluating Claude Sonnet 5.5. Sonnet 4.6 is most useful for existing deployments, migration testing, pre-Claude-5 compatibility, historical evaluations, and workloads that still depend on its older reasoning or tokenization behavior.
Model information
Last updated
Specifications, pricing, lifecycle, thinking behavior, effort controls, context limits, vision differences, tokenization changes, and migration guidance on this page are based on Anthropic’s official Claude Sonnet 4.6 model documentation, effort documentation, release notes, thinking documentation, model-versioning guide, and Sonnet 5.5 migration guide.