Claude Sonnet

Claude Sonnet 4.5

by Claude · Legacy; still available as of 2026-09-24; migration to Claude Sonnet 5 recommended

Claude Sonnet 4.5 is Anthropic’s coding- and agent-focused model for complex reasoning, software development, computer use, research, and visual document analysis. It accepts text and images, returns text, supports extended thinking and tools, and offers a 200,000-token context window with up to 64,000 output tokens. API pricing is $3 per million input tokens and $15 per million output tokens. The model remains available but is classified as legacy, with migration to Claude Sonnet 5 recommended for new deployments.

Text Reasoning Coding
Claude Sonnet 4.5 is a high-capability model from Anthropic designed especially for software development, coding agents, computer-use workflows, research, and other multi-step tasks. It can analyze text and images, reason through complex problems, use tools supplied by an application, and return structured text or JSON-like results. The model is still available through Anthropic’s developer platform and partner cloud services, but its legacy status means new projects should compare it carefully with Anthropic’s newer Sonnet models.
Outputs

What Claude Sonnet 4.5 can produce

Text
Inputs

What it can understand

Text Images Multimodal input
Capabilities

Supported features

Tool use Web search Streaming Structured output Prompt caching Batch API
Model profile

Performance characteristics

9/10 Reasoning
10/10 Coding
8/10 Speed
6/10 Cost efficiency
Specifications

Technical details

Model family Claude Sonnet
Model type Coding
Context window 200K tokens
Maximum output 64K tokens
Knowledge cutoff January 2025
Release date 2025-09-29
Status Legacy; still available as of 2026-09-24; migration to Claude Sonnet 5 recommended
Knowledge cutoff notes

Anthropic lists January 2025 as the reliable knowledge cutoff for Claude Sonnet 4.5. Web search and other retrieval tools can provide newer information during use but do not change the underlying model cutoff.

Model notes

The canonical dated API model ID is claude-sonnet-4-5-20250929. The claude-sonnet-4-5 alias points to the most recent dated snapshot for this minor version. The model supports extended thinking, text and image input, text output, client and server tools, structured outputs, prompt caching, streaming, and Message Batches. Anthropic's documentation classifies it as legacy and recommends migration to Claude Sonnet 5. The lifecycle page states retirement will occur no sooner than September 29, 2026, but does not provide an exact shutdown date. The reliable knowledge cutoff is January 2025. JSON mode is left unknown because structured outputs are documented separately and do not automatically establish a distinct legacy JSON-mode capability.

Cost

Model pricing

Input $3 per million tokens; cache write $3.75 per million tokens for 5 minutes or $6 per million tokens for 1 hour; cache read $0.30 per million tokens
Output $15 per million tokens; Message Batches API offers a 50% discount on input and output pricing
Model guide

Claude Sonnet 4.5: Anthropic’s Coding and Agent-Focused Model

Claude Sonnet 4.5 is Anthropic’s Claude 4.5 model for coding, long-running agents, computer use, research, and complex professional workflows. It accepts text and images, produces text, supports extended thinking, tool use, structured outputs, prompt caching, and batch processing, and provides a 200,000-token context window with up to 64,000 output tokens. It remains available but is classified as legacy, so Anthropic recommends evaluating Claude Sonnet 5 for new deployments.

What is Claude Sonnet 4.5?

Claude Sonnet 4.5 is Anthropic’s Claude 4.5 Sonnet model, released on September 29, 2025. Anthropic positioned it for coding, complex agents, computer use, reasoning, and long-horizon tasks. In practical terms, it is intended for work that requires more than a single short answer: maintaining context across a large repository, planning and executing several development steps, reviewing visual material, or coordinating actions through external tools.

The model belongs to Anthropic’s Claude family and is accessed primarily through the Claude API and supported cloud platforms. It is not a standalone image, audio, or video generator. Its multimodal capability refers to its ability to accept both text and images while returning text.

Anthropic currently classifies Claude Sonnet 4.5 as a legacy model and recommends migration to Claude Sonnet 5. It remains available according to the supplied lifecycle information, with retirement scheduled no sooner than September 29, 2026; Anthropic has not published an exact shutdown date.

Key specifications at a glance

SpecificationClaude Sonnet 4.5
ProviderAnthropic
Release dateSeptember 29, 2025
Canonical API model IDclaude-sonnet-4-5-20250929
Context window200,000 tokens
Maximum output64,000 tokens
InputText and images
OutputText, including structured responses
ReasoningExtended thinking supported
StatusLegacy, but still available according to the supplied research

A token is a unit of text used for model processing; it may represent a word, part of a word, punctuation, or other text fragment. The 200,000-token context window covers the material the model can consider in a request and conversation, while the 64,000-token maximum applies to the response it can produce in one request.

Why it is suited to coding and agent work

Claude Sonnet 4.5’s main distinction is its focus on coding and long-running agentic workflows. A coding agent is an application that combines a language model with tools such as a file system, terminal, browser, issue tracker, or source-control workflow. Sonnet 4.5 can help inspect a repository, identify relevant files, propose a plan, write or modify code, interpret test output, and continue through several related steps.

Its large context window is useful when a task involves extensive source code, technical documentation, logs, or design material. The model can also analyze screenshots, diagrams, interfaces, and visual PDF content when those are supplied as supported image input. This makes it suitable for tasks such as explaining an interface screenshot, reviewing a visual error report, or combining a written specification with a diagram.

Tool use does not mean that the base model independently operates a computer or gains unrestricted access to external systems. An application must provide client-side tools or enable compatible Anthropic server tools. Developers remain responsible for permissions, authentication, sandboxing, confirmation steps, and validation of any action that could modify data or affect an external system.

Reasoning, tools, and structured output

Claude Sonnet 4.5 supports extended thinking, which allows the model to spend additional processing effort on difficult problems before producing its answer. This is useful for multi-step coding tasks, debugging, planning, research synthesis, and problems where a quick response is more likely to miss dependencies or constraints. Extended thinking can increase the amount of processing required, so it should be enabled selectively when the additional reasoning is worth the cost or latency.

The model supports client-side and server-side tool use, streaming responses, structured outputs, strict tool use, prompt caching, and Anthropic’s Message Batches API. Structured outputs are useful when an application needs predictable fields rather than free-form prose, such as extracting an issue list, returning a test plan, or classifying documents. The supplied research does not verify a separate legacy “JSON mode” capability, so structured outputs should not automatically be treated as the same feature.

Prompt caching can reduce repeated processing for stable instructions or large prefixes that are reused across requests. The Message Batches API offers a 50% discount on input and output token pricing, according to the supplied pricing information, but batch processing is intended for workloads that do not require an immediate individual response.

Supported modalities and practical limits

  • Text input: Supported.
  • Image input: Supported for visual analysis alongside text.
  • Text output: Supported, including code and structured responses.
  • Image, audio, and video output: Not natively supported.
  • Context: Up to 200,000 tokens.
  • Maximum response: Up to 64,000 output tokens.

These limits make Sonnet 4.5 appropriate for substantial documents, codebases, and multi-step conversations, but they do not guarantee that every task will fit comfortably in one request. Large inputs still need careful organization, and application developers must account for tool results, instructions, conversation history, and the requested response within the available context.

API pricing

Anthropic lists Claude Sonnet 4.5 at $3 per million input tokens and $15 per million output tokens. Input and output are priced separately, so an application that generates long responses can spend considerably more on output than on the corresponding input.

Usage typePrice
Input tokens$3 per million tokens
Output tokens$15 per million tokens
Five-minute prompt-cache write$3.75 per million tokens
One-hour prompt-cache write$6 per million tokens
Prompt-cache read$0.30 per million tokens
Message Batches API50% discount on input and output token pricing

Prompt caching is most relevant when the same large instructions or reference material are sent repeatedly. The cache-write rates are higher than the normal input rate, while cache reads are substantially cheaper. The right choice therefore depends on how often the cached material is reused and how long it remains valid.

Strengths and trade-offs

The supplied research rates Sonnet 4.5 highly for coding and complex reasoning, but those scores are editorial evaluations rather than Anthropic-published specifications. The concrete strengths supported by the documentation are its 200,000-token context, 64,000-token output limit, image understanding, extended thinking, tool support, structured outputs, prompt caching, and batch processing.

Its principal trade-off is cost and latency relative to smaller or simpler models. The model is designed for difficult tasks rather than the lowest-cost response to every request. Extended thinking, large tool traces, and long generated outputs can increase both processing time and token usage. Applications that only need short classification, simple extraction, or brief routine replies may not benefit from using a model aimed at complex agent workflows.

Sonnet 4.5 also cannot replace a native media-generation model. It can analyze images, but it does not natively create images, audio, or video. Its context window is large but does not exceed 200,000 tokens, so tasks requiring a larger working context may need a different option or a retrieval and summarization strategy.

Best use cases

  • Repository-scale coding, refactoring, debugging, and code review.
  • Software agents that need to inspect files, run tools, and maintain a multi-step plan.
  • Computer-use or browser workflows where the application controls permissions and actions.
  • Technical research and document analysis involving substantial written material.
  • Visual analysis of screenshots, diagrams, interfaces, and supported PDF content.
  • Applications requiring structured responses or strict tool invocation.
  • Long-running professional workflows where consistency across many steps matters.

For example, a development assistant could receive a feature request, inspect relevant files through application-provided tools, propose an implementation plan, modify code, interpret test failures, and return a structured summary of the changes. Human review and execution safeguards remain important, especially when tools can change production systems or access sensitive data.

When to choose Claude Sonnet 4.5

Choose Claude Sonnet 4.5 when the workload benefits from strong coding and reasoning performance, image understanding, a large context window, tool use, and sustained multi-step interaction. It is a reasonable fit for teams that need a capable general model for software agents and complex analysis and that can accept its higher output price compared with lighter models.

Consider another option when the primary requirement is the lowest latency or lowest cost, when the task needs native image, audio, or video generation, or when a context window larger than 200,000 tokens is essential. New deployments should also evaluate Claude Sonnet 5 because Anthropic recommends migration to that newer Sonnet model. Sonnet 4.5 may still be appropriate where an existing application depends on its behavior, pricing, integrations, or validated workflow, but its legacy status makes lifecycle planning important.

Availability and model identity

The dated API identifier for the model is claude-sonnet-4-5-20250929. Anthropic also documents claude-sonnet-4-5 as a convenience alias that points to the most recent dated snapshot for this minor version. Using the dated identifier can make deployments more reproducible, while an alias may follow Anthropic’s designated snapshot for the model family.

Sonnet 4.5 is available through Anthropic’s developer platform and partner cloud services according to the supplied research. Because it is classified as legacy and has a retirement window beginning no sooner than September 29, 2026, developers should check Anthropic’s lifecycle documentation before committing it to a new long-lived system.


Answers to Frequently Asked Questions

Is Claude Sonnet 4.5 still available, and should developers use it for new projects?
Claude Sonnet 4.5 is classified as a legacy model but remains available according to the supplied lifecycle information, with retirement scheduled no sooner than September 29, 2026. Anthropic recommends migrating to Claude Sonnet 5, so developers starting new long-lived projects should evaluate the newer model and review lifecycle documentation.
What can Claude Sonnet 4.5 be used for?
Claude Sonnet 4.5 is suitable for repository-scale coding, debugging, refactoring, code review, technical research, document analysis, screenshot and diagram interpretation, structured data extraction, and software agents that use files, terminals, browsers, or other application-provided tools.
How much does Claude Sonnet 4.5 cost?
Anthropic lists Claude Sonnet 4.5 at $3 per million input tokens and $15 per million output tokens. Prompt-cache writes cost $3.75 per million tokens for five-minute caching or $6 per million tokens for one-hour caching, while cache reads cost $0.30 per million tokens. The Message Batches API provides a 50% discount on input and output token pricing.
What is Claude Sonnet 4.5?
Claude Sonnet 4.5 is Anthropic’s model for coding, complex reasoning, software agents, computer use, visual analysis, and long-running multi-step tasks. It accepts text and images and produces text, including code and structured responses.
What are Claude Sonnet 4.5’s context window and output limits?
Claude Sonnet 4.5 supports a context window of up to 200,000 tokens and can generate up to 64,000 output tokens in a single request.


Sources 9
Provider

About Claude