ModelRefs / Claude Sonnet 4 by Anthropic — Benchmarks, Pricing & Review…
Claude Sonnet 4 by Anthropic — Benchmarks, Pricing & Review…
Claude Sonnet 4 (Anthropic): Claude Sonnet 4 by Anthropic.. 200K context. Pricing: from $0.00300/1K in. Specs, benchmarks and code examples.
What this reference supports
Claude Sonnet 4 is an Anthropic Claude 4 model for coding, reasoning, tool use, and general language applications, positioned as a mid-tier option in Anthropic's Claude 4 generation between Opus and smaller Haiku-class models, intended for coding assistants and agent workflows rather than the highest-complexity reasoning tier or the lowest-latency tier.
Use this page to compare Claude Sonnet 4's context window, pricing, and modalities against other Claude models and competing providers, and to check which coding and agent benchmarks — such as SWE-bench and OSWorld — apply to this specific version rather than a newer Claude release.
Because later Sonnet generations exist, this is a version-specific reference rather than a recommendation to start a new deployment on this model. Pin an explicit model identifier, confirm current channel availability across Anthropic and supported cloud partners, and evaluate migration paths to a current-generation model before committing, especially for long-running agent deployments that are costly to re-validate.
Benchmark & Evaluation
ModelRefs currently has partial, narrow benchmark coverage for Claude Sonnet 4. Treat the available benchmark evidence as one input to the decision, not a guarantee that Claude Sonnet 4 is the strongest option for your workload, and evaluate it on representative workloads before selecting it.
- Provider-reported benchmark results should be interpreted with methodology, dataset, prompting, tool, sampling, and recency limitations in mind.
- The Claude 4 release and system-card index document provider-run evaluation context and limitations.
Implementation considerations
- Use an explicit model identifier and evaluate migration paths.
- Test extended thinking, tool schemas, prompt caching, output controls, and task-specific error modes.
- Originally available through Anthropic and supported cloud channels.
- Channel-specific model IDs, regions, features, and lifecycle dates must be checked.
Risks and limitations
- Outputs can be incorrect or unsuitable for the intended task; use task-specific evaluation, grounding, and human review where consequences are material.
- API availability, model aliases, rate limits, data controls, regions, and prices are mutable and differ by product channel.
Source coverage
This reference is Provisional. Model behavior, access, pricing, limits, and lifecycle can change; verify the linked provider documentation and run task-specific evaluations before implementation.
Known coverage gaps:
- Current model-retirement details need channel-specific confirmation.
- Independent evaluations are not attached at claim level.
Sources
Continue your research
Use these connected ModelRefs sections to compare alternatives, inspect implementation paths, and review the evidence and governance boundaries relevant to Claude Sonnet 4 by Anthropic — Benchmarks, Pricing & Review….