ModelRefs / Claude 3.7 Sonnet by Anthropic — Benchmarks, Pricing & Revi…
Claude 3.7 Sonnet by Anthropic — Benchmarks, Pricing & Revi…
Claude 3.7 Sonnet (Anthropic): Claude 3.7 Sonnet by Anthropic.. 200K context. Pricing: from $0.00300/1K in. Specs, benchmarks and code examples.
What this reference supports
Claude 3.7 Sonnet is an Anthropic hosted model introducing extended-thinking support and strengthened agentic coding in the Claude 3.7 line. Extended-thinking budgets change cost and latency and should be tuned per task; verify current feature and snapshot availability for your API tier before relying on parity with other surfaces.
Claude 3.7 Sonnet is attributed to Anthropic in ModelRefs' canonical registry. Tracked modalities: Text input and output, Image input. Primary use cases considered on ModelRefs: Repository-scale coding assistants and agentic coding loops; Tool-heavy workflows where configurable extended thinking is useful.
This ModelRefs profile is Provisional and pending review — decision-support material, not a final or universal ranking. Confirm current behavior, access, pricing, limits, licensing, and lifecycle in Anthropic's own documentation, and evaluate Claude 3.7 Sonnet on representative workloads before implementation.
Benchmark & Evaluation
ModelRefs currently has partial, narrow benchmark coverage for Claude 3.7 Sonnet. Treat the available benchmark evidence as one input to the decision, not a guarantee that Claude 3.7 Sonnet is the strongest option for your workload, and evaluate it on representative workloads before selecting it.
- No benchmark score is imported into this editorial record. Canonical benchmark runs and scores are governed separately with their own provenance and render only through those records; coverage in ModelRefs is currently narrow (partial), so any scored comparison must show its coverage limits.
- ModelRefs holds canonical run evidence on software-engineering evaluation (SWE-bench); coverage is narrow and provider-reported evaluations are not independently reproduced here.
Implementation considerations
- Decide extended-thinking budgets explicitly; thinking tokens change cost and latency and should be tuned per task.
- Pin a dated snapshot and verify computer-use / tool-use feature availability for your API tier and region.
- Hosted through the Anthropic API and eligible cloud partners (per Anthropic's current documentation).
- Feature flags, rate limits, and data-retention controls differ by platform and account; check the models overview for the current matrix.
Risks and limitations
- Hosted-model behavior, quotas, pricing, and data controls can change without a client-side version pin unless a dated snapshot is used.
- Provider-reported capabilities require task-specific evaluation before production reliance.
Source coverage
This reference is Provisional. Model behavior, access, pricing, limits, licensing, and lifecycle can change; verify the linked provider documentation and run task-specific evaluations before implementation.
Known coverage gaps:
- Independent reproduction of provider-reported coding/agentic evaluations is not attached.
- Extended-thinking cost-quality guidance needs task-level evidence.
Sources
Continue your research
Use these connected ModelRefs sections to compare alternatives, inspect implementation paths, and review the evidence and governance boundaries relevant to Claude 3.7 Sonnet by Anthropic — Benchmarks, Pricing & Revi….