This model is deprecated
Claude Sonnet 4 was deprecated on Jun 15, 2026 and can no longer be run here. Its evaluation results and details remain available for reference. Try Claude Sonnet 5 instead.
Claude 4 Sonnet, released by Anthropic in May 2025, is the mid-tier model in the Claude 4 family, designed to balance capability, cost, and speed. It is multimodal, accepting both text and images, and extends beyond prior versions with improved “computer use” support, allowing API-driven interaction with desktop-like interfaces. By default, it supports 200,000 tokens of context, but as of August 2025, it also offers a 1 million-token context window in public beta—making it one of the most context-capable models available for processing entire codebases or large document sets in a single request.
Sonnet 4 is significantly cheaper than the flagship Opus while still demonstrating strong reasoning, coding, and instruction-following ability with reduced hallucinations. Its extended context capabilities and lower latency make it well-suited for enterprise-scale knowledge management, software development, research assistants, and productivity automation where both cost efficiency and high reliability are essential.
—
Usage
Past 30 DaysNot available
Not in Playground
Claude Sonnet 4 has been deprecated by its provider and can no longer be evaluated on the current benchmark. The legacy Vision Evals results below are preserved for reference. See the current Vision Evals
| Category | Passed | Score |
|---|---|---|
| Document Understanding | 8 / 9 | 88.9% |
| Defect Detection | 12 / 15 | 80% |
| Object Understanding | 11 / 14 | 78.6% |
| Spatial Understanding | 13 / 19 | 68.4% |
| Object Counting | 2 / 10 | 20% |
Scores based on a single evaluation run · Methodology
View all legacy Vision Evals results →Claude Sonnet 4 costs $3.00 per 1M input tokens and $15.00 per 1M output tokens.
Pricing updated Aug 5, 2026
Estimated cost per task vs. Visual Understanding score, for this model and others ranked near it. Upper-left is the sweet spot (high quality, low cost). Based on Vision Evals (legacy) results.
10 of 11 models plotted · 1 not yet evaluated
| Model | Score | Median tokens | Est. cost / task | Compare |
|---|---|---|---|---|
| Claude Sonnet 5 | 70.2% | 2.2K | $0.0048 | Compare |
| Claude Sonnet 4.6 | 70.2% | 2.3K | $0.0080 | Compare |
| GPT-5.6 Luna | 70.2% | 1.5K | $0.0002 | Compare |
| Gemini 2.5 Pro | 70.2% | 856 | $0.0060 | Compare |
| Gemini 3.1 Flash-Lite | 68.7% | 1.1K | $0.0003 | Compare |
| Claude Sonnet 4(this model) | 68.7% | — | — | — |
| Gemma 4 26B A4B | 68.7% | 531 | $0.0001 | Compare |
| Qwen3.6 Plus | 68.7% | 1.6K | $0.0005 | Compare |
| Claude Opus 4.8 | 67.2% | 2.2K | $0.012 | Compare |
| Claude Opus 4.7 | 67.2% | 2.6K | $0.015 | Compare |
| Gemma 4 31B | 67.2% | 467 | $0.0001 | Compare |
Other models worth comparing for similar use cases.
Other versions in the same family as Claude Sonnet 4.
License terms and commercial-use guidance for Claude Sonnet 4.
This model is proprietary. The author retains all rights, and use of the model is governed by their specific terms of service or license agreement.
Commercial use depends on the terms set by the model author. Most proprietary commercial models require a paid subscription, API key, or per-call billing. Check the provider’s pricing and terms-of-service for details.
License information is provided as a guide and is not legal advice.
Yes. Claude Sonnet 4 accepts image input, and on Roboflow's previous vision benchmark it passed 68.7% of visual understanding tasks (#25 of 77).
Claude Sonnet 4 has been deprecated by its provider and can no longer be run, so it is not part of Roboflow's current Vision Evals. Its results from the previous benchmark are preserved on this page for reference.