This model is deprecated
Claude 3.7 Sonnet and can no longer be run here. Its evaluation results and details remain available for reference. Try Claude Sonnet 5 instead.
Claude 3.7 Sonnet, released by Anthropic in February 2025, is the company’s first hybrid reasoning model, combining fast response generation with an optional “extended thinking mode” that reveals longer, step-by-step reasoning. Like its predecessors, it is multimodal, handling both text and images, but expands its usability with up to 200,000 input tokens and up to 128,000 output tokens (64K generally available, 128K in beta). This makes it well-suited for analyzing large documents, codebases, or multi-turn conversations.
Typical applications include software development, research workflows, extended reasoning tasks, and enterprise-scale knowledge work where a trade-off between speed and visible reasoning is valuable.
—
Usage
Past 30 DaysNot available
Not in Playground
Claude 3.7 Sonnet has been deprecated by its provider and can no longer be evaluated on the current benchmark. The legacy Vision Evals results below are preserved for reference. See the current Vision Evals
| Category | Passed | Score |
|---|---|---|
| Document Understanding | 8 / 9 | 88.9% |
| Defect Detection | 10 / 15 | 66.7% |
| Object Understanding | 9 / 14 | 64.3% |
| Spatial Understanding | 11 / 19 | 57.9% |
| Object Counting | 2 / 10 | 20% |
Scores based on a single evaluation run · Methodology
View all legacy Vision Evals results →Estimated cost per task vs. Visual Understanding score, for this model and others ranked near it. Upper-left is the sweet spot (high quality, low cost). Based on Vision Evals (legacy) results.
10 of 11 models plotted · 1 not yet evaluated
| Model | Score | Median tokens | Est. cost / task | Compare |
|---|---|---|---|---|
| Claude Opus 4.6 | 64.2% | 2.3K | $0.014 | Compare |
| GPT-5.4 Nano | 62.7% | 1.8K | $0.0004 | Compare |
| Llama 4 Maverick | 59.7% | 2.4K | $0.0005 | Compare |
| Claude Sonnet 4.5 | 59.7% | 2.3K | $0.0092 | Compare |
| Claude Opus 4.1 | 59.7% | 2.1K | $0.040 | Compare |
| Claude 3.7 Sonnet(this model) | 59.7% | — | — | — |
| Claude Haiku 4.5 | 58.2% | 2.3K | $0.0030 | Compare |
| GPT-5 Nano | 58.2% | 2.7K | $0.0003 | Compare |
| Qwen3.5 397B A17B | 58.2% | 1.5K | $0.0006 | Compare |
| Gemini 2.5 Flash | 55.2% | 476 | $0.0005 | Compare |
| Gemini 2.5 Flash-Lite | 53.7% | 301 | <$0.0001 | Compare |
Other models worth comparing for similar use cases.
Other versions in the same family as Claude 3.7 Sonnet.
License terms and commercial-use guidance for Claude 3.7 Sonnet.
This model is proprietary. The author retains all rights, and use of the model is governed by their specific terms of service or license agreement.
Commercial use depends on the terms set by the model author. Most proprietary commercial models require a paid subscription, API key, or per-call billing. Check the provider’s pricing and terms-of-service for details.
License information is provided as a guide and is not legal advice.
Yes. Claude 3.7 Sonnet accepts image input, and on Roboflow's previous vision benchmark it passed 59.7% of visual understanding tasks (#46 of 77).
Claude 3.7 Sonnet has been deprecated by its provider and can no longer be run, so it is not part of Roboflow's current Vision Evals. Its results from the previous benchmark are preserved on this page for reference.