This model is deprecated
Claude Opus 4.1 was deprecated on Aug 5, 2026 and can no longer be run here. Its evaluation results and details remain available for reference. Try Claude Opus 5 instead.
Claude 4.1 Opus, released by Anthropic in August 2025, is the upgraded flagship of the Claude 4 family, building on Opus 4 with stronger reasoning and agentic capabilities. Like its predecessor, it is multimodal and optimized for text, code, and tool use, with support for large context windows suited to multi-file codebases, technical workflows, and long-horizon problem solving.
On benchmarks, Opus 4.1 improves coding performance, reaching ~74.5% on SWE-Bench Verified compared to Opus 4’s ~72.5%. It demonstrates more precise debugging, refactoring, and orchestration of agentic tasks while maintaining similar safety and alignment safeguards. It is best suited for enterprise-scale software development, research automation, and advanced reasoning workflows where reliability and depth of analysis are critical.
—
Usage
Past 30 DaysNot available
Not in Playground
Claude Opus 4.1 has been deprecated by its provider and can no longer be evaluated on the current benchmark. The legacy Vision Evals results below are preserved for reference. See the current Vision Evals
| Category | Passed | Score |
|---|---|---|
| Document Understanding | 8 / 9 | 88.9% |
| Defect Detection | 11 / 15 | 73.3% |
| Object Understanding | 9 / 14 | 64.3% |
| Spatial Understanding | 12 / 19 | 63.2% |
| Object Counting | 0 / 10 | 0% |
| Category | Passed | Score |
|---|---|---|
| Text Recognition | 24 / 30 | 80% |
| Focused Scene OCR | 73 / 99 | 73.7% |
| VQA & Extraction | 41 / 60 | 68.3% |
| License Plate Recognition | 16 / 30 | 53.3% |
| Handwritten Math | 3 / 10 | 30% |
Scores based on a single evaluation run · Methodology
View all legacy Vision Evals results →Claude Opus 4.1 costs $15.00 per 1M input tokens and $75.00 per 1M output tokens.
Pricing updated Aug 12, 2026
Estimated cost per task vs. Visual Understanding score, for this model and others ranked near it. Upper-left is the sweet spot (high quality, low cost). Based on Vision Evals (legacy) results.
11 of 11 models plotted
| Model | Score | Median tokens | Est. cost / task | Compare |
|---|---|---|---|---|
| Gemma 4 31B | 67.2% | 467 | $0.0001 | Compare |
| Claude Opus 4.6 | 64.2% | 2.3K | $0.014 | Compare |
| GPT-5.4 Nano | 62.7% | 1.8K | $0.0004 | Compare |
| Llama 4 Maverick | 59.7% | 2.4K | $0.0005 | Compare |
| Claude Sonnet 4.5 | 59.7% | 2.3K | $0.0092 | Compare |
| Claude Opus 4.1(this model) | 59.7% | 2.1K | $0.040 | — |
| Claude Haiku 4.5 | 58.2% | 2.3K | $0.0030 | Compare |
| GPT-5 Nano | 58.2% | 2.7K | $0.0003 | Compare |
| Qwen3.5 397B A17B | 58.2% | 1.5K | $0.0008 | Compare |
| Gemini 2.5 Flash | 55.2% | 476 | $0.0005 | Compare |
| Gemini 2.5 Flash-Lite | 53.7% | 301 | <$0.0001 | Compare |
Other models worth comparing for similar use cases.
Other versions in the same family as Claude Opus 4.1.
Claude Opus 4.1 is proprietary: the weights are not distributed, and the Claude Opus 4.1 license is the vendor's commercial terms of service that you accept when you call the API.
Vendor terms govern data retention, whether your inputs can be trained on, rate limits, and regional availability, and they can change with notice. Review them if you handle regulated or customer data.
Proprietary terms are set by the vendor rather than negotiated per project, and no open-source obligation attaches to your code. If you would rather deploy a model whose commercial license is included in your plan — on Roboflow Managed Cloud or a Self-Hosted Inference Server — Roboflow's licensing page lists the supported alternatives to Claude Opus 4.1.
Do not hesitate to reach out with questions for your commercial project — our team will help you start solving business problems on the first call. See Roboflow commercial licensing for the models included in each plan.
Talk to salesThis model is proprietary. The author retains all rights, and use of the model is governed by their specific terms of service or license agreement.
Commercial use depends on the terms set by the model author. Most proprietary commercial models require a paid subscription, API key, or per-call billing. Check the provider’s pricing and terms-of-service for details.
License information is provided as a guide and is not legal advice.
Yes. Claude Opus 4.1 accepts image input, and on Roboflow's previous vision benchmark it passed 59.7% of visual understanding tasks (#46 of 77) and scored 68.6% on OCR.
Claude Opus 4.1 has been deprecated by its provider and can no longer be run, so it is not part of Roboflow's current Vision Evals. Its results from the previous benchmark are preserved on this page for reference.