Compare the best 72 image classification models and try 61 of them on your own image, free in the Roboflow Playground. 33 are open-weight, so you can self-host them for free under their licenses.
72 models · 33 open-weight · 61 free to try · prices synced Jul 28, 2026
We haven't benchmarked image classification yet; scores and rankings appear on this site only where we've measured them. Until then, compare the models below, and for fixed categories in production expect a model fine-tuned on your own data to win.
33 models with downloadable weights you can self-host under their licenses (Modified MIT, Apache 2.0, and Proprietary). 22 run live in the Playground through hosted APIs, so self-hosting is optional.
Kimi K3Moonshot AI NEW | 2.8T | Modified MIT | Jul 2026 | |
Qwen3.6 27BQwen | 27B | Apache 2.0 | Apr 2026 | |
Qwen3.6 35B A3BQwen | 35B total, 3B active | Apache 2.0 | Apr 2026 | |
Gemma 4 26B A4BGoogle | 25.2B | Apache 2.0 | Apr 2026 | |
Gemma 4 31BGoogle | 31B | Apache 2.0 | Apr 2026 | |
Qwen3.5 9bQwen | 9B | Apache 2.0 | Mar 2026 | |
| 122B | Apache 2.0 | Feb 2026 | ||
Qwen3.5 27BQwen | 27B | Apache 2.0 | Feb 2026 | |
Qwen3.5 35B A3BQwen | 35B | Apache 2.0 | Feb 2026 | |
| 397B | Apache 2.0 | Feb 2026 | ||
Kimi K2.5Moonshot AI | 1T | Modified MIT | Jan 2026 | |
| 8.8B | Apache 2.0 | Oct 2025 | ||
| 31B | Apache 2.0 | Oct 2025 | ||
| 235B | Apache 2.0 | Sep 2025 | ||
Llama 4 MaverickMeta | 400B | Proprietary | Apr 2025 | |
Llama 4 ScoutMeta | 109B | Proprietary | Apr 2025 | |
Mistral Small 3.1 24BMistral | 24B | Apache 2.0 | Mar 2025 | |
Gemma 3 12BGoogle | 12B | Proprietary | Mar 2025 | |
Gemma 3 27BGoogle | — | Proprietary | Mar 2025 | |
Gemma 3 4BGoogle | 4B | Proprietary | Mar 2025 | |
YOLOv12THU-MIG | 2.6M-59.1M | AGPL 3.0 | Feb 2025 | |
| 7B | Apache 2.0 | Jan 2025 | ||
Pixtral 12BMistral | 12B | Apache 2.0 | Sep 2024 | |
SAM-CLIPApple | — | Custom | Oct 2023 | |
DINOv2Meta | 21M-1.1B | Apache 2.0 | Apr 2023 | |
SigLIPGoogle | 200M-900M | Apache 2.0 | Mar 2023 | |
YOLOv8 ClassificationUltralytics | — | AGPL 3.0 | Jan 2023 | |
CLIPOpenAI | — | MIT | Feb 2021 | |
Vision Transformer (ViT)Google | 86M-632M | Apache 2.0 | Oct 2020 | |
MobileNetV2Google | ~3.4M | Apache 2.0 | Jan 2018 | |
ResNet-32Meta | 0.46M | MIT | Dec 2015 | |
ResNet-34Meta | 21.8M | MIT | Dec 2015 | |
ResNet-50Microsoft | 25.6M | MIT | Dec 2015 |
39 proprietary models where the weights aren't downloadable: access is through each provider's API and billed by them. Try all of them free in the Playground.
Qwen3.7 FlashQwen NEW | $0.030 | $0.13 | 1M | Jul 2026 | |
Claude Opus 5Anthropic NEW | $5.00 | $25.00 | 1M | Jul 2026 | |
Gemini 3.5 Flash-LiteGoogle NEW | $0.30 | $2.50 | 1.0M | Jul 2026 | |
Gemini 3.6 FlashGoogle NEW | $1.50 | $7.50 | 1M | Jul 2026 | |
GPT-5.6 LunaOpenAI NEW | $0.50 | $3.00 | 1.5M | Jul 2026 | |
GPT-5.6 SolOpenAI NEW | $5.00 | $30.00 | 1.5M | Jul 2026 | |
GPT-5.6 TerraOpenAI NEW | $1.25 | $7.50 | 1.1M | Jul 2026 | |
Muse Spark 1.1Meta NEW | $1.25 | $4.25 | 1.0M | Jul 2026 | |
Claude Sonnet 5Anthropic NEW | $2.00 | $10.00 | 1M | Jun 2026 | |
Claude Fable 5Anthropic NEW | $10.00 | $50.00 | 1M | Jun 2026 | |
Claude Opus 4.8Anthropic | $5.00 | $25.00 | 1M | May 2026 | |
Gemini 3.5 FlashGoogle | $1.50 | $9.00 | 1.0M | May 2026 | |
GPT-5.5OpenAI | $5.00 | $30.00 | 1M | Apr 2026 | |
Claude Opus 4.7Anthropic | $5.00 | $25.00 | 1M | Apr 2026 | |
Qwen3.6 FlashQwen | $0.19 | $1.13 | 1M | Apr 2026 | |
Qwen3.6 PlusQwen | $0.33 | $1.95 | 1M | Apr 2026 | |
GPT-5.4 MiniOpenAI | $0.75 | $4.50 | 400K | Mar 2026 | |
GPT-5.4 NanoOpenAI | $0.20 | $1.25 | 400K | Mar 2026 | |
GPT-5.4OpenAI | $2.50 | $15.00 | 1.1M | Mar 2026 | |
Gemini 3.1 Flash-LiteGoogle | $0.25 | $1.50 | 1M | Mar 2026 | |
Gemini 3.1 ProGoogle | $2.00 | $12.00 | 1M | Feb 2026 | |
Claude Sonnet 4.6Anthropic | $3.00 | $15.00 | 1M | Feb 2026 | |
Claude Opus 4.6 Anthropic | $5.00 | $25.00 | 1M | Feb 2026 | |
Gemini 3 FlashGoogle | $0.50 | $3.00 | 1M | Dec 2025 | |
GPT-5.2OpenAI | $1.75 | $14.00 | 400K | Dec 2025 | |
Claude Opus 4.5Anthropic | $5.00 | $25.00 | 200K | Nov 2025 | |
GPT-5.1OpenAI | $1.25 | $10.00 | 196K | Nov 2025 | |
Claude Haiku 4.5Anthropic | $1.00 | $5.00 | 200K | Oct 2025 | |
Claude Sonnet 4.5Anthropic | $3.00 | $15.00 | 200K | Sep 2025 | |
Mistral Medium 3.1Mistral | $0.40 | $2.00 | 128K | Aug 2025 | |
GPT-5OpenAI | $1.25 | $10.00 | — | Aug 2025 | |
GPT-5 MiniOpenAI | $0.25 | $2.00 | 400K | Aug 2025 | |
GPT-5 NanoOpenAI | $0.050 | $0.40 | 400K | Aug 2025 | |
Claude Opus 4.1Anthropic | $15.00 | $75.00 | 200K | Aug 2025 | |
Gemini 2.5 Flash-LiteGoogle | $0.10 | $0.40 | 1M | Jul 2025 | |
Gemini 2.5 FlashGoogle | $0.30 | $2.50 | 1M | Jul 2025 | |
Grok 4xAI | — | — | — | Jul 2025 | |
Gemini 2.5 ProGoogle | $1.25 | $10.00 | 1M | Jun 2025 | |
Qwen VL MaxQwen | — | — | — | Feb 2025 |
Image classification models split into two groups, and the right choice comes down to two questions: are your categories fixed, and how much labeled data do you have?
CLIP-style models (CLIP, SigLIP) classify against text labels you provide at inference time, with no training: they embed the image and your candidate labels in the same space and pick the closest match. Vision language models do the same from a prompt and can also justify their answer. Use this group when categories change often, when you have no labeled data yet, or when you are validating an idea before investing in a dataset.
The tradeoff is accuracy on fine-grained or domain-specific classes: a general model knows what a beverage aisle is, but not your SKUs, and subtle distinctions (defect grades, near-identical variants) usually defeat zero-shot approaches.
When classes are fixed and accuracy matters, a classifier fine-tuned on your own images (architectures like ResNet, EfficientNet, and ViT) is smaller, faster, and more accurate on your domain than any general-purpose model. A few hundred labeled images per class is often enough to start, and the result deploys to the cloud or fully offline on the edge at a fraction of a VLM call cost.
Most production teams do both: validate the task zero-shot on this page, use the early predictions to bootstrap a labeled dataset, then train a compact classifier once the taxonomy settles. The zero-shot model stays useful for the categories that keep changing.
The bottom line: Changing labels or no data: start zero-shot. Fixed categories in production: train your own classifier. Most teams prototype with the first and ship the second.
Image classification is the task of assigning a label to an entire image from a set of categories, answering "what is this a picture of?". A model encodes the image, usually with a convolutional network like ResNet or EfficientNet or a vision transformer (ViT), and outputs a probability per class; the top score wins. CLIP-style models extend this to zero-shot classification: they embed the image and your candidate label texts in the same vector space and pick the closest match, so you can change categories without retraining. Accuracy is reported as top-1 or top-5 on the fixed label set. Classification differs from detection (which localizes objects) and tagging (which applies many loose keywords), and it powers quality inspection, content moderation, product categorization, and medical screening. This page lists 72 image classification models, including 33 open-weight options you can self-host; 61 of them run live in the Playground so you can test them on your own images.
It depends on your task and constraints. For fixed categories in production, a model fine-tuned on your own data typically beats any general-purpose model. Compare the image classification models on this page and try them on your own image to see which fits.
Yes. 33 of the 72 image classification models here are open-weight (for example Kimi K3, Qwen3.6 27B, and Qwen3.6 35B A3B), free to self-host under their licenses (Modified MIT, Apache 2.0, and Proprietary).
Yes. You can run 61 of them in the Roboflow Playground for free. Upload an image and compare the models' output side by side, no setup required.
This page lists all 72 image classification models in the Roboflow Playground catalog: 33 open-weight models you can self-host and 39 proprietary models accessed through provider APIs; 61 of them run live in the Roboflow Playground on your own images. Compare licenses, parameters, API prices, and release dates side by side, or open any model page for full details.