Compare both Xiaomi vision models we track, all of them open-weight. Try them on your own images, free in the Roboflow Playground.
2 models · 2 open-weight · 2 free to try · Updated Sep 2026
Every Xiaomi vision model in our catalog is open-weight. Both publish downloadable weights, so you can self-host under their own licenses instead of paying per request. The open-weight side runs on MIT terms. The newest additions are MiMo V2.6 Flash (Sep 2026) and MiMo V2.6 Pro (Sep 2026).
Between them, the Xiaomi models we list cover 10 distinct vision tasks. The widest coverage is captioning (2 models), chart question answering (2), and classification (2). For captioning, start with MiMo V2.6 Flash or MiMo V2.6 Pro.
On the open-weight tier, MiMo V2.6 Flash is the largest at 309B total, 15B active parameters and MiMo V2.6 Pro the smallest at 1.02T total, 42B active. Both carry a permissive license, so commercial use is straightforward.
With every model here open-weight, the real trade is size against generality: the smaller checkpoints fine-tune and deploy cheaply on your own hardware, while the larger ones cover more ground out of the box. Both Xiaomi vision models run live in the Roboflow Playground, so you can run the same image through both and compare the answers before committing to one.
What each of the 2 Xiaomi vision models in our catalog is built for, and how you run it.
2 models with downloadable weights you can self-host under their licenses (MIT). All run live in the Playground through hosted APIs, so self-hosting is optional.
2 of the 2 Xiaomi vision models we track handle captioning: MiMo V2.6 Flash and MiMo V2.6 Pro. Each model page lists its full task coverage, license, and specs.
2 of the 2 Xiaomi vision models we track handle chart question answering: MiMo V2.6 Flash and MiMo V2.6 Pro. Each model page lists its full task coverage, license, and specs.
Yes. Both Xiaomi vision models we track publish downloadable weights you can self-host under their licenses (MIT).
We do not publish a Xiaomi-only ranking, so pick on constraints rather than a label. 2 Xiaomi models handle captioning; the most recent is MiMo V2.6 Flash (Sep 2026). For fixed categories in production, a model fine-tuned on your own data typically beats any general-purpose model.
We track 2 live Xiaomi vision models. The most recent addition is MiMo V2.6 Flash, released Sep 2026.
Yes. Both Xiaomi models run live in the Roboflow Playground. Upload your own image, run several models on it at once, and compare the outputs side by side. No setup and no account required.
This page lists both Xiaomi vision models in the Roboflow Playground catalog, all of them open-weight and free to self-host. They cover captioning, chart question answering, and classification, among other tasks. All of them run live in the Roboflow Playground on your own images. Compare licenses, parameters, prices, and release dates side by side, or open any model page for full details.