SmolVLM2-2.2B-Instruct
HuggingFaceTB/SmolVLM2-2.2B-Instruct
Tracked by OpenModelStats since August 24, 2026View on Hugging Face ↗
- Downloads 30DRolling recent-window download count as reported by the source platform.
- 202.8K
- Total downloadsCumulative all-time download counter reported by the source platform.
- 3.2M
- Gained 7DIncrease in cumulative total downloads between OpenModelStats' own observations seven days apart.
- —
- Gained 30DIncrease in cumulative total downloads between OpenModelStats' own observations roughly thirty days apart.
- —
- LikesCommunity likes reported by the source platform.
- 331
- Rank · tracked modelsPosition among all eligible models tracked by OpenModelStats, ordered by recent downloads. #241 among tracked Image Text to Text models.
- #1,537
- 934(down 934 places)this week
- Parameters
- 2.2B
- Last source update
- Apr 8, 2025
- It moved from #127 to #241 among tracked Image Text to Text models this week.
Downloads over time
Cumulative total downloads observed by OpenModelStatsNot enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.
Likes over time
Not enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.
Rank among tracked models
Lower is betterNot enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.
Overview
SmolVLM2-2.2B-Instruct is a image text to text model published by HuggingFaceTB. OpenModelStats has tracked the model since Aug 24, 2026, most recently observing it 3h ago.
- Published
- Feb 8, 2025
- Last updated
- Apr 8, 2025
- Library
- transformers
- License
- apache-2.0
- Task
- Image Text to Text
- Parameters
- 2,246,784,880
- Spaces
- —
- Derivative models
- —
- Velocity 7D/day
- —
transformerssafetensorssmolvlmimage-text-to-textvideo-text-to-textconversationalendataset:HuggingFaceM4/the_cauldrondataset:HuggingFaceM4/Docmatixdataset:lmms-lab/LLaVA-OneVision-Datadataset:lmms-lab/M4-Instruct-Datadataset:HuggingFaceFV/finevideo
Model family
Based on base-model relationships reported by the source's metadata.
Base model
- SmolVLM-Instruct27.1K
Similar models
| Model | DL 30D |
|---|---|
| SmolVLM2-2.2B-Instruct-Agentic-GUIsmolagents | 96 |
| SmolVLM-InstructHuggingFaceTB | 27.1K |
| SmolVLM-BaseHuggingFaceTB | 6,348 |
| Qwen2-VL-2B-Instruct-unsloth-bnb-4bitunsloth | 3,501 |
| Qwen2-VL-2B-Instruct-bnb-4bitunsloth | 2,548 |
| Ovis2-2B-hfthisisiron | 2,700 |