SmolVLM2-256M-Video-Instruct
HuggingFaceTB/SmolVLM2-256M-Video-Instruct
Tracked by OpenModelStats since September 8, 2026View on Hugging Face ↗
- Downloads 30DRolling recent-window download count as reported by the source platform.
- 38K
- Total downloadsCumulative all-time download counter reported by the source platform.
- 2.2M
- Gained 7DIncrease in cumulative total downloads between OpenModelStats' own observations seven days apart.
- +8,714(increase)
- 0.4%(up) vs baseline
- Gained 30DIncrease in cumulative total downloads between OpenModelStats' own observations roughly thirty days apart.
- —
- LikesCommunity likes reported by the source platform.
- 115
- +1 in 7D
- Rank · tracked modelsPosition among all eligible models tracked by OpenModelStats, ordered by recent downloads. #549 among tracked Image Text to Text models.
- #4,291
- 50(down 50 places)this week
- Parameters
- 256M
- Last source update
- Apr 8, 2025
- MomentumOpenModelStats' composite rising-model signal (0–100): a percentile blend of absolute 7-day download gains, capped percentage growth, and likes gained. See Methodology.
- 78.7
- SmolVLM2-256M-Video-Instruct gained 8,714 downloads during the last seven tracked days.
- It moved from #538 to #549 among tracked Image Text to Text models this week.
Downloads over time
Cumulative total downloads observed by OpenModelStatsLikes over time
Rank among tracked models
Lower is betterOverview
SmolVLM2-256M-Video-Instruct is a image text to text model published by HuggingFaceTB. OpenModelStats has tracked the model since Sep 8, 2026, most recently observing it 11h ago.
- Published
- Feb 11, 2025
- Last updated
- Apr 8, 2025
- Library
- transformers
- License
- apache-2.0
- Task
- Image Text to Text
- Parameters
- 256,484,928
- Spaces
- —
- Derivative models
- —
- Velocity 7D/day
- 1,245
transformersonnxsafetensorssmolvlmimage-text-to-textconversationalendataset:HuggingFaceM4/the_cauldrondataset:HuggingFaceM4/Docmatixdataset:lmms-lab/LLaVA-OneVision-Datadataset:lmms-lab/M4-Instruct-Datadataset:HuggingFaceFV/finevideo
Publisher
HuggingFaceTBView publisher statistics →Derived from HuggingFaceTB/SmolVLM-256M-Instruct (reported by the source; parent not yet tracked).
Similar models
| Model | DL 30D |
|---|---|
| SmolVLM-256M-InstructHuggingFaceTB | 380.8K |
| SmolDocling-256M-previewdocling-project | 26.8K |
| SmolDocling-256M-preview-mlx-bf16docling-project | 1,590 |
| SmolVLM-256M-BaseHuggingFaceTB | 489 |
| granite-docling-258Mibm-granite | 276K |
| granite-docling-2stage-258mdocling-project | 747 |