SmolVLM2-256M-Video-Instruct

HuggingFaceTB/SmolVLM2-256M-Video-Instruct

Image Text to Text256M parameterstransformersLicense: apache-2.0
Tracked by OpenModelStats since August 24, 2026View on Hugging Face ↗
Downloads 30D
39.2K
Total downloads
2.2M
Gained 7D
+8,714(increase)
0.4%(up) vs baseline
Gained 30D
—
Likes
115
+1 in 7D
Rank · tracked models
#4,241
34(down 34 places)this week
Parameters
256M
Last source update
Apr 8, 2025
Momentum
78.7
  • SmolVLM2-256M-Video-Instruct gained 8,714 downloads during the last seven tracked days.
  • It moved from #533 to #538 among tracked Image Text to Text models this week.

Downloads over time

Cumulative total downloads observed by OpenModelStats

Likes over time

Rank among tracked models

Lower is better

Overview

SmolVLM2-256M-Video-Instruct is a image text to text model published by HuggingFaceTB. OpenModelStats has tracked the model since Aug 24, 2026, most recently observing it Oct 7, 2026.

Published
Feb 11, 2025
Last updated
Apr 8, 2025
Library
transformers
License
apache-2.0
Task
Image Text to Text
Parameters
256,484,928
Spaces
—
Derivative models
—
Velocity 7D/day
1,245
transformersonnxsafetensorssmolvlmimage-text-to-textconversationalendataset:HuggingFaceM4/the_cauldrondataset:HuggingFaceM4/Docmatixdataset:lmms-lab/LLaVA-OneVision-Datadataset:lmms-lab/M4-Instruct-Datadataset:HuggingFaceFV/finevideo

Publisher

HuggingFaceTBView publisher statistics →

Derived from HuggingFaceTB/SmolVLM-256M-Instruct (reported by the source; parent not yet tracked).

Similar models

Models similar to SmolVLM2-256M-Video-Instruct
ModelDL 30D
SmolVLM-256M-InstructHuggingFaceTB397.6K
SmolDocling-256M-previewdocling-project26.5K
SmolDocling-256M-preview-mlx-bf16docling-project1,585
SmolVLM-256M-BaseHuggingFaceTB492
granite-docling-258Mibm-granite274.9K
Florence-2-VQAJP2aipib143.3K
HuggingFaceTB/SmolVLM2-256M-Video-Instruct — Downloads, Growth & Stats | OpenModelStats