SmolVLM2-256M-Video-Instruct

HuggingFaceTB/SmolVLM2-256M-Video-Instruct

Image Text to Text256M parameterstransformersLicense: apache-2.0
Tracked by OpenModelStats since August 24, 2026View on Hugging Face ↗
Downloads 30D
59.8K
Total downloads
2.2M
Gained 7D
Gained 30D
Likes
113
Rank · tracked models
#3,463
this week
Parameters
256M
Last source update
Apr 8, 2025

Downloads over time

Daily gain between consecutive observations

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Likes over time

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Rank among tracked models

Lower is better

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Overview

SmolVLM2-256M-Video-Instruct is a image text to text model published by HuggingFaceTB. OpenModelStats has tracked the model since Aug 24, 2026, most recently observing it 3h ago.

Published
Feb 11, 2025
Last updated
Apr 8, 2025
Library
transformers
License
apache-2.0
Task
Image Text to Text
Parameters
256,484,928
Spaces
Derivative models
Velocity 7D/day
transformersonnxsafetensorssmolvlmimage-text-to-textconversationalendataset:HuggingFaceM4/the_cauldrondataset:HuggingFaceM4/Docmatixdataset:lmms-lab/LLaVA-OneVision-Datadataset:lmms-lab/M4-Instruct-Datadataset:HuggingFaceFV/finevideo

Publisher

HuggingFaceTBView publisher statistics →

Derived from HuggingFaceTB/SmolVLM-256M-Instruct (reported by the source; parent not yet tracked).

Similar models

Models similar to SmolVLM2-256M-Video-Instruct
ModelDL 30D
SmolVLM-256M-InstructHuggingFaceTB778.7K
SmolDocling-256M-previewdocling-project29.9K
SmolVLM-256M-BaseHuggingFaceTB3,031
SmolDocling-256M-preview-mlx-bf16docling-project1,506
granite-docling-258Mibm-granite434.8K
Florence-2-VQAJP2aipib173.1K