SmolVLM2-256M-Video-Instruct

HuggingFaceTB/SmolVLM2-256M-Video-Instruct

Image Text to Text256M parameterstransformersLicense: apache-2.0
Tracked by OpenModelStats since August 24, 2026View on Hugging Face ↗
Downloads 30D
59.8K
Total downloads
2.2M
Gained 7D
Gained 30D
Likes
113
Rank · tracked models
#3,463
this week
Parameters
256M
Last source update
Apr 8, 2025

Downloads over time

Cumulative total downloads observed by OpenModelStats

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Likes over time

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Rank among tracked models

Lower is better

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Overview

SmolVLM2-256M-Video-Instruct is a image text to text model published by HuggingFaceTB. OpenModelStats has tracked the model since Aug 24, 2026, most recently observing it 3 min ago.

Published
Feb 11, 2025
Last updated
Apr 8, 2025
Library
transformers
License
apache-2.0
Task
Image Text to Text
Parameters
256,484,928
Spaces
Derivative models
Velocity 7D/day
transformersonnxsafetensorssmolvlmimage-text-to-textconversationalendataset:HuggingFaceM4/the_cauldrondataset:HuggingFaceM4/Docmatixdataset:lmms-lab/LLaVA-OneVision-Datadataset:lmms-lab/M4-Instruct-Datadataset:HuggingFaceFV/finevideo

Publisher

HuggingFaceTBView publisher statistics →

Derived from HuggingFaceTB/SmolVLM-256M-Instruct (reported by the source; parent not yet tracked).

Similar models

Models similar to SmolVLM2-256M-Video-Instruct
ModelDL 30D
SmolVLM-256M-InstructHuggingFaceTB778.7K
SmolDocling-256M-previewdocling-project29.9K
SmolVLM-256M-BaseHuggingFaceTB3,031
SmolDocling-256M-preview-mlx-bf16docling-project1,506
granite-docling-258Mibm-granite434.8K
Florence-2-VQAJP2aipib173.1K