SmolVLM-500M-Instruct

HuggingFaceTB/SmolVLM-500M-Instruct

Image Text to Text507M parameterstransformersLicense: apache-2.0
Tracked by OpenModelStats since August 24, 2026View on Hugging Face ↗
Downloads 30D
166.8K
Total downloads
1.9M
Gained 7D
Gained 30D
Likes
196
Rank · tracked models
#1,767
this week
Parameters
507M
Last source update
Apr 8, 2025

Downloads over time

Cumulative total downloads observed by OpenModelStats

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Likes over time

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Rank among tracked models

Lower is better

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Overview

SmolVLM-500M-Instruct is a image text to text model published by HuggingFaceTB. OpenModelStats has tracked the model since Aug 24, 2026, most recently observing it 5 min ago.

Published
Jan 20, 2025
Last updated
Apr 8, 2025
Library
transformers
License
apache-2.0
Task
Image Text to Text
Parameters
507,482,304
Spaces
Derivative models
Velocity 7D/day
transformersonnxsafetensorsidefics3image-text-to-textconversationalendataset:HuggingFaceM4/the_cauldrondataset:HuggingFaceM4/Docmatixarxiv:2504.05299base_model:HuggingFaceTB/SmolLM2-360M-Instructbase_model:quantized:HuggingFaceTB/SmolLM2-360M-Instruct

Model family

Based on base-model relationships reported by the source's metadata.

Base model

Derivative models

Similar models

Models similar to SmolVLM-500M-Instruct
ModelDL 30D
SmolVLM2-500M-Video-InstructHuggingFaceTB1.5M
SmolVLM-500M-BaseHuggingFaceTB1,394
VisionPsy-Nano-460Mqvac2,178
VisionPsy-Nano-460M-Flashqvac508
LFM2.5-VL-3B-OptiQ-4bitmlx-community1,191
gemma-4-31B-it-qat-q4_0-unquantized-assistantgoogle19.6K