VideoLLaMA3-2B
DAMO-NLP-SG/VideoLLaMA3-2B
Tracked by OpenModelStats since August 24, 2026View on Hugging Face ↗
- Downloads 30DRolling recent-window download count as reported by the source platform.
- 2,800
- Total downloadsCumulative all-time download counter reported by the source platform.
- 116.1K
- Gained 7DIncrease in cumulative total downloads between OpenModelStats' own observations seven days apart.
- —
- Gained 30DIncrease in cumulative total downloads between OpenModelStats' own observations roughly thirty days apart.
- —
- LikesCommunity likes reported by the source platform.
- 21
- Rank · tracked modelsPosition among all eligible models tracked by OpenModelStats, ordered by recent downloads. #13 among tracked Video Text To Text models.
- #14,530
- –this week
- Parameters
- 2B
- Last source update
- Sep 3, 2025
Downloads over time
Cumulative total downloads observed by OpenModelStatsNot enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.
Likes over time
Not enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.
Rank among tracked models
Lower is betterNot enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.
Overview
VideoLLaMA3-2B is a video text to text model published by DAMO-NLP-SG. OpenModelStats has tracked the model since Aug 24, 2026, most recently observing it 8h ago.
- Published
- Jan 21, 2025
- Last updated
- Sep 3, 2025
- Library
- transformers
- License
- apache-2.0
- Task
- Video Text To Text
- Parameters
- 1,959,993,584
- Spaces
- —
- Derivative models
- —
- Velocity 7D/day
- —
transformerssafetensorsvideollama3_qwen2text-generationmulti-modallarge-language-modelvideo-language-modelvideo-text-to-textcustom_codeendataset:lmms-lab/LLaVA-OneVision-Datadataset:allenai/pixmo-docs
Publisher
DAMO-NLP-SGView publisher statistics →Derived from DAMO-NLP-SG/VideoLLaMA3-2B-Image (reported by the source; parent not yet tracked).
Similar models
| Model | DL 30D |
|---|---|
| Marlin-2BNemoStation | 6,598 |
| Marlin-2B-ungatedlunahr | 2,316 |
| VideoChat3-4BMCG-NJU | 2,927 |
| Spatial-MLLM-v1.1-Instruct-135KDiankun | 2,744 |
| LLaVA-NeXT-Video-7Blmms-lab | 227 |
| LLaVA-NeXT-Video-7B-hfllava-hf | 122.2K |