VideoLLaMA3-2B

DAMO-NLP-SG/VideoLLaMA3-2B

Video Text To Text2B parameterstransformersLicense: apache-2.0
Tracked by OpenModelStats since September 8, 2026View on Hugging Face ↗
Downloads 30D
3,978
Total downloads
120.7K
Gained 7D
+2,470(increase)
2.1%(up) vs baseline
Gained 30D
—
Likes
21
+0 in 7D
Rank · tracked models
#13,422
17(down 17 places)this week
Parameters
2B
Last source update
Sep 3, 2025
Momentum
60.8
  • VideoLLaMA3-2B gained 2,470 downloads during the last seven tracked days.
  • It moved from #9 to #8 among tracked Video Text To Text models this week.

Downloads over time

Cumulative total downloads observed by OpenModelStats

Likes over time

Rank among tracked models

Lower is better

Overview

VideoLLaMA3-2B is a video text to text model published by DAMO-NLP-SG. OpenModelStats has tracked the model since Sep 8, 2026, most recently observing it 4h ago.

Published
Jan 21, 2025
Last updated
Sep 3, 2025
Library
transformers
License
apache-2.0
Task
Video Text To Text
Parameters
1,959,993,584
Spaces
—
Derivative models
—
Velocity 7D/day
353
transformerssafetensorsvideollama3_qwen2text-generationmulti-modallarge-language-modelvideo-language-modelvideo-text-to-textcustom_codeendataset:lmms-lab/LLaVA-OneVision-Datadataset:allenai/pixmo-docs

Publisher

DAMO-NLP-SGView publisher statistics →

Derived from DAMO-NLP-SG/VideoLLaMA3-2B-Image (reported by the source; parent not yet tracked).

Similar models

Models similar to VideoLLaMA3-2B
ModelDL 30D
VideoChat-Flash-Qwen2_5-2B_res448OpenGVLab305
Marlin-2BNemoStation2,749
Marlin-2B-ungatedlunahr462
TimeLens2-2BMCG-NJU455
GenS-qwen2d5-vl-3byaolily4,815
CASA-Qwen2_5-VL-3B-LiveCCkyutai2,413
DAMO-NLP-SG/VideoLLaMA3-2B — Downloads, Growth & Stats | OpenModelStats