VideoLLaMA3-7B

DAMO-NLP-SG/VideoLLaMA3-7B

Video Text To Text8B parameterstransformersLicense: apache-2.0
Tracked by OpenModelStats since September 8, 2026View on Hugging Face ↗
Downloads 30D
3,895
Total downloads
1.3M
Gained 7D
+1,862(increase)
0.1%(up) vs baseline
Gained 30D
—
Likes
78
+0 in 7D
Rank · tracked models
#13,705
144(down 144 places)this week
Parameters
8B
Last source update
Sep 2, 2025
Momentum
51.9
  • VideoLLaMA3-7B gained 1,862 downloads during the last seven tracked days.

Downloads over time

Cumulative total downloads observed by OpenModelStats

Likes over time

Rank among tracked models

Lower is better

Overview

VideoLLaMA3-7B is a video text to text model published by DAMO-NLP-SG. OpenModelStats has tracked the model since Sep 8, 2026, most recently observing it 11h ago.

Published
Jan 21, 2025
Last updated
Sep 2, 2025
Library
transformers
License
apache-2.0
Task
Video Text To Text
Parameters
8,044,744,944
Spaces
—
Derivative models
—
Velocity 7D/day
266
transformerssafetensorsvideollama3_qwen2text-generationmulti-modallarge-language-modelvideo-language-modelvideo-text-to-textcustom_codeendataset:lmms-lab/LLaVA-OneVision-Datadataset:allenai/pixmo-docs

Publisher

DAMO-NLP-SGView publisher statistics →

Derived from DAMO-NLP-SG/VideoLLaMA3-7B-Image (reported by the source; parent not yet tracked).

Similar models

Models similar to VideoLLaMA3-7B
ModelDL 30D
LLaVA-Video-7B-Qwen2lmms-lab12.4K
Video-XL-2BAAI266
VideoChat-Flash-Qwen2-7B_res448OpenGVLab2,134
VideoScore-v1.1TIGER-Lab1,050
VideoChat-R1_7B_captionOpenGVLab44
TimeLens-7BTencentARC924
DAMO-NLP-SG/VideoLLaMA3-7B — Downloads, Growth & Stats | OpenModelStats