VideoLLaMA3-7B

DAMO-NLP-SG/VideoLLaMA3-7B

Video Text To Text8B parameterstransformersLicense: apache-2.0
Tracked by OpenModelStats since August 24, 2026View on Hugging Face ↗
Downloads 30D
6,983
Total downloads
1.3M
Gained 7D
Gained 30D
Likes
76
Rank · tracked models
#9,512
this week
Parameters
8B
Last source update
Sep 2, 2025

Downloads over time

Cumulative total downloads observed by OpenModelStats

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Likes over time

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Rank among tracked models

Lower is better

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Overview

VideoLLaMA3-7B is a video text to text model published by DAMO-NLP-SG. OpenModelStats has tracked the model since Aug 24, 2026, most recently observing it 5h ago.

Published
Jan 21, 2025
Last updated
Sep 2, 2025
Library
transformers
License
apache-2.0
Task
Video Text To Text
Parameters
8,044,744,944
Spaces
Derivative models
Velocity 7D/day
transformerssafetensorsvideollama3_qwen2text-generationmulti-modallarge-language-modelvideo-language-modelvideo-text-to-textcustom_codeendataset:lmms-lab/LLaVA-OneVision-Datadataset:allenai/pixmo-docs

Publisher

DAMO-NLP-SGView publisher statistics →

Derived from DAMO-NLP-SG/VideoLLaMA3-7B-Image (reported by the source; parent not yet tracked).

Similar models

Models similar to VideoLLaMA3-7B
ModelDL 30D
LLaVA-Video-7B-Qwen2lmms-lab14.4K
Video-XL-2BAAI348
VideoChat-Flash-Qwen2-7B_res448OpenGVLab1,288
VideoScore-v1.1TIGER-Lab1,325
qwen2.5-vl-7b-cam-motionchancharikm1,039
SkyCaptioner-V1Skywork169