VideoLLaMA3-7B-Image

DAMO-NLP-SG/VideoLLaMA3-7B-Image

Visual Question Answering8B parameterstransformersLicense: apache-2.0
Tracked by OpenModelStats since October 7, 2026View on Hugging Face ↗
Downloads 30D
146
Total downloads
23K
Gained 7D
—
Gained 30D
—
Likes
11
Rank · tracked models
—
–this week
Parameters
8B
Last source update
Mar 20, 2025

Downloads over time

Rolling recent-window downloads as reported

Likes over time

Rank among tracked models

Lower is better

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-10-07. Charts appear as daily observations accumulate.

Overview

VideoLLaMA3-7B-Image is a visual question answering model published by DAMO-NLP-SG. OpenModelStats has tracked the model since Oct 7, 2026, most recently observing it 3h ago.

Published
Jan 21, 2025
Last updated
Mar 20, 2025
Library
transformers
License
apache-2.0
Task
Visual Question Answering
Parameters
8,044,744,944
Spaces
—
Derivative models
—
Velocity 7D/day
—
transformerssafetensorsvideollama3_qwen2text-generationmulti-modallarge-language-modelvideo-language-modelvisual-question-answeringcustom_codeendataset:lmms-lab/LLaVA-OneVision-Datadataset:allenai/pixmo-docs

Publisher

DAMO-NLP-SGView publisher statistics →

Derived from Qwen/Qwen2.5-7B-Instruct (reported by the source; parent not yet tracked).

Similar models

Models similar to VideoLLaMA3-7B-Image
ModelDL 30D
VideoScoreTIGER-Lab59
VideoScore2TIGER-Lab4,008
VideoLLaMA2.1-7B-AVDAMO-NLP-SG2,501
llava-med-v1.5-mistral-7b-hfchaoyinshe1,465
MiniCPM-Llama3-V-2_5-int4openbmb195
MechVL-4B-RLXiaofengAlg756