VideoLLaMA3-7B-Image

DAMO-NLP-SG/VideoLLaMA3-7B-Image

Visual Question Answering8B parameterstransformersLicense: apache-2.0
Tracked by OpenModelStats since October 7, 2026View on Hugging Face ↗
Downloads 30D
146
Total downloads
23K
Gained 7D
—
Gained 30D
—
Likes
11
Rank · tracked models
—
–this week
Parameters
8B
Last source update
Mar 20, 2025

Downloads over time

Cumulative total downloads observed by OpenModelStats

Likes over time

Rank among tracked models

Lower is better

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-10-07. Charts appear as daily observations accumulate.

Overview

VideoLLaMA3-7B-Image is a visual question answering model published by DAMO-NLP-SG. OpenModelStats has tracked the model since Oct 7, 2026, most recently observing it 3h ago.

Published
Jan 21, 2025
Last updated
Mar 20, 2025
Library
transformers
License
apache-2.0
Task
Visual Question Answering
Parameters
8,044,744,944
Spaces
—
Derivative models
—
Velocity 7D/day
—
transformerssafetensorsvideollama3_qwen2text-generationmulti-modallarge-language-modelvideo-language-modelvisual-question-answeringcustom_codeendataset:lmms-lab/LLaVA-OneVision-Datadataset:allenai/pixmo-docs

Publisher

DAMO-NLP-SGView publisher statistics →

Derived from Qwen/Qwen2.5-7B-Instruct (reported by the source; parent not yet tracked).

Similar models

Models similar to VideoLLaMA3-7B-Image
ModelDL 30D
VideoScoreTIGER-Lab59
VideoScore2TIGER-Lab4,008
VideoLLaMA2.1-7B-AVDAMO-NLP-SG2,501
llava-med-v1.5-mistral-7b-hfchaoyinshe1,465
MiniCPM-Llama3-V-2_5-int4openbmb195
MechVL-4B-RLXiaofengAlg756