VideoLLaMA2.1-7B-AV

DAMO-NLP-SG/VideoLLaMA2.1-7B-AV

Visual Question Answering8.5B parameterstransformersLicense: apache-2.0
Tracked by OpenModelStats since August 24, 2026View on Hugging Face ↗
Downloads 30D
1,303
Total downloads
135.7K
Gained 7D
Gained 30D
Likes
16
Rank · tracked models
#22,379
this week
Parameters
8.5B
Last source update
Oct 25, 2024

Downloads over time

Cumulative total downloads observed by OpenModelStats

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Likes over time

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Rank among tracked models

Lower is better

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Overview

VideoLLaMA2.1-7B-AV is a visual question answering model published by DAMO-NLP-SG. OpenModelStats has tracked the model since Aug 24, 2026, most recently observing it 8h ago.

Published
Oct 21, 2024
Last updated
Oct 25, 2024
Library
transformers
License
apache-2.0
Task
Visual Question Answering
Parameters
8,527,095,391
Spaces
Derivative models
Velocity 7D/day
transformerssafetensorsvideollama2_qwen2text-generationAudio-visual Question AnsweringAudio Question Answeringmultimodal large language modelvisual-question-answeringendataset:lmms-lab/ClothoAQAdataset:Loie/VGGSoundarxiv:2406.07476

Similar models