VideoLLaMA2.1-7B-16F-Base

DAMO-NLP-SG/VideoLLaMA2.1-7B-16F-Base

Visual Question AnsweringtransformersLicense: apache-2.0
Tracked by OpenModelStats since August 24, 2026View on Hugging Face ↗
Downloads 30D
41
Total downloads
4,588
Gained 7D
+8(increase)
0.2%(up) vs baseline
Gained 30D
—
Likes
1
+0 in 7D
Rank · tracked models
#17,257
125(up 125 places)this week
Parameters
—
Last source update
Oct 21, 2024
  • VideoLLaMA2.1-7B-16F-Base gained 8 downloads during the last seven tracked days.

Downloads over time

Cumulative total downloads observed by OpenModelStats

Likes over time

Rank among tracked models

Lower is better

Overview

VideoLLaMA2.1-7B-16F-Base is a visual question answering model published by DAMO-NLP-SG. OpenModelStats has tracked the model since Aug 24, 2026, most recently observing it Oct 1, 2026.

Published
Oct 14, 2024
Last updated
Oct 21, 2024
Library
transformers
License
apache-2.0
Task
Visual Question Answering
Parameters
—
Spaces
0
Derivative models
—
Velocity 7D/day
1
transformersvideollama2_qwen2text-generationmultimodal large language modellarge video-language modelvisual-question-answeringendataset:OpenGVLab/VideoChat2-ITdataset:Lin-Chen/ShareGPT4Vdataset:liuhaotian/LLaVA-Instruct-150Karxiv:2406.07476arxiv:2306.02858

Similar models

Models similar to VideoLLaMA2.1-7B-16F-Base
ModelDL 30D
blip-vqa-baseSalesforce539.2K
vilt-b32-finetuned-vqadandelin71.4K
deplotgoogle17.3K
Qwen3.5-2B-MedVLOpenMed13.9K
blip-vqa-capfilt-largeSalesforce12.7K
MiniCPM-V-4-GGUFsecond-state9,399
DAMO-NLP-SG/VideoLLaMA2.1-7B-16F-Base — Downloads, Growth & Stats | OpenModelStats