Video-LLaVA-7B-hf

LanguageBind/Video-LLaVA-7B-hf

Image Text to Text7.4B parameterstransformers
Tracked by OpenModelStats since September 4, 2026View on Hugging Face ↗
Downloads 30D
20.1K
Total downloads
498.4K
Gained 7D
+4,747(increase)
1.0%(up) vs baseline
Gained 30D
—
Likes
51
+0 in 7D
Rank · tracked models
#5,748
67(down 67 places)this week
Parameters
7.4B
Last source update
May 16, 2024
Momentum
59.5
  • Video-LLaVA-7B-hf gained 4,747 downloads during the last seven tracked days.
  • It moved from #712 to #727 among tracked Image Text to Text models this week.

Downloads over time

Cumulative total downloads observed by OpenModelStats

Likes over time

Rank among tracked models

Lower is better

Overview

Video-LLaVA-7B-hf is a image text to text model published by LanguageBind. OpenModelStats has tracked the model since Sep 4, 2026, most recently observing it Oct 3, 2026.

Published
May 9, 2024
Last updated
May 16, 2024
Library
transformers
License
—
Task
Image Text to Text
Parameters
7,366,279,168
Spaces
—
Derivative models
—
Velocity 7D/day
678
transformerssafetensorsvideo_llavaimage-text-to-textarxiv:2311.10122arxiv:2310.01852endpoints_compatibleregion:us

Similar models

Models similar to Video-LLaVA-7B-hf
ModelDL 30D
deepseek-vl-7b-chatdeepseek-ai5,984
deepseek-vl-7b-chatdeepseek-community18.7K
Qwen3.8-27B-INT4-AWQ-GPTQ-gdn4TelperionAI11.8K
ChatRex-7BIDEA-Research1,003
llava-1.5-7b-hf-bnb-4bitunsloth1,325
Ornith-1.5-35B-A3B-AutoRound-W4A16-sym-G128-MTP-BF16SergiioB450
LanguageBind/Video-LLaVA-7B-hf — Downloads, Growth & Stats | OpenModelStats