Video-LLaVA-7B-hf

LanguageBind/Video-LLaVA-7B-hf

Image Text to Text7.4B parameterstransformers
Tracked by OpenModelStats since September 25, 2026View on Hugging Face ↗
Downloads 30D
20.9K
Total downloads
497.4K
Gained 7D
+4,747(increase)
1.0%(up) vs baseline
Gained 30D
—
Likes
51
+0 in 7D
Rank · tracked models
#5,657
23(down 23 places)this week
Parameters
7.4B
Last source update
May 16, 2024
Momentum
59.5
  • Video-LLaVA-7B-hf gained 4,747 downloads during the last seven tracked days.
  • It moved from #706 to #709 among tracked Image Text to Text models this week.

Downloads over time

Cumulative total downloads observed by OpenModelStats

Likes over time

Rank among tracked models

Lower is better

Overview

Video-LLaVA-7B-hf is a image text to text model published by LanguageBind. OpenModelStats has tracked the model since Sep 25, 2026, most recently observing it Oct 1, 2026.

Published
May 9, 2024
Last updated
May 16, 2024
Library
transformers
License
—
Task
Image Text to Text
Parameters
7,366,279,168
Spaces
—
Derivative models
—
Velocity 7D/day
678
transformerssafetensorsvideo_llavaimage-text-to-textarxiv:2311.10122arxiv:2310.01852endpoints_compatibleregion:us

Similar models

Models similar to Video-LLaVA-7B-hf
ModelDL 30D
deepseek-vl-7b-chatdeepseek-ai5,999
deepseek-vl-7b-chatdeepseek-community18.8K
Qwen3.8-27B-INT4-AWQ-GPTQ-gdn4TelperionAI11.8K
ChatRex-7BIDEA-Research6,271
llava-1.5-7b-hf-bnb-4bitunsloth1,487
Ornith-1.5-35B-A3B-AutoRound-W4A16-sym-G128-MTP-BF16SergiioB567
LanguageBind/Video-LLaVA-7B-hf — Downloads, Growth & Stats | OpenModelStats