VITA-1.5

VITA-MLLM/VITA-1.5

Video Text To Text8.3B parameters
Tracked by OpenModelStats since August 24, 2026View on Hugging Face ↗
Downloads 30D
10.3K
Total downloads
16.9K
Gained 7D
Gained 30D
Likes
50
Rank · tracked models
#7,698
this week
Parameters
8.3B
Last source update
Jan 16, 2025

Downloads over time

Cumulative total downloads observed by OpenModelStats

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Likes over time

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Rank among tracked models

Lower is better

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Overview

VITA-1.5 is a video text to text model published by VITA-MLLM. OpenModelStats has tracked the model since Aug 24, 2026, most recently observing it 42 min ago.

Published
Dec 18, 2024
Last updated
Jan 16, 2025
Library
License
Task
Video Text To Text
Parameters
8,322,947,232
Spaces
Derivative models
Velocity 7D/day
safetensorsvita-Qwen2video-text-to-textarxiv:2501.01957region:us

Similar models

Models similar to VITA-1.5
ModelDL 30D
qwen2.5-vl-7b-cam-motionchancharikm1,039
SkyCaptioner-V1Skywork169
VideoScore-v1.1TIGER-Lab1,325
InternVideo2_5_Chat_8BOpenGVLab2,742
Video-XL-2BAAI348
VideoLLaMA3-7BDAMO-NLP-SG6,983