VideoChat-R1_7B_caption

OpenGVLab/VideoChat-R1_7B_caption

Video Text To Text8.3B parameterstransformersLicense: apache-2.0
Tracked by OpenModelStats since September 6, 2026View on Hugging Face ↗
Downloads 30D
44
Total downloads
12.3K
Gained 7D
0
0% vs baseline
Gained 30D
—
Likes
6
+0 in 7D
Rank · tracked models
#54,061
459(down 459 places)this week
Parameters
8.3B
Last source update
Apr 22, 2025
Momentum
4.4
  • It moved from #54 to #56 among tracked Video Text To Text models this week.

Downloads over time

Rolling recent-window downloads as reported

Likes over time

Rank among tracked models

Lower is better

Overview

VideoChat-R1_7B_caption is a video text to text model published by OpenGVLab. OpenModelStats has tracked the model since Sep 6, 2026, most recently observing it Sep 13, 2026.

Published
Apr 22, 2025
Last updated
Apr 22, 2025
Library
transformers
License
apache-2.0
Task
Video Text To Text
Parameters
8,291,375,616
Spaces
—
Derivative models
—
Velocity 7D/day
0
transformerssafetensorsqwen2_vlimage-text-to-textmultimodalvideo-text-to-textenarxiv:2504.06958base_model:Qwen/Qwen2-VL-7B-Instructbase_model:finetune:Qwen/Qwen2-VL-7B-Instructlicense:apache-2.0text-generation-inference

Publisher

OpenGVLabView publisher statistics →

Derived from Qwen/Qwen2-VL-7B-Instruct (reported by the source; parent not yet tracked).

Similar models

Models similar to VideoChat-R1_7B_caption
ModelDL 30D
TimeLens-7BTencentARC924
SkyCaptioner-V1Skywork390
qwen2.5-vl-7b-cam-motionchancharikm259
VideoScore-v1.1TIGER-Lab1,029
VITA-1.5VITA-MLLM382
InternVideo2_5_Chat_8BOpenGVLab4,099
OpenGVLab/VideoChat-R1_7B_caption — Downloads, Growth & Stats | OpenModelStats