Sa2VA-Qwen2_5-VL-7B

ByteDance/Sa2VA-Qwen2_5-VL-7B

Image Text to Text8.5B parameterstransformersLicense: apache-2.0
Tracked by OpenModelStats since September 25, 2026View on Hugging Face ↗
Downloads 30D
443
Total downloads
1,868
Gained 7D
—
Gained 30D
—
Likes
5
Rank · tracked models
#44,255
252(down 252 places)this week
Parameters
8.5B
Last source update
Oct 16, 2025
  • It moved from #4192 to #4215 among tracked Image Text to Text models this week.

Downloads over time

Daily gain between consecutive observations

Likes over time

Rank among tracked models

Lower is better

Overview

Sa2VA-Qwen2_5-VL-7B is a image text to text model published by ByteDance. OpenModelStats has tracked the model since Sep 25, 2026, most recently observing it 30h ago.

Published
Oct 16, 2025
Last updated
Oct 16, 2025
Library
transformers
License
apache-2.0
Task
Image Text to Text
Parameters
8,527,538,994
Spaces
—
Derivative models
—
Velocity 7D/day
—
transformerssafetensorssa2va_chatfeature-extractionSa2VAcustom_codeimage-text-to-textconversationalmultilingualarxiv:2501.04001base_model:OpenGVLab/InternVL3-8Bbase_model:merge:OpenGVLab/InternVL3-8B

Publisher

ByteDanceView publisher statistics →

Derived from OpenGVLab/InternVL3-8B (reported by the source; parent not yet tracked).

Similar models

Models similar to Sa2VA-Qwen2_5-VL-7B
ModelDL 30D
LLaVA-OneVision-1.5-8B-Instructlmms-lab10.9K
LLaVA-OneVision-2-8B-Instructlmms-lab-encoder8,763
InternVL3_5-8BOpenGVLab40.6K
InternVL3_5-8B-HFOpenGVLab30K
InternVL3_5-8B-InstructOpenGVLab2,695
CapRL-InternVL3.5-8Binternlm402
ByteDance/Sa2VA-Qwen2_5-VL-7B — Downloads, Growth & Stats | OpenModelStats