Sa2VA-Qwen3-VL-4B-SAM3

ByteDance/Sa2VA-Qwen3-VL-4B-SAM3

Image Text to Text5.3B parameterstransformersLicense: apache-2.0
Tracked by OpenModelStats since September 6, 2026View on Hugging Face ↗
Downloads 30D
835
Total downloads
1,597
Gained 7D
+76(increase)
Gained 30D
—
Likes
2
+0 in 7D
Rank · tracked models
#35,088
339(up 339 places)this week
Parameters
5.3B
Last source update
Jul 28, 2026
Momentum
19.0
  • Sa2VA-Qwen3-VL-4B-SAM3 gained 76 downloads during the last seven tracked days.
  • It moved from #3347 to #3310 among tracked Image Text to Text models this week.

Downloads over time

Rolling recent-window downloads as reported

Likes over time

Rank among tracked models

Lower is better

Overview

Sa2VA-Qwen3-VL-4B-SAM3 is a image text to text model published by ByteDance. OpenModelStats has tracked the model since Sep 6, 2026, most recently observing it Oct 6, 2026.

Published
Jun 10, 2026
Last updated
Jul 28, 2026
Library
transformers
License
apache-2.0
Task
Image Text to Text
Parameters
5,306,227,586
Spaces
—
Derivative models
—
Velocity 7D/day
11
transformerssafetensorssa2va_chatfeature-extractionsa2vareferring-segmentationreferring-video-segmentationqwen3-vlsam3image-text-to-textconversationalcustom_code

Similar models

Models similar to Sa2VA-Qwen3-VL-4B-SAM3
ModelDL 30D
chandra-ocr-2datalab-to2M
NV-Reason-CTnvidia3,102
Youtu-VL-4B-Instructtencent1,534
Qwen3.8-Flash-Next-4bit-pagedGreenBitAI1,067
llama-3.2-Korean-Bllossom-AICA-5BBllossom199
gemma-3n-E2B-itgoogle174K
ByteDance/Sa2VA-Qwen3-VL-4B-SAM3 — Downloads, Growth & Stats | OpenModelStats