Sa2VA-Qwen3-VL-4B-SAM3

ByteDance/Sa2VA-Qwen3-VL-4B-SAM3

Image Text to Text5.3B parameterstransformersLicense: apache-2.0
Tracked by OpenModelStats since September 9, 2026View on Hugging Face ↗
Downloads 30D
837
Total downloads
1,630
Gained 7D
+76(increase)
Gained 30D
—
Likes
2
+0 in 7D
Rank · tracked models
#35,887
31(down 31 places)this week
Parameters
5.3B
Last source update
Jul 28, 2026
Momentum
19.0
  • Sa2VA-Qwen3-VL-4B-SAM3 gained 76 downloads during the last seven tracked days.
  • It moved from #3390 to #3399 among tracked Image Text to Text models this week.

Downloads over time

Cumulative total downloads observed by OpenModelStats

Likes over time

Rank among tracked models

Lower is better

Overview

Sa2VA-Qwen3-VL-4B-SAM3 is a image text to text model published by ByteDance. OpenModelStats has tracked the model since Sep 9, 2026, most recently observing it 1h ago.

Published
Jun 10, 2026
Last updated
Jul 28, 2026
Library
transformers
License
apache-2.0
Task
Image Text to Text
Parameters
5,306,227,586
Spaces
—
Derivative models
—
Velocity 7D/day
11
transformerssafetensorssa2va_chatfeature-extractionsa2vareferring-segmentationreferring-video-segmentationqwen3-vlsam3image-text-to-textconversationalcustom_code

Similar models

Models similar to Sa2VA-Qwen3-VL-4B-SAM3
ModelDL 30D
chandra-ocr-2datalab-to1.9M
NV-Reason-CTnvidia3,705
Youtu-VL-4B-Instructtencent1,545
Qwen3.8-Flash-Next-4bit-pagedGreenBitAI1,067
llama-3.2-Korean-Bllossom-AICA-5BBllossom202
gemma-3n-E2B-itgoogle170.8K
ByteDance/Sa2VA-Qwen3-VL-4B-SAM3 — Downloads, Growth & Stats | OpenModelStats