Sa2VA-Qwen3-VL-4B-SAM3

ByteDance/Sa2VA-Qwen3-VL-4B-SAM3

Image Text to Text5.3B parameterstransformersLicense: apache-2.0
Tracked by OpenModelStats since September 4, 2026View on Hugging Face ↗
Downloads 30D
812
Total downloads
1,535
Gained 7D
+76(increase)
Gained 30D
—
Likes
2
+0 in 7D
Rank · tracked models
#35,471
1175(down 1175 places)this week
Parameters
5.3B
Last source update
Jul 28, 2026
Momentum
19.0
  • Sa2VA-Qwen3-VL-4B-SAM3 gained 76 downloads during the last seven tracked days.
  • It moved from #3216 to #3354 among tracked Image Text to Text models this week.

Downloads over time

Cumulative total downloads observed by OpenModelStats

Likes over time

Rank among tracked models

Lower is better

Overview

Sa2VA-Qwen3-VL-4B-SAM3 is a image text to text model published by ByteDance. OpenModelStats has tracked the model since Sep 4, 2026, most recently observing it Oct 4, 2026.

Published
Jun 10, 2026
Last updated
Jul 28, 2026
Library
transformers
License
apache-2.0
Task
Image Text to Text
Parameters
5,306,227,586
Spaces
—
Derivative models
—
Velocity 7D/day
11
transformerssafetensorssa2va_chatfeature-extractionsa2vareferring-segmentationreferring-video-segmentationqwen3-vlsam3image-text-to-textconversationalcustom_code

Similar models

Models similar to Sa2VA-Qwen3-VL-4B-SAM3
ModelDL 30D
chandra-ocr-2datalab-to2M
NV-Reason-CTnvidia2,701
Youtu-VL-4B-Instructtencent1,460
Qwen3.8-Flash-Next-4bit-pagedGreenBitAI1,067
llama-3.2-Korean-Bllossom-AICA-5BBllossom194
gemma-3n-E2B-itgoogle174.7K