Sa2VA-Qwen3-VL-4B-SAM3
ByteDance/Sa2VA-Qwen3-VL-4B-SAM3
Tracked by OpenModelStats since September 4, 2026View on Hugging Face ↗
- Downloads 30DRolling recent-window download count as reported by the source platform.
- 812
- Total downloadsCumulative all-time download counter reported by the source platform.
- 1,535
- Gained 7DIncrease in cumulative total downloads between OpenModelStats' own observations seven days apart.
- +76(increase)
- Gained 30DIncrease in cumulative total downloads between OpenModelStats' own observations roughly thirty days apart.
- —
- LikesCommunity likes reported by the source platform.
- 2
- +0 in 7D
- Rank · tracked modelsPosition among all eligible models tracked by OpenModelStats, ordered by recent downloads. #3354 among tracked Image Text to Text models.
- #35,471
- 1175(down 1175 places)this week
- Parameters
- 5.3B
- Last source update
- Jul 28, 2026
- MomentumOpenModelStats' composite rising-model signal (0–100): a percentile blend of absolute 7-day download gains, capped percentage growth, and likes gained. See Methodology.
- 19.0
- Sa2VA-Qwen3-VL-4B-SAM3 gained 76 downloads during the last seven tracked days.
- It moved from #3216 to #3354 among tracked Image Text to Text models this week.
Downloads over time
Cumulative total downloads observed by OpenModelStatsLikes over time
Rank among tracked models
Lower is betterOverview
Sa2VA-Qwen3-VL-4B-SAM3 is a image text to text model published by ByteDance. OpenModelStats has tracked the model since Sep 4, 2026, most recently observing it Oct 4, 2026.
- Published
- Jun 10, 2026
- Last updated
- Jul 28, 2026
- Library
- transformers
- License
- apache-2.0
- Task
- Image Text to Text
- Parameters
- 5,306,227,586
- Spaces
- —
- Derivative models
- —
- Velocity 7D/day
- 11
transformerssafetensorssa2va_chatfeature-extractionsa2vareferring-segmentationreferring-video-segmentationqwen3-vlsam3image-text-to-textconversationalcustom_code
Publisher
ByteDanceView publisher statistics →Similar models
| Model | DL 30D |
|---|---|
| chandra-ocr-2datalab-to | 2M |
| NV-Reason-CTnvidia | 2,701 |
| Youtu-VL-4B-Instructtencent | 1,460 |
| Qwen3.8-Flash-Next-4bit-pagedGreenBitAI | 1,067 |
| llama-3.2-Korean-Bllossom-AICA-5BBllossom | 194 |
| gemma-3n-E2B-itgoogle | 174.7K |