UGC-VideoCaptioner
openinterx/UGC-VideoCaptioner
Tracked by OpenModelStats since September 7, 2026View on Hugging Face ↗
- Downloads 30DRolling recent-window download count as reported by the source platform.
- 1,405
- Total downloadsCumulative all-time download counter reported by the source platform.
- 3,800
- Gained 7DIncrease in cumulative total downloads between OpenModelStats' own observations seven days apart.
- 0
- 0% vs baseline
- Gained 30DIncrease in cumulative total downloads between OpenModelStats' own observations roughly thirty days apart.
- —
- LikesCommunity likes reported by the source platform.
- 4
- +0 in 7D
- Rank · tracked modelsPosition among all eligible models tracked by OpenModelStats, ordered by recent downloads. #18 among tracked Video Text To Text models.
- #28,794
- 119(down 119 places)this week
- Parameters
- 5.5B
- Last source update
- Jul 19, 2025
- MomentumOpenModelStats' composite rising-model signal (0–100): a percentile blend of absolute 7-day download gains, capped percentage growth, and likes gained. See Methodology.
- 4.4
Downloads over time
Cumulative total downloads observed by OpenModelStatsLikes over time
Rank among tracked models
Lower is betterOverview
UGC-VideoCaptioner is a video text to text model published by openinterx. OpenModelStats has tracked the model since Sep 7, 2026, most recently observing it Sep 14, 2026.
- Published
- Jul 9, 2025
- Last updated
- Jul 19, 2025
- Library
- transformers
- License
- mit
- Task
- Video Text To Text
- Parameters
- 5,537,120,672
- Spaces
- —
- Derivative models
- —
- Velocity 7D/day
- 0
transformerssafetensorsqwen2_5_omnitext-to-audiomultimodalvideo-captioningaudio-visualugcvideo-text-to-textdataset:openinterx/UGC-VideoCaparxiv:2507.11336base_model:Qwen/Qwen2.5-Omni-3B
Publisher
openinterxView publisher statistics →Derived from Qwen/Qwen2.5-Omni-3B (reported by the source; parent not yet tracked).
Similar models
| Model | DL 30D |
|---|---|
| Code-as-World-VL-4BMirroS-Lab | 489 |
| Molmo2-VideoPoint-4Ballenai | 199 |
| VideoChat3-4BMCG-NJU | 2,549 |
| TimeLens2-4BMCG-NJU | 1,332 |
| CASA-Qwen2_5-VL-3B-LiveCCkyutai | 2,447 |
| LLaVA-NeXT-Video-7Blmms-lab | 132 |