jina-vlm
jinaai/jina-vlm
Tracked by OpenModelStats since August 24, 2026View on Hugging Face ↗
- Downloads 30DRolling recent-window download count as reported by the source platform.
- 1,994
- Total downloadsCumulative all-time download counter reported by the source platform.
- 15.1K
- Gained 7DIncrease in cumulative total downloads between OpenModelStats' own observations seven days apart.
- —
- Gained 30DIncrease in cumulative total downloads between OpenModelStats' own observations roughly thirty days apart.
- —
- LikesCommunity likes reported by the source platform.
- 121
- Rank · tracked modelsPosition among all eligible models tracked by OpenModelStats, ordered by recent downloads. #1821 among tracked Image Text to Text models.
- #17,608
- 15467(down 15467 places)this week
- Parameters
- 2.5B
- Last source update
- Apr 2, 2026
- It moved from #375 to #1821 among tracked Image Text to Text models this week.
Downloads over time
Cumulative total downloads observed by OpenModelStatsNot enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.
Likes over time
Not enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.
Rank among tracked models
Lower is betterNot enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.
Overview
jina-vlm is a image text to text model published by jinaai. OpenModelStats has tracked the model since Aug 24, 2026, most recently observing it 3h ago.
- Published
- Nov 17, 2025
- Last updated
- Apr 2, 2026
- Library
- transformers
- License
- cc-by-nc-4.0
- Task
- Image Text to Text
- Parameters
- 2,481,403,760
- Spaces
- —
- Derivative models
- —
- Velocity 7D/day
- —
transformerssafetensorsjvlmtext-generationmultimodalmultilingualvlmvision-languageqwen3siglip2image-text-to-textconversational
Publisher
jinaaiView publisher statistics →Derived from Qwen/Qwen3-1.7B-Base (reported by the source; parent not yet tracked).
Similar models
| Model | DL 30D |
|---|---|
| North-Micro-Vision-InstructCohereLabs | 25.1K |
| InternVL3_5-2B-FlashOpenGVLab | 447 |
| Ovis2-2BATH-MaaS | 861 |
| UI-TARS-2B-SFTByteDance-Seed | 5,730 |
| Qwen2-VL-2B-Instruct-AWQQwen | 5,549 |
| ToriiGate-v0.4-2BMinthy | 3,543 |