jina-vlm

jinaai/jina-vlm

Image Text to Text2.5B parameterstransformersLicense: cc-by-nc-4.0
Tracked by OpenModelStats since August 24, 2026View on Hugging Face ↗
Downloads 30D
1,994
Total downloads
15.1K
Gained 7D
Gained 30D
Likes
121
Rank · tracked models
#17,608
15467(down 15467 places)this week
Parameters
2.5B
Last source update
Apr 2, 2026
  • It moved from #375 to #1821 among tracked Image Text to Text models this week.

Downloads over time

Rolling recent-window downloads as reported

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Likes over time

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Rank among tracked models

Lower is better

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Overview

jina-vlm is a image text to text model published by jinaai. OpenModelStats has tracked the model since Aug 24, 2026, most recently observing it 2h ago.

Published
Nov 17, 2025
Last updated
Apr 2, 2026
Library
transformers
License
cc-by-nc-4.0
Task
Image Text to Text
Parameters
2,481,403,760
Spaces
Derivative models
Velocity 7D/day
transformerssafetensorsjvlmtext-generationmultimodalmultilingualvlmvision-languageqwen3siglip2image-text-to-textconversational

Publisher

jinaaiView publisher statistics →

Derived from Qwen/Qwen3-1.7B-Base (reported by the source; parent not yet tracked).

Similar models

Models similar to jina-vlm
ModelDL 30D
North-Micro-Vision-InstructCohereLabs25.1K
InternVL3_5-2B-FlashOpenGVLab447
Ovis2-2BATH-MaaS861
UI-TARS-2B-SFTByteDance-Seed5,730
Qwen2-VL-2B-Instruct-AWQQwen5,549
ToriiGate-v0.4-2BMinthy3,543