LLaVA-OneVision-1.5-4B-Instruct

lmms-lab/LLaVA-OneVision-1.5-4B-Instruct

Image Text to Text4.7B parameterstransformersLicense: apache-2.0
Tracked by OpenModelStats since September 3, 2026View on Hugging Face ↗
Downloads 30D
1,000
Total downloads
37K
Gained 7D
+87(increase)
0.2%(up) vs baseline
Gained 30D
—
Likes
18
+0 in 7D
Rank · tracked models
#33,543
501(up 501 places)this week
Parameters
4.7B
Last source update
Jul 16, 2026
Momentum
29.5
  • LLaVA-OneVision-1.5-4B-Instruct gained 87 downloads during the last seven tracked days.
  • It moved from #3178 to #3149 among tracked Image Text to Text models this week.

Downloads over time

Rolling recent-window downloads as reported

Likes over time

Rank among tracked models

Lower is better

Overview

LLaVA-OneVision-1.5-4B-Instruct is a image text to text model published by lmms-lab. OpenModelStats has tracked the model since Sep 3, 2026, most recently observing it Oct 2, 2026.

Published
Sep 26, 2025
Last updated
Jul 16, 2026
Library
transformers
License
apache-2.0
Task
Image Text to Text
Parameters
4,741,610,528
Spaces
—
Derivative models
—
Velocity 7D/day
12
transformerstensorboardsafetensorsfeature-extractionimage-text-to-textconversationalcustom_codedataset:lmms-lab/LLaVA-One-Vision-1.5-Mid-Training-85Mdataset:lmms-lab/LLaVA-OneVision-1.5-Insturct-Datadataset:HuggingFaceM4/FineVisionarxiv:2509.23661base_model:DeepGlint-AI/rice-vit-large-patch14-560

Publisher

lmms-labView publisher statistics →

Derived from DeepGlint-AI/rice-vit-large-patch14-560 (reported by the source; parent not yet tracked).

Similar models

Models similar to LLaVA-OneVision-1.5-4B-Instruct
ModelDL 30D
Mage-VLmicrosoft11.5K
Mage-VL-8bitmlx-community245
Qianfan-OCRbaidu248.3K
QianfanOCRbairongz1,120
Qwen3.5-4B-AWQ-BF16-INT8cyankiwi373
Qwen3.5-4B-AWQ-BF16-INT4cyankiwi345
lmms-lab/LLaVA-OneVision-1.5-4B-Instruct — Downloads, Growth & Stats | OpenModelStats