LLaVA-OneVision-1.5-4B-Instruct

lmms-lab/LLaVA-OneVision-1.5-4B-Instruct

Image Text to Text4.7B parameterstransformersLicense: apache-2.0
Tracked by OpenModelStats since August 24, 2026View on Hugging Face ↗
Downloads 30D
704
Total downloads
35.8K
Gained 7D
Gained 30D
Likes
18
Rank · tracked models
#26,497
this week
Parameters
4.7B
Last source update
Jul 16, 2026

Downloads over time

Cumulative total downloads observed by OpenModelStats

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Likes over time

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Rank among tracked models

Lower is better

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Overview

LLaVA-OneVision-1.5-4B-Instruct is a image text to text model published by lmms-lab. OpenModelStats has tracked the model since Aug 24, 2026, most recently observing it 12h ago.

Published
Sep 26, 2025
Last updated
Jul 16, 2026
Library
transformers
License
apache-2.0
Task
Image Text to Text
Parameters
4,741,610,528
Spaces
Derivative models
Velocity 7D/day
transformerstensorboardsafetensorsfeature-extractionimage-text-to-textconversationalcustom_codedataset:lmms-lab/LLaVA-One-Vision-1.5-Mid-Training-85Mdataset:lmms-lab/LLaVA-OneVision-1.5-Insturct-Datadataset:HuggingFaceM4/FineVisionarxiv:2509.23661base_model:DeepGlint-AI/rice-vit-large-patch14-560

Publisher

lmms-labView publisher statistics →

Derived from DeepGlint-AI/rice-vit-large-patch14-560 (reported by the source; parent not yet tracked).

Similar models

Models similar to LLaVA-OneVision-1.5-4B-Instruct
ModelDL 30D
Mage-VLmicrosoft494.8K
Mage-VL-FP8-W8A8-W8A16ajh-code960
Qianfan-OCRbaidu97.6K
QianfanOCRbairongz1,716
Qwen3.5-4B-AWQ-BF16-INT4cyankiwi1,823
Qwen3.5-4B-AWQ-BF16-INT8cyankiwi1,366