ScreenVLM

docling-project/ScreenVLM

Image Text to Text315M parameterstransformersLicense: apache-2.0
Tracked by OpenModelStats since August 25, 2026View on Hugging Face ↗
Downloads 30D
990
Total downloads
2,828
Gained 7D
Gained 30D
Likes
6
Rank · tracked models
this week
Parameters
315M
Last source update
May 29, 2026

Downloads over time

Cumulative total downloads observed by OpenModelStats

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-25. Charts appear as daily observations accumulate.

Likes over time

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-25. Charts appear as daily observations accumulate.

Rank among tracked models

Lower is better

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-25. Charts appear as daily observations accumulate.

Overview

ScreenVLM is a image text to text model published by docling-project. OpenModelStats has tracked the model since Aug 25, 2026, most recently observing it Aug 25, 2026.

Published
Feb 11, 2026
Last updated
May 29, 2026
Library
transformers
License
apache-2.0
Task
Image Text to Text
Parameters
315,319,872
Spaces
Derivative models
Velocity 7D/day
transformerssafetensorsidefics3image-text-to-texttext-generationscreen-parsingui-understandingobject-detectiongroundingwebscreentagdocling

Similar models

Models similar to ScreenVLM
ModelDL 30D
granite-docling-258M-mlxibm-granite3,049
Qwen3.5-0.8B-8bitmlx-community4,680
Qwen3.5-0.8B-MLX-8bitlmstudio-community3,336
OvisOCR2-8bitmlx-community427
texifyvikp58.8K
Qwen3-VL-2B-Instructmobilint1,259