pix2struct-docvqa-large

google/pix2struct-docvqa-large

Visual Question AnsweringtransformersLicense: apache-2.0
Tracked by OpenModelStats since September 21, 2026View on Hugging Face ↗
Downloads 30D
444
Total downloads
39.2K
Gained 7D
—
Gained 30D
—
Likes
36
Rank · tracked models
#43,381
229(down 229 places)this week
Parameters
—
Last source update
May 19, 2023
Momentum
0.0

Downloads over time

Daily gain between consecutive observations

Likes over time

Rank among tracked models

Lower is better

Overview

pix2struct-docvqa-large is a visual question answering model published by google. OpenModelStats has tracked the model since Sep 21, 2026, most recently observing it Sep 28, 2026.

Published
Mar 21, 2023
Last updated
May 19, 2023
Library
transformers
License
apache-2.0
Task
Visual Question Answering
Parameters
—
Spaces
—
Derivative models
—
Velocity 7D/day
—
transformerspytorchpix2structimage-text-to-textvisual-question-answeringenfrrodemultilingualarxiv:2210.03347license:apache-2.0

Similar models

Models similar to pix2struct-docvqa-large
ModelDL 30D
blip-vqa-baseSalesforce552.8K
vilt-b32-finetuned-vqadandelin71.3K
deplotgoogle18.4K
Qwen3.5-2B-MedVLOpenMed13.9K
blip-vqa-capfilt-largeSalesforce12.6K
MiniCPM-V-2openbmb11.4K
google/pix2struct-docvqa-large — Downloads, Growth & Stats | OpenModelStats