pix2struct-docvqa-large
google/pix2struct-docvqa-large
Tracked by OpenModelStats since September 21, 2026View on Hugging Face ↗
- Downloads 30DRolling recent-window download count as reported by the source platform.
- 444
- Total downloadsCumulative all-time download counter reported by the source platform.
- 39.2K
- Gained 7DIncrease in cumulative total downloads between OpenModelStats' own observations seven days apart.
- —
- Gained 30DIncrease in cumulative total downloads between OpenModelStats' own observations roughly thirty days apart.
- —
- LikesCommunity likes reported by the source platform.
- 36
- Rank · tracked modelsPosition among all eligible models tracked by OpenModelStats, ordered by recent downloads. #32 among tracked Visual Question Answering models.
- #43,381
- 229(down 229 places)this week
- Parameters
- —
- Last source update
- May 19, 2023
- MomentumOpenModelStats' composite rising-model signal (0–100): a percentile blend of absolute 7-day download gains, capped percentage growth, and likes gained. See Methodology.
- 0.0
Downloads over time
Rolling recent-window downloads as reportedLikes over time
Rank among tracked models
Lower is betterOverview
pix2struct-docvqa-large is a visual question answering model published by google. OpenModelStats has tracked the model since Sep 21, 2026, most recently observing it Sep 28, 2026.
- Published
- Mar 21, 2023
- Last updated
- May 19, 2023
- Library
- transformers
- License
- apache-2.0
- Task
- Visual Question Answering
- Parameters
- —
- Spaces
- —
- Derivative models
- —
- Velocity 7D/day
- —
transformerspytorchpix2structimage-text-to-textvisual-question-answeringenfrrodemultilingualarxiv:2210.03347license:apache-2.0
Publisher
googleView publisher statistics →Similar models
| Model | DL 30D |
|---|---|
| blip-vqa-baseSalesforce | 552.8K |
| vilt-b32-finetuned-vqadandelin | 71.3K |
| deplotgoogle | 18.4K |
| Qwen3.5-2B-MedVLOpenMed | 13.9K |
| blip-vqa-capfilt-largeSalesforce | 12.6K |
| MiniCPM-V-2openbmb | 11.4K |