visionlanguageTransformer
Joe99/visionlanguageTransformer
Tracked by OpenModelStats since October 6, 2026View on Hugging Face ↗
- Downloads 30DRolling recent-window download count as reported by the source platform.
- 13
- Total downloadsCumulative all-time download counter reported by the source platform.
- 150
- Gained 7DIncrease in cumulative total downloads between OpenModelStats' own observations seven days apart.
- —
- Gained 30DIncrease in cumulative total downloads between OpenModelStats' own observations roughly thirty days apart.
- —
- LikesCommunity likes reported by the source platform.
- 1
- Rank · tracked modelsPosition among all eligible models tracked by OpenModelStats, ordered by recent downloads.
- —
- –this week
- Parameters
- —
- Last source update
- Aug 15, 2023
Downloads over time
Rolling recent-window downloads as reportedLikes over time
Rank among tracked models
Lower is betterNot enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-10-06. Charts appear as daily observations accumulate.
Overview
visionlanguageTransformer is a visual question answering model published by Joe99. OpenModelStats has tracked the model since Oct 6, 2026, most recently observing it 2h ago.
- Published
- Aug 15, 2023
- Last updated
- Aug 15, 2023
- Library
- transformers
- License
- apache-2.0
- Task
- Visual Question Answering
- Parameters
- —
- Spaces
- —
- Derivative models
- —
- Velocity 7D/day
- —
transformerspytorchviltvisual-question-answeringenarxiv:2102.03334license:apache-2.0endpoints_compatibleregion:us
Publisher
Joe99View publisher statistics →Similar models
| Model | DL 30D |
|---|---|
| blip-vqa-baseSalesforce | 500.7K |
| vilt-b32-finetuned-vqadandelin | 75.4K |
| Qwen3.5-2B-MedVLOpenMed | 13.7K |
| blip-vqa-capfilt-largeSalesforce | 11.9K |
| deplotgoogle | 11.5K |
| internlm-xcomposer2-vl-7binternlm | 10.7K |