vit-gpt2-image-captioning

nlpconnect/vit-gpt2-image-captioning

Image to TexttransformersLicense: apache-2.0
Tracked by OpenModelStats since August 24, 2026View on Hugging Face ↗
Downloads 30D
107.6K
Total downloads
60.3M
Gained 7D
Gained 30D
Likes
933
Rank · tracked models
#2,362
this week
Parameters
Last source update
Feb 27, 2023

Downloads over time

Cumulative total downloads observed by OpenModelStats

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Likes over time

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Rank among tracked models

Lower is better

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Overview

vit-gpt2-image-captioning is a image to text model published by nlpconnect. OpenModelStats has tracked the model since Aug 24, 2026, most recently observing it 1 min ago.

Published
Mar 2, 2022
Last updated
Feb 27, 2023
Library
transformers
License
apache-2.0
Task
Image to Text
Parameters
Spaces
100
Derivative models
Velocity 7D/day
transformerspytorchvision-encoder-decoderimage-text-to-textimage-to-textimage-captioningdoi:10.57967/hf/0222license:apache-2.0endpoints_compatibleregion:us

Similar models

Models similar to vit-gpt2-image-captioning
ModelDL 30D
blip-image-captioning-baseSalesforce2.1M
PP-OCRv5_server_detPaddlePaddle768.2K
manga-ocr-basekha-white659.4K
en_PP-OCRv5_mobile_recPaddlePaddle524.9K
trocr-small-handwrittenmicrosoft489.2K
blip-image-captioning-largeSalesforce484.2K