vit-gpt2-image-captioning

nlpconnect/vit-gpt2-image-captioning

Image to TexttransformersLicense: apache-2.0
Tracked by OpenModelStats since August 28, 2026View on Hugging Face ↗
Downloads 30D
83.5K
Total downloads
60.4M
Gained 7D
+19.8K(increase)
0.0%(up) vs baseline
Gained 30D
—
Likes
935
+0 in 7D
Rank · tracked models
#2,989
14(down 14 places)this week
Parameters
—
Last source update
Feb 27, 2023
Momentum
56.1
  • vit-gpt2-image-captioning gained 19.8K downloads during the last seven tracked days.

Downloads over time

Daily gain between consecutive observations

Likes over time

Rank among tracked models

Lower is better

Overview

vit-gpt2-image-captioning is a image to text model published by nlpconnect. OpenModelStats has tracked the model since Aug 28, 2026, most recently observing it Sep 27, 2026.

Published
Mar 2, 2022
Last updated
Feb 27, 2023
Library
transformers
License
apache-2.0
Task
Image to Text
Parameters
—
Spaces
—
Derivative models
1
Velocity 7D/day
2,828
transformerspytorchvision-encoder-decoderimage-text-to-textimage-to-textimage-captioningdoi:10.57967/hf/0222license:apache-2.0endpoints_compatibleregion:us

Model family

Based on base-model relationships reported by the source's metadata.

Derivative models

Similar models

Models similar to vit-gpt2-image-captioning
ModelDL 30D
GLM-OCRzai-org1.8M
blip-image-captioning-baseSalesforce1.6M
manga-ocr-basekha-white1.1M
PP-OCRv5_server_detPaddlePaddle610.6K
blip-image-captioning-largeSalesforce574.1K
UVDocPaddlePaddle476K
nlpconnect/vit-gpt2-image-captioning — Downloads, Growth & Stats | OpenModelStats