prismatic-vlms

TRI-ML/prismatic-vlms

Image to TextLicense: mit
Tracked by OpenModelStats since August 24, 2026View on Hugging Face ↗
Downloads 30D
0
Total downloads
0
Gained 7D
Gained 30D
Likes
29
Rank · tracked models
#32,700
29450(down 29450 places)this week
Parameters
Last source update
May 6, 2024
  • It moved from #14 to #167 among tracked Image to Text models this week.

Downloads over time

Rolling recent-window downloads as reported

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Likes over time

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Rank among tracked models

Lower is better

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Overview

prismatic-vlms is a image to text model published by TRI-ML. OpenModelStats has tracked the model since Aug 24, 2026, most recently observing it 2h ago.

Published
Feb 13, 2024
Last updated
May 6, 2024
Library
License
mit
Task
Image to Text
Parameters
Spaces
Derivative models
Velocity 7D/day
computer visionnatural language processingvision language modelsmultimodal modelsimage-to-textenarxiv:2402.07865license:mitregion:us

Similar models

Models similar to prismatic-vlms
ModelDL 30D
blip-image-captioning-baseSalesforce2.1M
PP-OCRv5_server_detPaddlePaddle768.2K
manga-ocr-basekha-white659.4K
en_PP-OCRv5_mobile_recPaddlePaddle524.9K
trocr-small-handwrittenmicrosoft489.2K
blip-image-captioning-largeSalesforce484.2K