Qwen3-VisionCaption-2B

prithivMLmods/Qwen3-VisionCaption-2B

Image Text to Text2.1B parameterstransformersLicense: apache-2.0
Tracked by OpenModelStats since October 3, 2026View on Hugging Face ↗
Downloads 30D
70
Total downloads
654
Gained 7D
—
Gained 30D
—
Likes
6
Rank · tracked models
#55,923
654(down 654 places)this week
Parameters
2.1B
Last source update
Nov 29, 2025
  • It moved from #5055 to #5086 among tracked Image Text to Text models this week.

Downloads over time

Daily gain between consecutive observations

Likes over time

Rank among tracked models

Lower is better

Overview

Qwen3-VisionCaption-2B is a image text to text model published by prithivMLmods. OpenModelStats has tracked the model since Oct 3, 2026, most recently observing it 3h ago.

Published
Nov 28, 2025
Last updated
Nov 29, 2025
Library
transformers
License
apache-2.0
Task
Image Text to Text
Parameters
2,127,532,032
Spaces
—
Derivative models
—
Velocity 7D/day
—
transformerssafetensorsqwen3_vlimage-text-to-texttext-generation-inferenceimage-captionabliterateduncensoredllama.cppconversationalenzh

Publisher

prithivMLmodsView publisher statistics →

Derived from prithivMLmods/Qwen3-VL-2B-Instruct-abliterated-v1 (reported by the source; parent not yet tracked).

Similar models

Models similar to Qwen3-VisionCaption-2B
ModelDL 30D
Qwen3-VL-2B-InstructQwen2.7M
typhoon-ocr1.5-2btyphoon-ai349K
Qwen3-VL-2B-ThinkingQwen114.9K
Qwen3-VL-2B-Instructunsloth6,741
Qwen3-VL-2B-Instruct-4bitmlx-community3,259
Qwen3-VL-2B-Instructfuriosa-ai2,374