VibeVoice-ASR-HF
microsoft/VibeVoice-ASR-HF
Tracked by OpenModelStats since August 24, 2026View on Hugging Face ↗
- Downloads 30DRolling recent-window download count as reported by the source platform.
- 85.4K
- Total downloadsCumulative all-time download counter reported by the source platform.
- 2M
- Gained 7DIncrease in cumulative total downloads between OpenModelStats' own observations seven days apart.
- —
- Gained 30DIncrease in cumulative total downloads between OpenModelStats' own observations roughly thirty days apart.
- —
- LikesCommunity likes reported by the source platform.
- 175
- Rank · tracked modelsPosition among all eligible models tracked by OpenModelStats, ordered by recent downloads. #6 among tracked Audio Text To Text models.
- #2,945
- –this week
- Parameters
- 8.3B
- Last source update
- Mar 9, 2026
Downloads over time
Cumulative total downloads observed by OpenModelStatsNot enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.
Likes over time
Not enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.
Rank among tracked models
Lower is betterNot enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.
Overview
VibeVoice-ASR-HF is a audio text to text model published by microsoft. OpenModelStats has tracked the model since Aug 24, 2026, most recently observing it 3h ago.
- Published
- Mar 2, 2026
- Last updated
- Mar 9, 2026
- Library
- transformers
- License
- mit
- Task
- Audio Text To Text
- Parameters
- 8,330,325,888
- Spaces
- —
- Derivative models
- —
- Velocity 7D/day
- —
transformerssafetensorsvibevoice_asrautomatic-speech-recognitionASRDiarizationSpeech-to-TextTranscriptionaudio-text-to-textenzhes
Publisher
microsoftView publisher statistics →Similar models
| Model | DL 30D |
|---|---|
| midashenglm-7b-0804-fp32mispeech | 65.1K |
| midashenglm-7b-1021-bf16mispeech | 1,614 |
| audio-flamingo-3-hfnvidia | 135.1K |
| music-flamingo-2601-hfnvidia | 25.5K |
| audio-flamingo-next-hfnvidia | 11.6K |
| audio-flamingo-next-captioner-hfnvidia | 10.4K |