VibeVoice-ASR-HF

microsoft/VibeVoice-ASR-HF

Audio Text To Text8.3B parameterstransformersLicense: mit
Tracked by OpenModelStats since September 6, 2026View on Hugging Face ↗
Downloads 30D
69K
Total downloads
2.1M
Gained 7D
+15.7K(increase)
0.8%(up) vs baseline
Gained 30D
—
Likes
182
+0 in 7D
Rank · tracked models
#3,539
32(up 32 places)this week
Parameters
8.3B
Last source update
Mar 9, 2026
Momentum
60.9
  • VibeVoice-ASR-HF gained 15.7K downloads during the last seven tracked days.

Downloads over time

Cumulative total downloads observed by OpenModelStats

Likes over time

Rank among tracked models

Lower is better

Overview

VibeVoice-ASR-HF is a audio text to text model published by microsoft. OpenModelStats has tracked the model since Sep 6, 2026, most recently observing it Oct 6, 2026.

Published
Mar 2, 2026
Last updated
Mar 9, 2026
Library
transformers
License
mit
Task
Audio Text To Text
Parameters
8,330,325,888
Spaces
—
Derivative models
—
Velocity 7D/day
2,245
transformerssafetensorsvibevoice_asrautomatic-speech-recognitionASRDiarizationSpeech-to-TextTranscriptionaudio-text-to-textenzhes

Similar models

Models similar to VibeVoice-ASR-HF
ModelDL 30D
midashenglm-7b-0804-fp32mispeech16.2K
midashenglm-7b-1021-bf16mispeech414
audio-flamingo-3-hfnvidia52.2K
music-flamingo-2601-hfnvidia18K
audio-flamingo-next-captioner-hfnvidia14.7K
music-flamingo-hfnvidia4,235
microsoft/VibeVoice-ASR-HF — Downloads, Growth & Stats | OpenModelStats