ast-finetuned-speech-commands-v2
MIT/ast-finetuned-speech-commands-v2
Tracked by OpenModelStats since August 24, 2026View on Hugging Face ↗
- Downloads 30DRolling recent-window download count as reported by the source platform.
- 7,401
- Total downloadsCumulative all-time download counter reported by the source platform.
- 519.4K
- Gained 7DIncrease in cumulative total downloads between OpenModelStats' own observations seven days apart.
- —
- Gained 30DIncrease in cumulative total downloads between OpenModelStats' own observations roughly thirty days apart.
- —
- LikesCommunity likes reported by the source platform.
- 18
- Rank · tracked modelsPosition among all eligible models tracked by OpenModelStats, ordered by recent downloads. #48 among tracked Audio Classification models.
- #10,635
- –this week
- Parameters
- 85.4M
- Last source update
- Sep 10, 2023
Downloads over time
Cumulative total downloads observed by OpenModelStatsNot enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.
Likes over time
Not enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.
Rank among tracked models
Lower is betterNot enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.
Overview
ast-finetuned-speech-commands-v2 is a audio classification model published by MIT. OpenModelStats has tracked the model since Aug 24, 2026, most recently observing it 1h ago.
- Published
- Nov 14, 2022
- Last updated
- Sep 10, 2023
- Library
- transformers
- License
- bsd-3-clause
- Task
- Audio Classification
- Parameters
- 85,395,491
- Spaces
- —
- Derivative models
- —
- Velocity 7D/day
- —
transformerspytorchsafetensorsaudio-spectrogram-transformeraudio-classificationdataset:speech_commandsarxiv:2104.01778license:bsd-3-clausemodel-indexendpoints_compatibleregion:usdeploy:azure
Publisher
MITView publisher statistics →Similar models
| Model | DL 30D |
|---|---|
| dasheng-basemispeech | 3,591 |
| Bird-MAE-BaseDBD-research-group | 1,180 |
| ced-basemispeech | 12.2K |
| ast-finetuned-audioset-16-16-0.442MIT | 2,252 |
| ast-finetuned-audioset-14-14-0.443MIT | 36.4K |
| ast-finetuned-audioset-10-10-0.4593MIT | 1.4M |