Frame_VAD_Multilingual_MarbleNet_v2.0
nvidia/Frame_VAD_Multilingual_MarbleNet_v2.0
Tracked by OpenModelStats since August 24, 2026View on Hugging Face ↗
- Downloads 30DRolling recent-window download count as reported by the source platform.
- 3,142
- Total downloadsCumulative all-time download counter reported by the source platform.
- 116.8K
- Gained 7DIncrease in cumulative total downloads between OpenModelStats' own observations seven days apart.
- —
- Gained 30DIncrease in cumulative total downloads between OpenModelStats' own observations roughly thirty days apart.
- —
- LikesCommunity likes reported by the source platform.
- 47
- Rank · tracked modelsPosition among all eligible models tracked by OpenModelStats, ordered by recent downloads. #14 among tracked Voice Activity Detection models.
- #13,660
- –this week
- Parameters
- —
- Last source update
- May 14, 2025
Downloads over time
Cumulative total downloads observed by OpenModelStatsNot enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.
Likes over time
Not enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.
Rank among tracked models
Lower is betterNot enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.
Overview
Frame_VAD_Multilingual_MarbleNet_v2.0 is a voice activity detection model published by nvidia. OpenModelStats has tracked the model since Aug 24, 2026, most recently observing it 3h ago.
- Published
- May 8, 2025
- Last updated
- May 14, 2025
- Library
- nemo
- License
- other
- Task
- Voice Activity Detection
- Parameters
- —
- Spaces
- —
- Derivative models
- —
- Velocity 7D/day
- —
nemoMultilingualMarbleNetpytorchspeechaudioVADonnxonnxruntimevoice-activity-detectionenes
Publisher
nvidiaView publisher statistics →Similar models
| Model | DL 30D |
|---|---|
| segmentation-3.0pyannote | 6M |
| segmentationpyannote | 4.7M |
| Namo-Turn-Detector-v1-Koreanvideosdk-live | 411.1K |
| Silero-VAD-v5-MLXaufklarer | 38.8K |
| speaker-diarization-precision-2pyannote | 36.8K |
| silero-vad-coremlFluidInference | 30.6K |