espnet AI Models
Tracked by OpenModelStats since Aug 23, 2026 · View on Hugging Face ↗
- Tracked modelsRepositories from this publisher currently tracked by OpenModelStats.
- 590
- Downloads 30DSum of each tracked model's rolling recent-downloads counter. A sum of repository counters is not a count of unique users.
- 44.8K
- Total downloadsSum of cumulative all-time download counters across tracked repositories.
- 1.5M
- Total likes
- 654
- Gained 7DCombined seven-day cumulative-download gains measured by OpenModelStats.
- 10.9K
Most downloaded
- 1owsm_ctc_v4_1BSep 20, 202611.2K
- 2fastspeech2_conformerOct 6, 20236,276
- 3fastspeech2_conformer_with_hifiganOct 5, 20235,878
- 4hubert_dummySep 18, 20264,869
- 5voxcelebs12_rawnet3Sep 18, 20264,302
- 6voxcelebs12_ecapa_wavlm_jointSep 18, 20264,182
- 7kamo-naoyuki-mini_an4_asr_train_raw_bpe_valid.acc.bestSep 19, 20261,752
- 8owsm_v4_small_370MSep 19, 2026226
- 9fastspeech2_conformer_hifiganSep 18, 2026214
- 10powsmJan 21, 2026203
Fastest growing
- 1owsm_ctc_v4_1B+3,217 in 7D11.2K
- 2hubert_dummy+2,082 in 7D4,869
- 3fastspeech2_conformer+1,739 in 7D6,276
- 4fastspeech2_conformer_with_hifigan+1,583 in 7D5,878
- 5voxcelebs12_rawnet3+1,281 in 7D4,302
- 6kamo-naoyuki-mini_an4_asr_train_raw_bpe_valid.acc.best+707 in 7D1,752
- 7kan-bayashi_ljspeech_vits+52 in 7D95
- 8xeus+17 in 7D89
Recently updated
- 1must_c_esp2_st_train_st_conformerSep 24, 20265
- 2must_c_st_train_st_conformerSep 24, 20269
- 3meld_esp2_cls_wavlm_base_plusSep 21, 202614
- 4multi-talker-whisper-small-amiSep 20, 202653
- 5anogkongda_librimix_enh_train_raw_valid.si_snr.aveSep 20, 202619
- 6dns_icassp21_enh_train_enh_tcn_tf_rawSep 20, 202612
- 7chenda-li-wsj0_2mix_enh_train_enh_conv_tasnet_raw_valid.si_snr.aveSep 20, 202614
- 8anogkongda-librimix_enh_train_raw_valid.si_snr.aveSep 20, 202615
- 9Chenda_Li_wsj0_2mix_enh_train_enh_conv_tasnet_raw_valid.si_snr.aveSep 20, 202617
- 10tedlium2_streaming_transformerSep 20, 202621
All models
| Model | DL 30D | 7D |
|---|---|---|
| dac_44k_all_single_surveyespnet/dac_44k_all_single_survey | 1 | — |
| owsm_pure_codec_v1.1_16kespnet/owsm_pure_codec_v1.1_16k | 1 | — |
| audioset_dac_16kespnet/audioset_dac_16k | 1 | — |
| diar_ami_eend_edaespnet/diar_ami_eend_eda | 1 | — |
| YushiUeda_librimix_diar_enh_2_3_spkespnet/yushiueda_librimix_diar_enh_2_3_spk | 1 | — |
| owsm_dac_v2_16kespnet/owsm_dac_v2_16k | 1 | — |
| YushiUeda_librimix_diar_enh_2_3_spk_lmfespnet/yushiueda_librimix_diar_enh_2_3_spk_lmf | 1 | — |
| libritts_dac_16kespnet/libritts_dac_16k | 1 | — |
| opencpop_svs_train_toksing_300epoch-multi_hl6_wl6_wl23espnet/opencpop_svs_train_toksing_300epoch-multi_hl6_wl6_wl23 | 1 | — |
| dac_44k_speech_surveyespnet/dac_44k_speech_survey | 1 | — |
| mls-english_encodec_16k_360epochespnet/mls-english_encodec_16k_360epoch | 1 | — |
| libritts_encodec_24kespnet/libritts_encodec_24k | 1 | — |
| BEATs-BEAN.CornellBirdIdentificationespnet/beats-bean.cornellbirdidentification | 1 | — |
| s3prl_adapter_modelespnet/s3prl_adapter_model | 1 | — |
| speechlm_tts_ls_giga_mlsen_amuse_speech_multiscaleespnet/speechlm_tts_ls_giga_mlsen_amuse_speech_multiscale | 1 | — |
| OpenBEATS-Large-i3-humbugdbespnet/openbeats-large-i3-humbugdb | 1 | — |
| mls-audioset_encodec_16kespnet/mls-audioset_encodec_16k | 1 | — |
| dongwei_ami_vad_rnnespnet/dongwei_ami_vad_rnn | 0 | — |
| proyecto_nahuatlespnet/proyecto_nahuatl | 0 | — |
| dongwei_tedlium3_asr_conformer_external_lmespnet/dongwei_tedlium3_asr_conformer_external_lm | 0 | — |
| speech_translation_mboshi_french_transformer_baselineespnet/speech_translation_mboshi_french_transformer_baseline | 0 | — |
| mediaspeech-fr-hubertespnet/mediaspeech-fr-hubert | 0 | — |
| asr_train_ogi_kids_speech_branchformer_transformerespnet/asr_train_ogi_kids_speech_branchformer_transformer | 0 | — |
| YosukeHiguchi_espnet2_wsj_asr_transformer_maskctcespnet/yosukehiguchi_espnet2_wsj_asr_transformer_maskctc | 0 | — |
| owls_1b_11kespnet/owls_1b_11k | 0 | — |