RedHatAI AI Models

Tracked by OpenModelStats since Aug 23, 2026 · View on Hugging Face ↗

Tracked models
282
Downloads 30D
8.9M
Total downloads
120.9M
Total likes
2,311
Gained 7D
2.4M

Most downloaded

  1. 1Qwen3.6-35B-A3B-NVFP4Aug 13, 20261M
  2. 2NVIDIA-Nemotron-3.5-Lightning-30B-A3B-FP8Sep 17, 2026950.3K
  3. 3gemma-4-26B-A4B-it-FP8-dynamicAug 13, 2026890.2K
  4. 4Qwen3.8-27B-INT4Sep 28, 2026569.9K
  5. 5Llama-3.2-3B-Instruct-FP8-dynamicOct 9, 2024518.2K
  6. 6gemma-4-31B-it-FP8-blockAug 13, 2026433K
  7. 7Qwen3-32B-NVFP4Nov 21, 2025411.8K
  8. 8gemma-4-26B-A4B-it-NVFP4Aug 13, 2026315.2K
  9. 9gemma-3-27b-it-quantized.w4a16Jun 9, 2025293.7K
  10. 10Qwen3-Coder-Next-FP8-dynamicFeb 27, 2026288.1K

Fastest growing

  1. 1Qwen3.6-35B-A3B-NVFP4+314.9K in 7D1M
  2. 2NVIDIA-Nemotron-3.5-Lightning-30B-A3B-FP8+238.5K in 7D950.3K
  3. 3gemma-4-26B-A4B-it-FP8-dynamic+218K in 7D890.2K
  4. 4Qwen3-Coder-Next-FP8-dynamic+162.9K in 7D288.1K
  5. 5Llama-3.2-3B-Instruct-FP8-dynamic+127.5K in 7D518.2K
  6. 6gemma-4-31B-it-FP8-block+118.7K in 7D433K
  7. 7Qwen3.8-27B-INT4+117.4K in 7D569.9K
  8. 8Qwen3-32B-NVFP4+106.3K in 7D411.8K
  9. 9Llama-3.2-3B-Instruct-FP8+94.2K in 7D184.9K
  10. 10gemma-3-27b-it-quantized.w4a16+88.9K in 7D293.7K

Recently updated

  1. 1Qwen3.8-27B-MXFP4Oct 8, 2026853
  2. 2gemma-4-31B-it-speculator.dflashOct 7, 2026646
  3. 3Qwen3.5-35B-A3B-quantized.w8a8Oct 6, 2026406
  4. 4Laguna-S-2.1-NVFP4Oct 6, 2026316
  5. 5Laguna-S-2.1-FP8Oct 6, 2026193
  6. 6Qwen3.8-Flash-Next-NVFP4Oct 5, 20261,503
  7. 7wolf-defender-prompt-injectionOct 5, 202687
  8. 8DeepSeek-V4-Flash-0731-NVFP4Oct 2, 2026625
  9. 9granite-3.2-2b-instructOct 1, 2026309
  10. 10Qwen3.8-27B-NVFP4Sep 30, 202634.5K

All models

All tracked models published by RedHatAI
ModelDL 30D7D
Llama-Guard-4-12B-quantized.w4a16redhatai/llama-guard-4-12b-quantized.w4a162,240+338
Llama-4-Scout-17B-16E-Instructredhatai/llama-4-scout-17b-16e-instruct2,235+442
Qwen2.5-VL-7B-Instruct-quantized.w4a16redhatai/qwen2.5-vl-7b-instruct-quantized.w4a162,205+270
Mistral-Small-3.1-24B-Instruct-2503redhatai/mistral-small-3.1-24b-instruct-25032,141+381
gemma-3-4b-it-quantized.w8a8redhatai/gemma-3-4b-it-quantized.w8a82,098+1,112
DeepSeek-R1-Distill-Llama-70B-FP8-dynamicredhatai/deepseek-r1-distill-llama-70b-fp8-dynamic2,094+420
NVIDIA-Nemotron-3-Ultra-550B-A55B-FP8-blockredhatai/nvidia-nemotron-3-ultra-550b-a55b-fp8-block2,088+742
gemma-3-12b-it-FP8-dynamicredhatai/gemma-3-12b-it-fp8-dynamic1,958+559
llama2.c-stories110M-pruned50redhatai/llama2.c-stories110m-pruned501,954+434
Llama-3.1-8B-Instructredhatai/llama-3.1-8b-instruct1,917+452
Inkling-FP8-dynamicredhatai/inkling-fp8-dynamic1,911+1,782
gemma-4-26B-A4B-itredhatai/gemma-4-26b-a4b-it1,903+328
Meta-Llama-3.1-405B-Instruct-FP8-dynamicredhatai/meta-llama-3.1-405b-instruct-fp8-dynamic1,902+514
Mistral-7B-Instruct-v0.3-FP8redhatai/mistral-7b-instruct-v0.3-fp81,883+0
Qwen3.8-27Bredhatai/qwen3.8-27b1,860+329
Qwen3.6-35B-A3B-FP8-dynamicredhatai/qwen3.6-35b-a3b-fp8-dynamic1,781+430
whisper-large-v3-turbo-quantized.w8a8redhatai/whisper-large-v3-turbo-quantized.w8a81,763+79
Muse-Glimmer-30Bredhatai/muse-glimmer-30b1,748+293
Qwen2-0.5B-Instruct-FP8redhatai/qwen2-0.5b-instruct-fp81,740+585
Qwen2.5-VL-72B-Instruct-FP8-dynamicredhatai/qwen2.5-vl-72b-instruct-fp8-dynamic1,662+138
Llama-3.3-70B-Instructredhatai/llama-3.3-70b-instruct1,656—
Qwen3-30B-A3B-quantized.w4a16redhatai/qwen3-30b-a3b-quantized.w4a161,650+505
Qwen3-4B-Instruct-2507-quantized.w8a8redhatai/qwen3-4b-instruct-2507-quantized.w8a81,639+211
Inkling-Small-FP8-dynamicredhatai/inkling-small-fp8-dynamic1,568+186
whisper-large-v3-quantized.w8a8redhatai/whisper-large-v3-quantized.w8a81,566+289