ggml-org AI Models

Tracked by OpenModelStats since Aug 23, 2026 · View on Hugging Face ↗

Tracked models
116
Downloads 30D
4.7M
Total downloads
80.4M
Total likes
2,103
Gained 7D

Most downloaded

  1. 1models-movedOct 28, 2025712.8K
  2. 2Qwen3.8-27B-GGUFAug 14, 2026608.7K
  3. 3embeddinggemma-300M-GGUFApr 29, 2026444.4K
  4. 4Qwen3.6-35B-A3B-GGUFJul 16, 2026343.2K
  5. 5embeddinggemma-300m-qat-q8_0-GGUFSep 15, 2025274.2K
  6. 6GLM-4.7-Flash-GGUFJan 23, 2026245.9K
  7. 7gemma-3-1b-it-GGUFMar 12, 2025220.5K
  8. 8NVIDIA-Nemotron-3.5-Lightning-30B-A3B-GGUFAug 23, 2026136.7K
  9. 9tinygemma3-GGUFMay 6, 2025104.6K
  10. 10gpt-oss-120b-GGUFJul 28, 2026103.7K

Recently updated

  1. 1dots3-note-prev-GGUFAug 24, 2026162
  2. 2SmolVLM2-256M-Video-Instruct-GGUFAug 23, 20262,444
  3. 3SmolLM2-135M-GGUFAug 23, 20260
  4. 4NVIDIA-Nemotron-3.5-Lightning-30B-A3B-GGUFAug 23, 2026136.7K
  5. 5Laguna-XS-2.1-GGUFAug 23, 20262,165
  6. 6Laguna-S-2.1-GGUFAug 23, 2026689
  7. 7gemma-4-12B-it-GGUFAug 23, 202641.5K
  8. 8Qwen3.8-27B-GGUFAug 14, 2026608.7K
  9. 9DeepSeek-V4-Flash-0731-GGUFAug 6, 202628.7K
  10. 10DeepSeek-V4-Flash-GGUFAug 6, 20264,251

All models

All tracked models published by ggml-org
ModelDL 30D7D
Qwen2.5-Coder-7B-Q8_0-GGUFggml-org/qwen2.5-coder-7b-q8_0-gguf2,542
Qwen2.5-Omni-7B-GGUFggml-org/qwen2.5-omni-7b-gguf2,450
GLM-4.6V-Flash-GGUFggml-org/glm-4.6v-flash-gguf2,426
pixtral-12b-GGUFggml-org/pixtral-12b-gguf2,400
Qwen3.5-35B-A3B-GGUFggml-org/qwen3.5-35b-a3b-gguf2,397
gemma-3-270m-it-qat-GGUFggml-org/gemma-3-270m-it-qat-gguf2,332
Qwen2.5-VL-72B-Instruct-GGUFggml-org/qwen2.5-vl-72b-instruct-gguf2,228
Laguna-XS-2.1-GGUFggml-org/laguna-xs-2.1-gguf2,165
Qwen2-VL-7B-Instruct-GGUFggml-org/qwen2-vl-7b-instruct-gguf2,156
ultravox-v0_5-llama-3_1-8b-GGUFggml-org/ultravox-v0_5-llama-3_1-8b-gguf2,128
LightOnOCR-2-1B-GGUFggml-org/lightonocr-2-1b-gguf2,106
jina-reranker-v1-turbo-en-GGUFggml-org/jina-reranker-v1-turbo-en-gguf2,048
Qwen3-VL-2B-Instruct-GGUFggml-org/qwen3-vl-2b-instruct-gguf1,829
Qwen2.5-Coder-3B-Q8_0-GGUFggml-org/qwen2.5-coder-3b-q8_0-gguf1,814
Qwen2.5-Coder-1.5B-Instruct-Q8_0-GGUFggml-org/qwen2.5-coder-1.5b-instruct-q8_0-gguf1,741
InternVL2_5-4B-GGUFggml-org/internvl2_5-4b-gguf1,693
Qwen3-4B-Instruct-2507-Q8_0-GGUFggml-org/qwen3-4b-instruct-2507-q8_0-gguf1,601
Qwen3-30B-A3B-GGUFggml-org/qwen3-30b-a3b-gguf1,549
AutoGLM-Phone-9B-GGUFggml-org/autoglm-phone-9b-gguf1,535
Meta-Llama-3.1-8B-Instruct-Q4_0-GGUFggml-org/meta-llama-3.1-8b-instruct-q4_0-gguf1,498
gemma-3-12b-it-qat-GGUFggml-org/gemma-3-12b-it-qat-gguf1,484
InternVL2_5-1B-GGUFggml-org/internvl2_5-1b-gguf1,473
NVIDIA-Nemotron-3-Nano-Omni-30B-A3B-GGUFggml-org/nvidia-nemotron-3-nano-omni-30b-a3b-gguf1,466
Voxtral-Mini-3B-2507-GGUFggml-org/voxtral-mini-3b-2507-gguf1,399
Qwen3-Coder-30B-A3B-Instruct-Q8_0-GGUFggml-org/qwen3-coder-30b-a3b-instruct-q8_0-gguf1,361