Qwen3-8B-DFlash-GPTQ-IMatrix-PerfectBlend-NVFP4-W4A4

inference-optimization/Qwen3-8B-DFlash-GPTQ-IMatrix-PerfectBlend-NVFP4-W4A4

Text Generation1B parametersspeculatorsLicense: apache-2.0
Tracked by OpenModelStats since September 28, 2026View on Hugging Face ↗
Downloads 30D
21
Total downloads
21
Gained 7D
—
Gained 30D
—
Likes
0
Rank · tracked models
—
–this week
Parameters
1B
Last source update
Sep 28, 2026

Downloads over time

Rolling recent-window downloads as reported

Likes over time

Rank among tracked models

Lower is better

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-09-28. Charts appear as daily observations accumulate.

Overview

Qwen3-8B-DFlash-GPTQ-IMatrix-PerfectBlend-NVFP4-W4A4 is a text generation model published by inference-optimization. OpenModelStats has tracked the model since Sep 28, 2026, most recently observing it Sep 29, 2026.

Published
Sep 28, 2026
Last updated
Sep 28, 2026
Library
speculators
License
apache-2.0
Task
Text Generation
Parameters
1,048,626,432
Spaces
—
Derivative models
—
Velocity 7D/day
—
speculatorssafetensorsspeculative-decodingdflashqwen3quantizednvfp4gptqimatrixperfectblend-calibrationtext-generationconversational

Publisher

inference-optimizationView publisher statistics →

Derived from Qwen/Qwen3-8B (reported by the source; parent not yet tracked).

Similar models

Models similar to Qwen3-8B-DFlash-GPTQ-IMatrix-PerfectBlend-NVFP4-W4A4
ModelDL 30D
LLaMA3.1-8B-Instruct-DFlash-UltraChatz-lab38.1K
Qwen3-8B-DFlash-b16z-lab14.4K
Qwen3.5-9B-abliterated-DFlashguglxni435
nanoLLaVAqnguyen313.2K
EAGLE3-Llama-3.1-8B-Instructmobilint429
APT3-1B-BaseAzurro4,793
inference-optimization/Qwen3-8B-DFlash-GPTQ-IMatrix-PerfectBlend-NVFP4-W4A4 — Downloads, Growth & Stats | OpenModelStats