Llama-3.1-8B-Instruct-NVFP4

nvidia/Llama-3.1-8B-Instruct-NVFP4

4.5B parametersModel OptimizerLicense: other
Tracked by OpenModelStats since August 24, 2026View on Hugging Face ↗
Downloads 30D
111.3K
Total downloads
1.3M
Gained 7D
Gained 30D
Likes
15
Rank · tracked models
#2,281
this week
Parameters
4.5B
Last source update
Sep 15, 2025

Downloads over time

Cumulative total downloads observed by OpenModelStats

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Likes over time

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Rank among tracked models

Lower is better

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Overview

Llama-3.1-8B-Instruct-NVFP4 is an AI model published by nvidia. OpenModelStats has tracked the model since Aug 24, 2026, most recently observing it 4 min ago.

Published
Sep 5, 2025
Last updated
Sep 15, 2025
Library
Model Optimizer
License
other
Task
Parameters
4,540,600,320
Spaces
Derivative models
Velocity 7D/day
Model OptimizersafetensorsllamanvidiaModelOptllama3quantizedFP4fp4base_model:meta-llama/Llama-3.1-8B-Instructbase_model:quantized:meta-llama/Llama-3.1-8B-Instructlicense:other

Model family

Based on base-model relationships reported by the source's metadata.

Similar models

Models similar to Llama-3.1-8B-Instruct-NVFP4
ModelDL 30D
Llama-3.1-Minitron-4B-Depth-Basenvidia5,225
webAI-ColVec1.1-4bwebAI-Official1,396
webAI-ColVec1-4bwebAI-Official658
colqwen3.5-4.5B-v3athrael-soju136.5K
VultronRetrieverCore-Qwen3.5-4.5Bvultr1,349
Qwen3.5-4B-quantized.w8a8RedHatAI15.2K