Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
nvidia
/
Qwen3.8-2.4T-A95B-NVFP4
like
10
Follow
NVIDIA
67.2k
Text Generation
Safetensors
Model Optimizer
qwen3_5_moe_text
NVIDIA
ModelOpt
Qwen3.8
quantized
4-bit precision
FP4
fp4
conversational
8-bit precision
modelopt
License:
nvidia-open-model-license
Model card
Files
Files and versions
xet
Community
1
Copy to bucket
new
refs/pr/1
Qwen3.8-2.4T-A95B-NVFP4
/
.quant_summary.txt
shengliangx
Qwen3.8-2.4T-A95B-NVFP4 (NVFP4 experts + FP8 attn/KV), complete with MTP
89ef427
6 days ago
Raw
Download with hf CLI
Copy download link
History
Safe
771 kB
File too large to display, you can
check the raw version
instead.