LLM EXPLORER 59,071 MODELS INDEXED

Qwen3 Next 80B A3B Instruct NVFP4 by nvidia

By nvidia · 13898 downloads

Qwen3 Next 80B A3B Instruct NVFP4 is an open-source language model by nvidia. Features: 80b LLM, VRAM: 50.7GB, Context: 256K, License: apache-2.0, Instruction-Based, LLM Explorer Score: 0.27.

Base model:quantized:qwen/qwen... Base model:qwen/qwen3-next-80b...   Conversational   Instruct   Model optimizer   Modelopt   Nvfp4   Nvidia   Quantized   Qwen3   Qwen3 next   Region:us   Safetensors   Sharded   Tensorflow

Qwen3 Next 80B A3B Instruct NVFP4 Parameters and Internals

LLM NameQwen3 Next 80B A3B Instruct NVFP4
Repository πŸ€—https://huggingface.co/nvidia/Qwen3-Next-80B-A3B-Instruct-NVFP4 
Base Model(s)  Qwen3 Next 80B A3B Instruct   Qwen/Qwen3-Next-80B-A3B-Instruct
Model Size80b
Required VRAM50.7 GB
Updated2026-05-21
Maintainernvidia
Model Typeqwen3_next
Instruction-BasedYes
Model Files  5.0 GB: 1-of-11   5.0 GB: 2-of-11   5.0 GB: 3-of-11   5.0 GB: 4-of-11   5.0 GB: 5-of-11   5.0 GB: 6-of-11   5.0 GB: 7-of-11   5.0 GB: 8-of-11   5.0 GB: 9-of-11   5.0 GB: 10-of-11   0.7 GB: 11-of-11
Model ArchitectureQwen3NextForCausalLM
Licenseapache-2.0
Context Length262144
Model Max Length262144
Transformers Version4.57.1
Tokenizer ClassQwen2Tokenizer
Padding Token<|im_end|>
Vocabulary Size151936
Errorsreplace

Best Alternatives to Qwen3 Next 80B A3B Instruct NVFP4

Best Alternatives
Context / RAM
Downloads
Likes
Qwen3 Next 80B A3B Instruct256K / 162.7 GB42289
Qwen3 Next 80B A3B Instruct256K / 162.7 GB2866791042
...wen3 Next 80B A3B Instruct FP8256K / 81.8 GB32958389
Qwen3 Next MoE256K / 0 GB1270904
...t 80B A3B Instruct FP8 Dynamic256K / 80.5 GB5064
...ext 80B A3B Instruct Mxfp4 Mlx256K / 42 GB2438
... Instruct Int4 Mixed AutoRound256K / 43.3 GB9724
...t 80B A3B Instruct Qx86 Hi Mlx256K / 73.5 GB1272
...0B A3B Instruct Int4 AutoRound256K / 42.5 GB12010
...en3 Next 80B A3B Instruct 8bit256K / 83.9 GB44438
Note: green Score (e.g. "73.2") means that the model is better than nvidia/Qwen3-Next-80B-A3B-Instruct-NVFP4.