LLM EXPLORER 55,709 MODELS INDEXED

Qwen3 14B NVFP4 by nvidia

By nvidia · 125214 downloads

Qwen3 14B NVFP4 is an open-source language model by nvidia. Features: 14b LLM, VRAM: 10.6GB, Context: 40K, License: apache-2.0, LLM Explorer Score: 0.28.

  8-bit Base model:quantized:qwen/qwen...   Base model:qwen/qwen3-14b   Conversational   Fp4   Model optimizer   Modelopt   Nvidia   Quantized   Qwen3   Region:us   Safetensors   Sharded   Tensorflow
Model Card on HF πŸ€—: https://huggingface.co/nvidia/Qwen3-14B-NVFP4 

Qwen3 14B NVFP4 Parameters and Internals

LLM NameQwen3 14B NVFP4
Repository πŸ€—https://huggingface.co/nvidia/Qwen3-14B-NVFP4 
Base Model(s)  Qwen3 14B   Qwen/Qwen3-14B
Model Size14b
Required VRAM10.6 GB
Updated2026-08-03
Maintainernvidia
Model Typeqwen3
Model Files  5.0 GB: 1-of-3   4.0 GB: 2-of-3   1.6 GB: 3-of-3
Model ArchitectureQwen3ForCausalLM
Licenseapache-2.0
Context Length40960
Model Max Length40960
Transformers Version4.53.1
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Vocabulary Size151936
Torch Data Typebfloat16
Errorsreplace

Best Alternatives to Qwen3 14B NVFP4

Best Alternatives
Context / RAM
Downloads
Likes
...JA Qwen3 14B Agentic 256K V0.1256K / 29.5 GB288
SimpleChat 14B V1195K / 29.5 GB102
...0528DistillQwen 14B V27.3 200K195K / 29.5 GB95
...uct 21B Brainstorm20x 128K Ctx128K / 84.1 GB80
NousCoder 14B80K / 29.5 GB379221
MiroThinker 14B DPO V0.264K / 29.7 GB286
UIGEN T3 14B Preview40K / 29.5 GB2422
Qwen3 14B40K / 29.7 GB2722973437
Hermes 4 14B40K / 29.5 GB194693169
Qwen3 14B FP840K / 16.4 GB45851548
Note: green Score (e.g. "73.2") means that the model is better than nvidia/Qwen3-14B-NVFP4.