LLM EXPLORER 61,060 MODELS INDEXED

Qwen3.8 Flash Next NVFP4 by primitive-ai

By primitive-ai · 7533 downloads

Qwen3.8 Flash Next NVFP4 is an open-source language model by primitive-ai. Features: 119.6b LLM, VRAM: 1.4GB, License: other, LLM Explorer Score: 0.35.

  8-bit Base model:quantized:qwen/qwen... Base model:qwen/qwen3.8-flash-...   Conversational   Endpoints compatible   Flash-next   Image-text-to-text   Modelopt   Nvfp4   Quantized   Qwen3.8   Qwen4 exp   Region:us   Safetensors   Single-gpu   Speculative-decoding   Vllm

Qwen3.8 Flash Next NVFP4 Parameters and Internals

LLM NameQwen3.8 Flash Next NVFP4
Repository πŸ€—https://huggingface.co/primitive-ai/Qwen3.8-Flash-Next-NVFP4 
Base Model(s)  Qwen/Qwen3.8-Flash-Next   Qwen/Qwen3.8-Flash-Next
Model Size119.6b
Required VRAM1.4 GB
Updated2026-08-30
Maintainerprimitive-ai
Model Typeqwen4_exp
Model Files  1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB   1.4 GB
Model ArchitectureQwen4ExpForConditionalGeneration
Licenseother
Transformers Version5.8.0.dev0

Best Alternatives to Qwen3.8 Flash Next NVFP4

Best Alternatives
Context / RAM
Downloads
Likes
Qwen3.8 Flash Next NVFP40K / 78.9 GB13321114
Qwen3.8 Flash Next NVFP40K / 0.4 GB229740
...8 Flash Next ABLITERATED NVFP40K / 0.4 GB143412
...Flash Next CYBERSECURITY NVFP40K / 0.4 GB9966
...3.8 Flash Next Mixed NVFP4 FP80K / 0.1 GB8158
...ash Next NVFP4 Reshard Mtp Fix0K / 1.1 GB1163
Qwen3.8 Flash Next NVFP4 FP80K / 0.4 GB3164
...xt RadixArk NVFP4 Hybrid Sharp0K / 0.4 GB3291
...n3.8 Flash Next DERISKED NVFP40K / 0.4 GB2702
Qwen3.8 Flash Next W4A16 NVFP40K / 0.4 GB1710
Note: green Score (e.g. "73.2") means that the model is better than primitive-ai/Qwen3.8-Flash-Next-NVFP4.