LLM EXPLORER 57,252 MODELS INDEXED

Qwen3.8 4B Distilled GGUF by Ma7ee7

By Ma7ee7 · 34085 downloads

Qwen3.8 4B Distilled GGUF is an open-source language model by Ma7ee7. Features: 4b LLM, VRAM: 2.5GB, Context: 256K, License: apache-2.0, Quantized, LLM Explorer Score: 0.39.

Base model:ma7ee7/qwen3.8 4b d... Base model:quantized:ma7ee7/qw...   Chain-of-thought   Conversational Dataset:r0b0tlab/qwen3.8-max-d...   Distillation   En   Endpoints compatible   Gguf   Knowledge-distillation   Llama.cpp   Quantized   Qwen   Qwen3   Qwen3.8   Reasoning   Region:us   Sequence-level-distillation   Thinking

Qwen3.8 4B Distilled GGUF Parameters and Internals

LLM NameQwen3.8 4B Distilled GGUF
Repository πŸ€—https://huggingface.co/Ma7ee7/Qwen3.8_4B_Distilled_GGUF 
Base Model(s)  Ma7ee7/Qwen3.8_4B_Distilled   Ma7ee7/Qwen3.8_4B_Distilled
Model Size4b
Required VRAM2.5 GB
Updated2026-08-09
MaintainerMa7ee7
Model Typeqwen3
Model Files  2.5 GB   2.9 GB   4.3 GB
Supported Languagesen
GGUF QuantizationYes
Quantization Typegguf
Model ArchitectureQwen3ForCausalLM
Licenseapache-2.0
Context Length262144
Model Max Length262144
Vocabulary Size151936
Torch Data Typebfloat16

Best Alternatives to Qwen3.8 4B Distilled GGUF

Best Alternatives
Context / RAM
Downloads
Likes
...wen3 4B Toolcalling Gguf Codex256K / 4.3 GB261355
...B Toolcall Gguf Llamacpp Codex256K / 4.3 GB13049
...wen3 4B Thinking 2507 Hermes 3256K / 8.1 GB1572
Qwen3 4B Tcomanr Merge V2.2256K / 8 GB282
Qwen3 4B Tcomanr Merge V2256K / 8 GB52
Qwen3 4B 128K GGUF128K / 1.1 GB133027
Qwen3 4B GGUF40K / 1.1 GB131929229
MiniAI Quata1 4B40K / 8.1 GB270
Hmanlab Ai V0.140K / 8.1 GB951
Qwen3 Hermes 4B40K / 8.1 GB5713