LLM EXPLORER 59,694 MODELS INDEXED

Qwen3.8 27B GPTQ INT4 FP8KV by abhishekchohan

By abhishekchohan · 6488 downloads

Qwen3.8 27B GPTQ INT4 FP8KV is an open-source language model by abhishekchohan. Features: 27b LLM, VRAM: 20.2GB, License: apache-2.0, Quantized, Instruction-Based, LLM Explorer Score: 0.37.

Base model:quantized:qwen/qwen...   Base model:qwen/qwen3.8-27b   Compressed-tensors   Conversational Dataset:nvidia/nemotron-sft-in... Dataset:nvidia/nemotron-sft-ma... Dataset:nvidia/nemotron-sft-mu... Dataset:nvidia/nemotron-sft-sc... Dataset:nvidia/nemotron-sft-sw...   Endpoints compatible   Fp8-kv-cache   Gptq   Image-text-to-text   Imatrix   Instruct   Int4   Llm-compressor   Long-context   Quantized   Qwen3 5   Region:us   Safetensors   Sharded   Speculative-decoding   Tensorflow   Vllm   W4a16

Qwen3.8 27B GPTQ INT4 FP8KV Parameters and Internals

LLM NameQwen3.8 27B GPTQ INT4 FP8KV
Repository πŸ€—https://huggingface.co/abhishekchohan/Qwen3.8-27B-GPTQ-INT4-FP8KV 
Base Model(s)  Qwen/Qwen3.8-27B   Qwen/Qwen3.8-27B
Model Size27b
Required VRAM20.2 GB
Updated2026-08-27
Maintainerabhishekchohan
Model Typeqwen3_5
Instruction-BasedYes
Model Files  2.5 GB: 1-of-6   4.0 GB: 2-of-6   4.0 GB: 3-of-6   4.0 GB: 4-of-6   4.0 GB: 5-of-6   1.7 GB: 6-of-6   0.8 GB
GPTQ QuantizationYes
Quantization Typegptq
Model ArchitectureQwen3_5ForConditionalGeneration
Licenseapache-2.0
Model Max Length262144
Transformers Version5.10.1
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Errorsreplace

Best Alternatives to Qwen3.8 27B GPTQ INT4 FP8KV

Best Alternatives
Context / RAM
Downloads
Likes
... 27B Brainwaves 2M Qx64 Hi Mlx0K / 21 GB5521
Tmax 27B0K / 53.8 GB463026
... 27B Brainwaves 1M Qx86 Hi Mlx0K / 27 GB4271
...1 Darker Hero GAIN B Mxfp8 Mlx0K / 28.7 GB6092
...sion GAIN V1.1 Fable Mxfp8 Mlx0K / 28.7 GB7202
Qwen3.8 27B Brainwaves0K / 55.5 GB4011
...en3.8 27B Brainwaves Mxfp8 Mlx0K / 28.7 GB3611
...ble Fusion F711 GAIN Mxfp4 Mlx0K / 15.1 GB4231
Qwen3.8 27B AWQ INT40K / 20.2 GB1100
Qwen3.8 27b Euskara0K / 55.6 GB1820