LLM EXPLORER 61,060 MODELS INDEXED

Qwen3.8 27B GPTQ INT4 FP8KV by abhishekchohan

By abhishekchohan · 6488 downloads

Qwen3.8 27B GPTQ INT4 FP8KV is an open-source language model by abhishekchohan. Features: 27b LLM, VRAM: 20.2GB, License: apache-2.0, Quantized, Instruction-Based, LLM Explorer Score: 0.33.

Base model:quantized:qwen/qwen...   Base model:qwen/qwen3.8-27b   Compressed-tensors   Conversational Dataset:nvidia/nemotron-sft-in... Dataset:nvidia/nemotron-sft-ma... Dataset:nvidia/nemotron-sft-mu... Dataset:nvidia/nemotron-sft-sc... Dataset:nvidia/nemotron-sft-sw...   Endpoints compatible   Fp8-kv-cache   Gptq   Image-text-to-text   Imatrix   Instruct   Int4   Llm-compressor   Long-context   Quantized   Qwen3 5   Region:us   Safetensors   Sharded   Speculative-decoding   Tensorflow   Vllm   W4a16

Qwen3.8 27B GPTQ INT4 FP8KV Parameters and Internals

LLM NameQwen3.8 27B GPTQ INT4 FP8KV
Repository πŸ€—https://huggingface.co/abhishekchohan/Qwen3.8-27B-GPTQ-INT4-FP8KV 
Base Model(s)  Qwen/Qwen3.8-27B   Qwen/Qwen3.8-27B
Model Size27b
Required VRAM20.2 GB
Updated2026-08-27
Maintainerabhishekchohan
Model Typeqwen3_5
Instruction-BasedYes
Model Files  2.5 GB: 1-of-6   4.0 GB: 2-of-6   4.0 GB: 3-of-6   4.0 GB: 4-of-6   4.0 GB: 5-of-6   1.7 GB: 6-of-6   0.8 GB
GPTQ QuantizationYes
Quantization Typegptq
Model ArchitectureQwen3_5ForConditionalGeneration
Licenseapache-2.0
Model Max Length262144
Transformers Version5.10.1
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Errorsreplace

Best Alternatives to Qwen3.8 27B GPTQ INT4 FP8KV

Best Alternatives
Context / RAM
Downloads
Likes
Qwopus3.8 27B Flash INT4 W4A160K / 18.4 GB732
Qwopus3.8 27B Flash0K / 55.6 GB23419
Qwopus3.8 27B Flash AWQ MTP0K / 19.1 GB5911
Tmax 27B0K / 53.8 GB463026
Qwen3.8 27B FULL NVFP40K / 18.8 GB3362
Salience 27B R60K / 55.6 GB457
...8 27B Brainwaves WFH Mxfp4 Mlx0K / 15.1 GB4901
...sion GAIN V1.1 Fable Mxfp8 Mlx0K / 28.7 GB7202
...1 Darker Hero GAIN B Mxfp8 Mlx0K / 28.7 GB6092
... 27B Brainwaves 2M Qx64 Hi Mlx0K / 21 GB5521