LLM EXPLORER 60,702 MODELS INDEXED

Qwen3.6 35B A3B W4a16 Llmcompressor by amd

By amd · 375 downloads

Qwen3.6 35B A3B W4a16 Llmcompressor is an open-source language model by amd. Features: 35b LLM, VRAM: 20.3GB, License: apache-2.0, LLM Explorer Score: 0.29.

  4-bit   Amd Base model:quantized:qwen/qwen... Base model:qwen/qwen3.6-35b-a3...   Compressed-tensors   Conversational   Cpu-inference   En   Endpoints compatible   Image-text-to-text   Int4   Llm-compressor   Quantized   Qwen3 5 moe   Region:us   Safetensors   W4a16   Weight-only   Zendnn

Qwen3.6 35B A3B W4a16 Llmcompressor Parameters and Internals

LLM NameQwen3.6 35B A3B W4a16 Llmcompressor
Repository πŸ€—https://huggingface.co/amd/Qwen3.6-35B-A3B-w4a16-llmcompressor 
Base Model(s)  Qwen/Qwen3.6-35B-A3B   Qwen/Qwen3.6-35B-A3B
Model Size35b
Required VRAM20.3 GB
Updated2026-09-08
Maintaineramd
Model Typeqwen3_5_moe
Model Files  20.3 GB
Supported Languagesen
Model ArchitectureQwen3_5MoeForConditionalGeneration
Licenseapache-2.0
Model Max Length262144
Transformers Version5.10.1
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Errorsreplace

Best Alternatives to Qwen3.6 35B A3B W4a16 Llmcompressor

Best Alternatives
Context / RAM
Downloads
Likes
Ornith 1.5 35B A3B NPU2256K /  GB1891
Qwen3.6 35B A3B0K / 72.2 GB55408552648
Qwen3.5 35B A3B0K / 72.3 GB21585111482
Ornith 1.0 35B0K / 70.4 GB2983134506
Qwen3.6 35B A3B NVFP40K / 23.4 GB8788684482
Qwen3.6 35B A3B FP80K / 0.8 GB8984811343
KAT Coder V2.5 Dev0K / 69.7 GB17885537
Qwen AgentWorld 35B A3B0K / 69.2 GB66155665
Qwen3.6 35B A3B NVFP4 Fast0K / 23.6 GB46366198
Ornith 1.5 35B A3B0K / 72.1 GB1713201
Note: green Score (e.g. "73.2") means that the model is better than amd/Qwen3.6-35B-A3B-w4a16-llmcompressor.