LLM EXPLORER 55,709 MODELS INDEXED

Qwen3.5 9B 4bit by mlx-community

By mlx-community · 17566 downloads

Qwen3.5 9B 4bit is an open-source language model by mlx-community. Features: 9b LLM, VRAM: 5.9GB, License: apache-2.0, Quantized.

  4-bit   4bit Base model:quantized:qwen/qwen... Base model:qwen/qwen3.5-9b-bas...   Conversational   Endpoints compatible   Image-text-to-text   Mlx   Quantized   Qwen3 5   Region:us   Safetensors   Sharded   Tensorflow

Qwen3.5 9B 4bit Parameters and Internals

LLM NameQwen3.5 9B 4bit
Repository πŸ€—https://huggingface.co/mlx-community/Qwen3.5-9B-4bit 
Base Model(s)  Qwen/Qwen3.5-9B-Base   Qwen/Qwen3.5-9B-Base
Model Size9b
Required VRAM5.9 GB
Updated2026-08-08
Maintainermlx-community
Model Typeqwen3_5
Model Files  5.3 GB: 1-of-2   0.6 GB: 2-of-2
Quantization Type4bit
Model ArchitectureQwen3_5ForConditionalGeneration
Licenseapache-2.0
Model Max Length262144
Transformers Version4.57.0.dev0
Tokenizer ClassTokenizersBackend
Padding Token<|endoftext|>
Errorsreplace

Best Alternatives to Qwen3.5 9B 4bit

Best Alternatives
Context / RAM
Downloads
Likes
Ornith 1.0 9B 4bit0K / 5.9 GB97919
Qwen3.5 9B MLX 4bit0K / 5.9 GB22827151
Qwen3.5 9B OptiQ 4bit0K / 7.1 GB1013686
Ornith 1.0 9B 8bit0K / 10.4 GB20792
Ornith 1.0 9B 6bit0K / 8.1 GB8282
... Opus Reasoning Distilled 4bit0K / 5 GB137025
Qwopus3.5 9B V3 4bit0K / 5.9 GB652
Ornith 1.0 9B OptiQ 4bit0K / 7.1 GB16568
Qwen3.5 9B 6bit0K / 8.1 GB7905
Fara1.5 9B 8bit0K / 10.4 GB1201
Note: green Score (e.g. "73.2") means that the model is better than mlx-community/Qwen3.5-9B-4bit.