LLM EXPLORER 58,975 MODELS INDEXED

Qwen3 Embedding 0.6B 4bit DWQ by mlx-community

By mlx-community · 13142 downloads

Qwen3 Embedding 0.6B 4bit DWQ is an open-source language model by mlx-community. Features: 0.6b LLM, VRAM: 0.3GB, Context: 32K, License: apache-2.0, Quantized, LLM Explorer Score: 0.23.

  4-bit   4bit Base model:quantized:qwen/qwen... Base model:qwen/qwen3-embeddin...   Conversational   Endpoints compatible   Feature-extraction   Mlx   Quantized   Qwen3   Region:us   Safetensors   Sentence-similarity   Sentence-transformers

Qwen3 Embedding 0.6B 4bit DWQ Parameters and Internals

LLM NameQwen3 Embedding 0.6B 4bit DWQ
Repository πŸ€—https://huggingface.co/mlx-community/Qwen3-Embedding-0.6B-4bit-DWQ 
Base Model(s)  Qwen/Qwen3-Embedding-0.6B   Qwen/Qwen3-Embedding-0.6B
Model Size0.6b
Required VRAM0.3 GB
Updated2026-08-08
Maintainermlx-community
Model Typeqwen3
Model Files  0.3 GB
Quantization Type4bit
Model ArchitectureQwen3ForCausalLM
Licenseapache-2.0
Context Length32768
Model Max Length32768
Transformers Version4.51.3
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Vocabulary Size151669
Torch Data Typebfloat16
Errorsreplace

Best Alternatives to Qwen3 Embedding 0.6B 4bit DWQ

Best Alternatives
Context / RAM
Downloads
Likes
Agent 0.6B125K / 1.2 GB140
Fizik 0.6B Preview40K / 1.2 GB90
Qwen3 0.6B Unsloth Bnb 4bit40K / 0.6 GB14692025
Qwen3 0.6B 4bit40K / 0.3 GB5031514
Qwentestnew140K / 1.2 GB660
Qwen3 0.6B 8bit40K / 0.6 GB413187
Qwen3 0.6B SFTchat Math40K / 1.2 GB150
Meet7 0.6B40K / 1.2 GB101
...d Qwen3 0.6B Gabliterated 4bit40K / 0.3 GB241
Qwen3 0.6B Gabliterated 4bit40K / 0.3 GB61
Note: green Score (e.g. "73.2") means that the model is better than mlx-community/Qwen3-Embedding-0.6B-4bit-DWQ.