LLM EXPLORER 56,604 MODELS INDEXED

SmolLM2 1.7B Instruct Q8 Mlx by HuggingFaceTB

By HuggingFaceTB · 258 downloads

SmolLM2 1.7B Instruct Q8 Mlx is an open-source language model by HuggingFaceTB. Features: 1.7b LLM, VRAM: 1.8GB, Context: 8K, License: apache-2.0, Quantized, Instruction-Based, ELO: 1062.

  8-bit Base model:huggingfacetb/smoll... Base model:quantized:huggingfa...   Conversational   En   Endpoints compatible   Instruct   Llama   Mlx   Mlx-my-repo   Onnx   Q8   Quantized   Region:us   Safetensors   Transformers.js

SmolLM2 1.7B Instruct Q8 Mlx Benchmarks

nn.n% — How the model compares to the reference models: Anthropic Sonnet 3.5 ("so35"), GPT-4o ("gpt4o") or GPT-4 ("gpt4").

SmolLM2 1.7B Instruct Q8 Mlx Parameters and Internals

LLM NameSmolLM2 1.7B Instruct Q8 Mlx
Repository πŸ€—https://huggingface.co/HuggingFaceTB/SmolLM2-1.7B-Instruct-Q8-mlx 
Base Model(s)  HuggingFaceTB/SmolLM2-1.7B-Instruct   HuggingFaceTB/SmolLM2-1.7B-Instruct
Model Size1.7b
Required VRAM1.8 GB
Updated2026-08-09
MaintainerHuggingFaceTB
Model Typellama
Instruction-BasedYes
Model Files  1.8 GB
Supported Languagesen
Quantization Typeq8
Model ArchitectureLlamaForCausalLM
Licenseapache-2.0
Context Length8192
Model Max Length8192
Transformers Version4.42.3
Tokenizer ClassGPT2Tokenizer
Padding Token<|im_end|>
Vocabulary Size49152
Torch Data Typebfloat16

Best Alternatives to SmolLM2 1.7B Instruct Q8 Mlx

Best Alternatives
Context / RAM
Downloads
Likes
SmolLM2 1.7B Instruct Bnb 4bit8K / 1 GB4072
SmolLM 1.7B Bnb 4bit2K / 1 GB342
SmolLM 1.7B Instruct 4bit2K / 1 GB702
SmolLM 1.7B Instruct 8bit2K / 1.8 GB91
SmolLM2 1.7B Instruct 16K16K / 3.4 GB10310
SmolLM2 1.7B Instruct8K / 3.4 GB154013742
SmolLM2 1.7B Instruct8K / 3.4 GB49096
...ygnis Alpha 1.7B V0.1 Instruct8K / 3.4 GB112
SmolTulu 1.7B Reinforced8K / 3.4 GB125
SmolLM2 1.7 Persona8K / 3.5 GB70
Note: green Score (e.g. "73.2") means that the model is better than HuggingFaceTB/SmolLM2-1.7B-Instruct-Q8-mlx.