LLM EXPLORER 58,297 MODELS INDEXED

Mellum2 12B A2.5B Instruct 4bit by mlx-community

By mlx-community · 14 downloads

Mellum2 12B A2.5B Instruct 4bit is an open-source language model by mlx-community. Features: 12b LLM, VRAM: 6.8GB, Context: 128K, License: apache-2.0, Quantized, Instruction-Based, LLM Explorer Score: 0.28.

  4-bit   4bit Base model:jetbrains/mellum2-1... Base model:quantized:jetbrains...   Conversational   En   Instruct   Mellum   Mlx   Moe   Quantized   Region:us   Safetensors   Sharded   Tensorflow

Mellum2 12B A2.5B Instruct 4bit Parameters and Internals

LLM NameMellum2 12B A2.5B Instruct 4bit
Repository πŸ€—https://huggingface.co/mlx-community/Mellum2-12B-A2.5B-Instruct-4bit 
Base Model(s)  JetBrains/Mellum2-12B-A2.5B-Instruct   JetBrains/Mellum2-12B-A2.5B-Instruct
Model Size12b
Required VRAM6.8 GB
Updated2026-08-13
Maintainermlx-community
Model Typemellum
Instruction-BasedYes
Model Files  5.3 GB: 1-of-2   1.5 GB: 2-of-2
Supported Languagesen
Quantization Type4bit
Model ArchitectureMellumForCausalLM
Licenseapache-2.0
Context Length131072
Model Max Length131072
Transformers Version5.8.1
Tokenizer ClassTokenizersBackend
Padding Token<|endoftext|>
Vocabulary Size98304

Best Alternatives to Mellum2 12B A2.5B Instruct 4bit

Best Alternatives
Context / RAM
Downloads
Likes
Mellum2 12B A2.5B Instruct128K / 24.3 GB449882
...llum2 12B A2.5B Instruct Mxfp4128K / 6.4 GB1661
Note: green Score (e.g. "73.2") means that the model is better than mlx-community/Mellum2-12B-A2.5B-Instruct-4bit.