LLM EXPLORER 59,420 MODELS INDEXED

CodeLlama 70B Instruct Hf 4bit MLX by mlx-community

By mlx-community · 417 downloads

CodeLlama 70B Instruct Hf 4bit MLX is an open-source language model by mlx-community. Features: 70b LLM, VRAM: 39.1GB, Context: 4K, License: llama2, Quantized, Instruction-Based, Code Generating, LLM Explorer Score: 0.11.

  4bit   Code   Codegen   Conversational   Instruct   Llama   Llama2   Mlx   Quantized   Region:us   Sharded   Tensorflow

CodeLlama 70B Instruct Hf 4bit MLX Parameters and Internals

Model Type 
text generation, code generation
Additional Notes 
Model converted to MLX format for use with 'mlx-lm' package.
Input Output 
Accepted Modalities:
code
LLM NameCodeLlama 70B Instruct Hf 4bit MLX
Repository πŸ€—https://huggingface.co/mlx-community/CodeLlama-70b-Instruct-hf-4bit-MLX 
Model Size70b
Required VRAM39.1 GB
Updated2026-07-21
Maintainermlx-community
Model Typellama
Instruction-BasedYes
Model Files  5.3 GB: 1-of-8   5.3 GB: 2-of-8   5.3 GB: 3-of-8   5.3 GB: 4-of-8   5.3 GB: 5-of-8   5.3 GB: 6-of-8   5.3 GB: 7-of-8   2.0 GB: 8-of-8
Supported Languagescode
Quantization Type4bit
Generates CodeYes
Model ArchitectureLlamaForCausalLM
Licensellama2
Context Length4096
Model Max Length4096
Transformers Version4.36.2
Vocabulary Size32016
Torch Data Typebfloat16

Best Alternatives to CodeLlama 70B Instruct Hf 4bit MLX

Best Alternatives
Context / RAM
Downloads
Likes
...70B Instruct Nf4 Fp16 Upscaled4K / 138.7 GB101
...70B Instruct Hf 5.0bpw H6 EXL22K / 43.6 GB56
...70B Instruct Hf 2.4bpw H6 EXL22K / 21.3 GB61
...0B Instruct Hf 2.65bpw H6 EXL22K / 23.4 GB53
...70B Instruct Hf 4.0bpw H6 EXL22K / 35.1 GB61
CodeLlama 70B Instruct Hf4K / 72.3 GB72924
Code Llama 70B Python Instruct4K / 138.1 GB51
CodeLlama 70B Instruct Hf4K / 72.3 GB463210
CodeLlama 70B Instruct Hf GGUF4K / 25.5 GB4872
CodeLlama 70B Instruct AWQ4K / 36.6 GB29713
Note: green Score (e.g. "73.2") means that the model is better than mlx-community/CodeLlama-70b-Instruct-hf-4bit-MLX.