LLM EXPLORER 59,420 MODELS INDEXED

Codellama 70B Instruct Nf4 Fp16 Upscaled by arnavgrg

By arnavgrg · 10 downloads

Codellama 70B Instruct Nf4 Fp16 Upscaled is an open-source language model by arnavgrg. Features: 70b LLM, VRAM: 138.7GB, Context: 4K, License: apache-2.0, Quantized, Instruction-Based, Code Generating, LLM Explorer Score: 0.11.

  Codegen   Conversational   Endpoints compatible   Fp16   Instruct   Llama   Quantized   Region:us   Safetensors   Sharded   Tensorflow

Codellama 70B Instruct Nf4 Fp16 Upscaled Parameters and Internals

Model Type 
text generation
Additional Notes 
Quantization to nf4 is not lossless, and model weights for linear layers are lossy compared to the official base model.
Training Details 
Methodology:
Upscaled fp16 variant after nf4 4-bit quantization
Input Output 
Accepted Modalities:
text
Performance Tips:
Upscaling helps avoid quantization/dequantization costs for each inference pass.
LLM NameCodellama 70B Instruct Nf4 Fp16 Upscaled
Repository πŸ€—https://huggingface.co/arnavgrg/codellama-70b-instruct-nf4-fp16-upscaled 
Model Size70b
Required VRAM138.7 GB
Updated2026-08-02
Maintainerarnavgrg
Model Typellama
Instruction-BasedYes
Model Files  4.7 GB: 1-of-29   4.7 GB: 2-of-29   5.0 GB: 3-of-29   5.0 GB: 4-of-29   4.7 GB: 5-of-29   4.7 GB: 6-of-29   4.7 GB: 7-of-29   5.0 GB: 8-of-29   5.0 GB: 9-of-29   4.7 GB: 10-of-29   4.7 GB: 11-of-29   4.7 GB: 12-of-29   5.0 GB: 13-of-29   5.0 GB: 14-of-29   4.7 GB: 15-of-29   4.7 GB: 16-of-29   4.7 GB: 17-of-29   5.0 GB: 18-of-29   5.0 GB: 19-of-29   4.7 GB: 20-of-29   4.7 GB: 21-of-29   4.7 GB: 22-of-29   5.0 GB: 23-of-29   5.0 GB: 24-of-29   4.7 GB: 25-of-29   4.7 GB: 26-of-29   4.7 GB: 27-of-29   5.0 GB: 28-of-29   3.8 GB: 29-of-29
Quantization Typefp16
Generates CodeYes
Model ArchitectureLlamaForCausalLM
Licenseapache-2.0
Context Length4096
Model Max Length4096
Transformers Version4.37.0
Tokenizer ClassLlamaTokenizer
Padding Token</s>
Vocabulary Size32016
Torch Data Typefloat16

Best Alternatives to Codellama 70B Instruct Nf4 Fp16 Upscaled

Best Alternatives
Context / RAM
Downloads
Likes
...Llama 70B Instruct Hf 4bit MLX4K / 39.1 GB41725
...70B Instruct Hf 5.0bpw H6 EXL22K / 43.6 GB56
...70B Instruct Hf 2.4bpw H6 EXL22K / 21.3 GB61
...0B Instruct Hf 2.65bpw H6 EXL22K / 23.4 GB53
...70B Instruct Hf 4.0bpw H6 EXL22K / 35.1 GB61
CodeLlama 70B Instruct Hf4K / 72.3 GB72924
Code Llama 70B Python Instruct4K / 138.1 GB51
CodeLlama 70B Instruct Hf4K / 72.3 GB463210
CodeLlama 70B Instruct Hf GGUF4K / 25.5 GB4872
CodeLlama 70B Instruct AWQ4K / 36.6 GB29713
Note: green Score (e.g. "73.2") means that the model is better than arnavgrg/codellama-70b-instruct-nf4-fp16-upscaled.