LLM EXPLORER 60,783 MODELS INDEXED

CodeLlama 34B Instruct Fp16 by TheBloke

By TheBloke · 131 downloads

CodeLlama 34B Instruct Fp16 is an open-source language model by TheBloke. Features: 34b LLM, VRAM: 67.5GB, Context: 16K, License: llama2, Quantized, Instruction-Based, Code Generating, LLM Explorer Score: 0.23, ELO: 1136, Arc: 40.8, HellaSwag: 35.7, MMLU: 39.7, GSM8K: 23.1.

  Codegen   Codellama   Custom code   Endpoints compatible   Fp16   Instruct   Llama   Llama2   Quantized   Region:us   Safetensors   Sharded   Tensorflow

CodeLlama 34B Instruct Fp16 Benchmarks

nn.n% — How the model compares to the reference models: Anthropic Sonnet 3.5 ("so35"), GPT-4o ("gpt4o") or GPT-4 ("gpt4").

CodeLlama 34B Instruct Fp16 Parameters and Internals

Model Type 
instruction following, code synthesis, text generation
Use Cases 
Areas:
Commercial, Research
Applications:
Code synthesis, Code understanding, Code assistive applications
Primary Use Cases:
Code assistant, Code generation
Limitations:
Use in languages other than English, Code use not covered by safety evaluations
Considerations:
Perform safety testing tailored to specific applications.
Additional Notes 
Model variants are designed for Python and safer deployment applications.
Supported Languages 
English (Full support), Python (Highly specialized support), Other programming languages (Partial support)
Training Details 
Methodology:
Fine-tuned with instruction data
Context Length:
100000
Hardware Used:
A100-80GB
Model Architecture:
Autoregressive, transformer architecture
LLM NameCodeLlama 34B Instruct Fp16
Repository πŸ€—https://huggingface.co/TheBloke/CodeLlama-34B-Instruct-fp16 
Base Model(s)  CodeLlama 34B Instruct Hf   premai-io/CodeLlama-34b-Instruct-hf
Model Size34b
Required VRAM67.5 GB
Updated2026-08-07
MaintainerTheBloke
Model Typellama
Instruction-BasedYes
Model Files  9.8 GB: 1-of-7   9.7 GB: 2-of-7   9.7 GB: 3-of-7   9.7 GB: 4-of-7   9.7 GB: 5-of-7   9.7 GB: 6-of-7   9.2 GB: 7-of-7
Quantization Typefp16
Generates CodeYes
Model ArchitectureLlamaForCausalLM
Licensellama2
Context Length16384
Model Max Length16384
Transformers Version4.32.0
Tokenizer ClassLlamaTokenizer
Beginning of Sentence Token<s>
End of Sentence Token</s>
Unk Token<unk>
Vocabulary Size32000
Torch Data Typefloat16

Best Alternatives to CodeLlama 34B Instruct Fp16

Best Alternatives
Context / RAM
Downloads
Likes
CodeLlama 34B Instruct Hf 4bit16K / 19.4 GB632
...gpt 32K Codellama 34B Instruct32K / 67.5 GB562
CodeLlama 34B Instruct Hf16K / 67.5 GB22872304
CodeLlama 34B Instruct Hf16K / 1.4 GB73
CodeLlama 34B Instruct Hf16K / 1.4 GB53
Speechless Codellama 34B V2.016K / 67.5 GB9617
CodeLlama 34B Instruct Hf16K / 67.5 GB62219
Speechless Codellama 34B V1.916K / 67.5 GB2860
XAgentLLaMa 34B Preview16K / 157.3 GB33
CodeFuse CodeLlama 34B16K / 67.5 GB35093
Note: green Score (e.g. "73.2") means that the model is better than TheBloke/CodeLlama-34B-Instruct-fp16.