LLM EXPLORER 59,516 MODELS INDEXED

Llama 3 15B Instruct Ft V2 AWQ by solidrust

By solidrust · 7 downloads

Llama 3 15B Instruct Ft V2 AWQ is an open-source language model by solidrust. Features: 15b LLM, VRAM: 9.4GB, Context: 8K, Quantized, Instruction-Based, LLM Explorer Score: 0.14, Arc: 61.4, HellaSwag: 78.8, MMLU: 67.3, GSM8K: 72.5.

  4-bit   Autotrain compatible   Awq Base model:elinas/llama-3-15b-... Base model:quantized:elinas/ll...   Conversational   Endpoints compatible   Instruct   Llama   Quantized   Region:us   Safetensors   Sharded   Tensorflow

Llama 3 15B Instruct Ft V2 AWQ Parameters and Internals

Model Type 
text-generation
Additional Notes 
AWQ is a low-bit weight quantization method for faster Transformers-based inference with 4-bit quantization.
Input Output 
Accepted Modalities:
text
LLM NameLlama 3 15B Instruct Ft V2 AWQ
Repository πŸ€—https://huggingface.co/solidrust/Llama-3-15B-Instruct-ft-v2-AWQ 
Base Model(s)  Llama 3 15B Instruct Ft V2   elinas/Llama-3-15B-Instruct-ft-v2
Model Size15b
Required VRAM9.4 GB
Updated2026-07-23
Maintainersolidrust
Model Typellama
Instruction-BasedYes
Model Files  5.0 GB: 1-of-2   4.4 GB: 2-of-2
AWQ QuantizationYes
Quantization Typeawq
Model ArchitectureLlamaForCausalLM
Context Length8192
Model Max Length8192
Transformers Version4.41.1
Tokenizer ClassPreTrainedTokenizerFast
Padding Token<|end_of_text|>
Vocabulary Size128256
Torch Data Typefloat16

Best Alternatives to Llama 3 15B Instruct Ft V2 AWQ

Best Alternatives
Context / RAM
Downloads
Likes
L3 Aethora 15B V2 EXL2 6.0bpw8K / 11.9 GB61
L3 Aethora 15B V2 EXL2 8.0bpw8K / 12.4 GB41
L3 Aethora 15B V2 EXL2 4.0bpw8K / 8.4 GB21
L3 Aethora 15B V2 EXL2 5.0bpw8K / 10.2 GB01
Hamanasu 15B Instruct16K / 29.4 GB4739
L3 Aethora 15B V28K / 30.1 GB2242
Meta Llama 3 15B Instruct8K / 23.9 GB31
OpenCrystal 15B L3 V38K / 30 GB1310
OpenCrystal L3 15B V2.18K / 30 GB35
OpenCrystal 15B L3 V28K / 30 GB36
Note: green Score (e.g. "73.2") means that the model is better than solidrust/Llama-3-15B-Instruct-ft-v2-AWQ.