LLM EXPLORER 59,420 MODELS INDEXED

RAFT Llama3 Quantized 5epochs 24 06 04 by Vipinap

By Vipinap · 7 downloads

RAFT Llama3 Quantized 5epochs 24 06 04 is an open-source language model by Vipinap. Features: 8.2b LLM, VRAM: 5.8GB, Context: 8K, LLM Explorer Score: 0.13.

  Arxiv:1910.09700   4-bit   Bitsandbytes   Conversational   Endpoints compatible   Llama   Region:us   Safetensors   Sharded   Tensorflow   Unsloth

RAFT Llama3 Quantized 5epochs 24 06 04 Parameters and Internals

Additional Notes 
This is the model card of a πŸ€— transformers model that has been pushed on the Hub. This model card has been automatically generated. The model description and other specific sections require more information to be filled in.
LLM NameRAFT Llama3 Quantized 5epochs 24 06 04
Repository πŸ€—https://huggingface.co/Vipinap/RAFT_llama3_quantized_5epochs_24_06_04 
Model Size8.2b
Required VRAM5.8 GB
Updated2026-07-23
MaintainerVipinap
Model Typellama
Model Files  4.7 GB: 1-of-2   1.1 GB: 2-of-2
Model ArchitectureLlamaForCausalLM
Context Length8192
Model Max Length8192
Transformers Version4.39.3
Tokenizer ClassPreTrainedTokenizerFast
Padding Token<|eot_id|>
Vocabulary Size128256
Torch Data Typebfloat16

Best Alternatives to RAFT Llama3 Quantized 5epochs 24 06 04

Best Alternatives
Context / RAM
Downloads
Likes
Ko Pt Model Test18K / 16.4 GB20270
Hola8K / 5.8 GB60
Llama SciQ 4bits8K / 5.8 GB120
RAFT Llama3 Version 24.06.018K / 5.8 GB50
Llama 3 Merged Linear8K / 5.8 GB160
Llama 3 Ko Luxia Instruct8K / 16.4 GB63
KernelLLM Bnb 4bit128K / 5.7 GB91
KernelLLM Unsloth Bnb 4bit128K / 6 GB62
Llama3.1 Merged128K / 5.8 GB60
EPFL TA Meister 4bit8K / 5.8 GB60
Note: green Score (e.g. "73.2") means that the model is better than Vipinap/RAFT_llama3_quantized_5epochs_24_06_04.