LLM EXPLORER 59,420 MODELS INDEXED

Llama3 Ko 4bit by vessl

By vessl · 5 downloads

Llama3 Ko 4bit is an open-source language model by vessl. Features: 8.2b LLM, VRAM: 5.8GB, Context: 8K, Quantized, Merged, LLM Explorer Score: 0.12.

  Merged Model   Arxiv:1910.09700   4-bit   4bit   Bitsandbytes   Endpoints compatible   Llama   Quantized   Region:us   Safetensors   Sharded   Tensorflow
Model Card on HF πŸ€—: https://huggingface.co/vessl/llama3-ko-4bit 

Llama3 Ko 4bit Parameters and Internals

LLM NameLlama3 Ko 4bit
Repository πŸ€—https://huggingface.co/vessl/llama3-ko-4bit 
Merged ModelYes
Model Size8.2b
Required VRAM5.8 GB
Updated2026-04-24
Maintainervessl
Model Typellama
Model Files  4.7 GB: 1-of-2   1.1 GB: 2-of-2
Quantization Type4bit
Model ArchitectureLlamaForCausalLM
Context Length8192
Model Max Length8192
Transformers Version4.38.2
Tokenizer ClassPreTrainedTokenizerFast
Padding Token<|end_of_text|>
Vocabulary Size128256
Torch Data Typefloat16

Best Alternatives to Llama3 Ko 4bit

Best Alternatives
Context / RAM
Downloads
Likes
KernelLLM Bnb 4bit128K / 5.7 GB91
KernelLLM Unsloth Bnb 4bit128K / 6 GB62
Llama3.1 Merged128K / 5.8 GB60
EPFL TA Meister 4bit8K / 5.8 GB60
Book 4bitV58K / 5.8 GB60
Book 4bit8K / 5.8 GB60
Ko Pt Model Test18K / 16.4 GB20270
Maqa Llama 4bit8K / 5.8 GB90
Hola8K / 5.8 GB60
Llama SciQ 4bits8K / 5.8 GB120
Note: green Score (e.g. "73.2") means that the model is better than vessl/llama3-ko-4bit.