LLM EXPLORER 59,420 MODELS INDEXED

Llama3 Finetune 4bit by Sreevadan

By Sreevadan · 0 downloads

Llama3 Finetune 4bit is an open-source language model by Sreevadan. Features: 4.6b LLM, VRAM: 6.1GB, Context: 8K, License: apache-2.0, Quantized, Fine-Tuned.

  Arxiv:1910.09700   4-bit   4bit   Autotrain compatible   Bitsandbytes   Dataset:open-orca/openorca   Endpoints compatible   Finetuned   License:apache-2.0   Llama   Quantized   Region:us   Safetensors   Sharded   Tensorflow

Llama3 Finetune 4bit Parameters and Internals

LLM NameLlama3 Finetune 4bit
Repository πŸ€—https://huggingface.co/Sreevadan/Llama3-finetune-4bit 
Model Size4.6b
Required VRAM6.1 GB
Updated2024-07-04
MaintainerSreevadan
Model Typellama
Model Files  5.0 GB: 1-of-2   1.1 GB: 2-of-2
Quantization Type4bit
Model ArchitectureLlamaForCausalLM
Licenseapache-2.0
Context Length8192
Model Max Length8192
Transformers Version4.41.2
Tokenizer ClassPreTrainedTokenizerFast
Padding Token<|reserved_special_token_250|>
Vocabulary Size128256
Torch Data Typefloat16

Best Alternatives to Llama3 Finetune 4bit

Best Alternatives
Context / RAM
Downloads
Likes
Q4 Llama 38K / 6.1 GB120
TinyLLama V02K / 0 GB23117645
SimpleLlamaSentences2K / 0 GB60
UniversalNER TinyLLama2K / 0 GB101
UniversalNER TinyLLama2K / 0 GB61
Note: green Score (e.g. "73.2") means that the model is better than Sreevadan/Llama3-finetune-4bit.