LLM EXPLORER 60,645 MODELS INDEXED

Llama 3 8B Instruct 64K by MaziyarPanahi

By MaziyarPanahi · 28 downloads

Llama 3 8B Instruct 64K is an open-source language model by MaziyarPanahi. Features: 8b LLM, VRAM: 16.1GB, Context: 8K, License: llama3, Instruction-Based, LLM Explorer Score: 0.12.

  Arxiv:2309.10400   64k   Axolotl Base model:finetune:winglian/l... Base model:winglian/llama-3-8b...   Conversational   Dataset:intel/orca dpo pairs   Deploy:azure   Dpo   En   Facebook   Finetuned   Instruct   Llama   Llama-3   Meta   Pose   Pytorch   Region:us   Safetensors   Sharded   Tensorflow

Llama 3 8B Instruct 64K Parameters and Internals

Model Type 
text generation
Additional Notes 
This model uses PoSE to extend Llama's context length from 8k to 64k.
Supported Languages 
en (proficient)
Training Details 
Data Sources:
RedPajama V1 dataset
Data Volume:
300M tokens
Methodology:
rank stabilized LoRA of rank 256
Context Length:
64000
Model Architecture:
Llama-3
LLM NameLlama 3 8B Instruct 64K
Repository πŸ€—https://huggingface.co/MaziyarPanahi/Llama-3-8B-Instruct-64k 
Model NameLlama-3-8B-Instruct-64k
Model CreatorMaziyarPanahi
Base Model(s)  Llama 3 8B 64K PoSE   winglian/Llama-3-8b-64k-PoSE
Model Size8b
Required VRAM16.1 GB
Updated2026-08-04
MaintainerMaziyarPanahi
Model Typellama
Instruction-BasedYes
Model Files  5.0 GB: 1-of-4   5.0 GB: 2-of-4   4.9 GB: 3-of-4   1.2 GB: 4-of-4
Supported Languagesen
Model ArchitectureLlamaForCausalLM
Licensellama3
Context Length8192
Model Max Length8192
Transformers Version4.40.0.dev0
Tokenizer ClassPreTrainedTokenizerFast
Vocabulary Size128256
Torch Data Typefloat16

Quantized Models of the Llama 3 8B Instruct 64K

Model
Likes
Downloads
VRAM
Llama 3 8B Instruct 64K GGUF14567403 GB
... Instruct 64K HQQ 1bit Smashed153 GB
Llama 3 8B Instruct 64K AWQ075 GB

Best Alternatives to Llama 3 8B Instruct 64K

Best Alternatives
Context / RAM
Downloads
Likes
...otron 8B UltraLong 4M Instruct4192K / 32.1 GB1135125
UltraLong Thinking4192K / 16.1 GB23
...a 3.1 8B UltraLong 4M Instruct4192K / 32.1 GB17624
...a 3.1 8B UltraLong 2M Instruct2096K / 32.1 GB8759
...otron 8B UltraLong 2M Instruct2096K / 32.1 GB12418
Cthulhu 8B V1.41048K / 16.1 GB1010
...raLong 1M Instruct Abliterated1048K / 32.1 GB49
...a 3.1 8B UltraLong 1M Instruct1048K / 32.1 GB138729
...otron 8B UltraLong 1M Instruct1048K / 32.1 GB70259
Zero Llama 3.1 8B Beta61048K / 16.1 GB71
Note: green Score (e.g. "73.2") means that the model is better than MaziyarPanahi/Llama-3-8B-Instruct-64k.