LLM EXPLORER 59,516 MODELS INDEXED

Meta Llama 3.1 8B Instruct Quantized.w8a16 by RedHatAI

By RedHatAI · 725 downloads

Meta Llama 3.1 8B Instruct Quantized.w8a16 is an open-source language model by RedHatAI. Features: 8b LLM, VRAM: 9.1GB, Context: 128K, License: meta, Instruction-Based, LLM Explorer Score: 0.14.

  Arxiv:2210.17323 Base model:meta-llama/llama-3.... Base model:quantized:meta-llam...   Compressed-tensors   Conversational   De   En   Endpoints compatible   Es   Fr   Hi   Instruct   Int8   It   Llama   Pt   Region:us   Safetensors   Sharded   Tensorflow   Th   Vllm

Meta Llama 3.1 8B Instruct Quantized.w8a16 Parameters and Internals

LLM NameMeta Llama 3.1 8B Instruct Quantized.w8a16
Repository πŸ€—https://huggingface.co/RedHatAI/Meta-Llama-3.1-8B-Instruct-quantized.w8a16 
Base Model(s)  meta-llama/Meta-Llama-3.1-8B-Instruct   meta-llama/Meta-Llama-3.1-8B-Instruct
Model Size8b
Required VRAM9.1 GB
Updated2026-08-07
MaintainerRedHatAI
Model Typellama
Instruction-BasedYes
Model Files  5.0 GB: 1-of-2   4.1 GB: 2-of-2
Supported Languagesen de fr it pt hi es th
Model ArchitectureLlamaForCausalLM
Licensemeta
Context Length131072
Model Max Length131072
Transformers Version4.43.1
Tokenizer ClassPreTrainedTokenizerFast
Vocabulary Size128256
Torch Data Typebfloat16

Best Alternatives to Meta Llama 3.1 8B Instruct Quantized.w8a16

Best Alternatives
Context / RAM
Downloads
Likes
...otron 8B UltraLong 4M Instruct4192K / 32.1 GB1135125
UltraLong Thinking4192K / 16.1 GB23
...a 3.1 8B UltraLong 4M Instruct4192K / 32.1 GB17624
...a 3.1 8B UltraLong 2M Instruct2096K / 32.1 GB8759
...otron 8B UltraLong 2M Instruct2096K / 32.1 GB12418
Cthulhu 8B V1.41048K / 16.1 GB1010
...raLong 1M Instruct Abliterated1048K / 32.1 GB49
...a 3.1 8B UltraLong 1M Instruct1048K / 32.1 GB138729
...otron 8B UltraLong 1M Instruct1048K / 32.1 GB70259
Zero Llama 3.1 8B Beta61048K / 16.1 GB71
Note: green Score (e.g. "73.2") means that the model is better than RedHatAI/Meta-Llama-3.1-8B-Instruct-quantized.w8a16.