LLM EXPLORER 60,645 MODELS INDEXED

Gradientai Llama 3 8B Instruct 262K AWQ 4bit Smashed by PrunaAI

By PrunaAI · 5 downloads

Gradientai Llama 3 8B Instruct 262K AWQ 4bit Smashed is an open-source language model by PrunaAI. Features: 8b LLM, VRAM: 5.8GB, Context: 256K, Quantized, Instruction-Based, LLM Explorer Score: 0.12.

  4-bit   4bit   Awq Base model:prunaai/gradientai-... Base model:quantized:prunaai/g...   Instruct   Llama   Pruna-ai   Quantized   Region:us   Safetensors   Sharded   Tensorflow

Gradientai Llama 3 8B Instruct 262K AWQ 4bit Smashed Parameters and Internals

Model Type 
compression, efficiency
Additional Notes 
The model uses safetensors format and calibration data from WikiText for compression. Different metrics for efficiency evaluation include inference latency, memory, energy consumption, and CO2 emissions. Efficiency evaluation results might vary based on the hardware and settings used.
LLM NameGradientai Llama 3 8B Instruct 262K AWQ 4bit Smashed
Repository πŸ€—https://huggingface.co/PrunaAI/gradientai-Llama-3-8B-Instruct-262k-AWQ-4bit-smashed 
Base Model(s)  ...Instruct 262K AWQ 4bit Smashed   PrunaAI/gradientai-Llama-3-8B-Instruct-262k-AWQ-4bit-smashed
Model Size8b
Required VRAM5.8 GB
Updated2025-10-02
MaintainerPrunaAI
Model Typellama
Instruction-BasedYes
Model Files  4.7 GB: 1-of-2   1.1 GB: 2-of-2
AWQ QuantizationYes
Quantization Typeawq|4bit
Model ArchitectureLlamaForCausalLM
Context Length262144
Model Max Length262144
Transformers Version4.40.0
Tokenizer ClassPreTrainedTokenizerFast
Vocabulary Size128256
Torch Data Typefloat16

Quantized Models of the Gradientai Llama 3 8B Instruct 262K AWQ 4bit Smashed

Model
Likes
Downloads
VRAM
...Instruct 262K AWQ 4bit Smashed455 GB

Best Alternatives to Gradientai Llama 3 8B Instruct 262K AWQ 4bit Smashed

Best Alternatives
Context / RAM
Downloads
Likes
...8B Instruct Gradient 1048K AWQ1024K / 5.8 GB80
...radient 1048K AWQ 4bit Smashed1024K / 5.8 GB71
...ta Llama 3 8B Instruct 64K AWQ64K / 5.8 GB60
... Instruct 8B 32K V0.1 4bit AWQ64K / 5.8 GB90
Llama 3 8B Instruct AWQ8K / 5.8 GB2545831
Meta Llama 3 8B Instruct AWQ8K / 5.8 GB60
Meta Llama 3 8B Instruct AWQ8K / 5.8 GB5630
Meta Llama 3 8B Instruct AWQ8K / 5.8 GB70
Meta Llama 3 8B Instruct AWQ8K / 5.8 GB4015
...eta Llama 3 8B Instruct Hf AWQ8K / 5.8 GB1799
Note: green Score (e.g. "73.2") means that the model is better than PrunaAI/gradientai-Llama-3-8B-Instruct-262k-AWQ-4bit-smashed.