LLM EXPLORER 59,420 MODELS INDEXED

Llama 3 70B Quantised by farhan-ahmad

By farhan-ahmad · 13 downloads

Llama 3 70B Quantised is an open-source language model by farhan-ahmad. Features: 70b LLM, VRAM: 48.7GB, Context: 8K, License: mit, Quantized, LLM Explorer Score: 0.13.

  Conversational   Endpoints compatible   Gguf   Llama   Quantized   Region:us

Llama 3 70B Quantised Parameters and Internals

LLM NameLlama 3 70B Quantised
Repository πŸ€—https://huggingface.co/farhan-ahmad/llama-3-70b-quantised 
Model Size70b
Required VRAM48.7 GB
Updated2026-08-06
Maintainerfarhan-ahmad
Model Typellama
Model Files  48.7 GB
GGUF QuantizationYes
Quantization Typegguf
Model ArchitectureLlamaForCausalLM
Licensemit
Context Length8192
Model Max Length8192
Transformers Version4.40.0.dev0
Vocabulary Size128256
Torch Data Typebfloat16

Best Alternatives to Llama 3 70B Quantised

Best Alternatives
Context / RAM
Downloads
Likes
Grok Oss Revenant 70B128K / 141.9 GB85032
...Seek R1 Distill Llama 70B GGUF128K / 15.9 GB27430120
Llama 3.3 70B Instruct GGUF128K / 15.9 GB23398124
R1 1776 Distill Llama 70B GGUF128K / 26.4 GB247824
Reflection Llama 3.1 70B Bf16128K / 141.9 GB2836
Reflection Llama 3.1 70B GGUF128K / 26.4 GB1155
...Horizon AI Korean Advanced 70B128K / 141.9 GB301
Midnight Miqu 70B V1.0 GGUF31K / 29.9 GB3984
...qu 1 70B 24GB VRAM IQ2 XS SOTA31K / 20.3 GB681
...ma3 70B Chinese Chat GGUF 4bit8K / 40 GB19618
Note: green Score (e.g. "73.2") means that the model is better than farhan-ahmad/llama-3-70b-quantised.