LLM EXPLORER 61,491 MODELS INDEXED

Yarn Llama 2 7B 128K GGML by TheBloke

By TheBloke · 3 downloads

Yarn Llama 2 7B 128K GGML is an open-source language model by TheBloke. Features: 7b LLM, VRAM: 2.9GB, License: llama2, Quantized, LLM Explorer Score: 0.08.

  Arxiv:2309.00071 Base model:finetune:nousresear... Base model:nousresearch/yarn-l...   Dataset:pg19   Ggml   Llama   Quantized   Region:us   Yarn

Yarn Llama 2 7B 128K GGML Parameters and Internals

Model Type 
llama
Additional Notes 
The GGML format has now been superseded by GGUF. As of August 21st 2023, llama.cpp no longer supports GGML models.
Training Details 
Data Sources:
PG19 dataset
Methodology:
Further pretrained on a subset of the PG19 dataset.
Context Length:
128000
LLM NameYarn Llama 2 7B 128K GGML
Repository πŸ€—https://huggingface.co/TheBloke/Yarn-Llama-2-7B-128K-GGML 
Model NameYarn Llama 2 7B 128K
Model CreatorNousResearch
Base Model(s)  Yarn Llama 2 7B 128K   NousResearch/Yarn-Llama-2-7b-128k
Model Size7b
Required VRAM2.9 GB
Updated2026-07-29
MaintainerTheBloke
Model Typellama
Model Files  2.9 GB   3.6 GB   3.3 GB   3.0 GB   3.8 GB   4.2 GB   4.1 GB   3.8 GB   4.7 GB   5.1 GB   4.8 GB   4.7 GB   5.5 GB   7.1 GB
GGML QuantizationYes
Quantization Typeggml
Model ArchitectureAutoModel
Licensellama2

Best Alternatives to Yarn Llama 2 7B 128K GGML

Best Alternatives
Context / RAM
Downloads
Likes
Llama 2 7B Chat GGML0K / 2.9 GB157872
Llama 2 GGML Medical Chatbot0K /  GB1445
Llama 2 7B GGML0K / 2.9 GB147219
Yarn Llama 2 7B 64K GGML0K / 2.9 GB53
CodeLlama 7B GGML0K / 3 GB1027
CodeLlama 7B Python GGML0K / 2.9 GB624
CodeLlama 7B Instruct GGML0K / 3 GB1020
Airoboros L2 7B 2.1 GGML0K / 2.9 GB51
Zarafusionex 1.1 L2 7B GGML0K / 2.9 GB62
EDGE 0 7B GGML0K / 2.9 GB21
Note: green Score (e.g. "73.2") means that the model is better than TheBloke/Yarn-Llama-2-7B-128K-GGML.