LLM EXPLORER 63,407 MODELS INDEXED

TinyLlama QuantumQuill Chat by Stefan171

By Stefan171 · 5 downloads

TinyLlama QuantumQuill Chat is an open-source language model by Stefan171. Features: LLM, VRAM: 2.2GB, Context: 15K, License: apache-2.0, Quantized, LLM Explorer Score: 0.16, Arc: 29.2, HellaSwag: 48.8, MMLU: 23.9, GSM8K: 4.9.

  4bit   Autotrain compatible Base model:finetune:unsloth/ti... Base model:unsloth/tinyllama-c...   Dataset:meta-math/metamathqa   En   Endpoints compatible   Llama   Pytorch   Quantized   Region:us   Sft   Trl   Unsloth

TinyLlama QuantumQuill Chat Parameters and Internals

Model Type 
text-generation-inference, transformers, unsloth, llama, trl, sft
Additional Notes 
This llama model was trained 2x faster with Unsloth and Huggingface's TRL library.
LLM NameTinyLlama QuantumQuill Chat
Repository πŸ€—https://huggingface.co/Stefan171/TinyLlama-QuantumQuill-chat 
Base Model(s)  Tinyllama Chat Bnb 4bit   unsloth/tinyllama-chat-bnb-4bit
Required VRAM2.2 GB
Updated2025-10-20
MaintainerStefan171
Model Typellama
Model Files  2.2 GB
Supported Languagesen
Quantization Type4bit
Model ArchitectureLlamaForCausalLM
Licenseapache-2.0
Context Length16000
Model Max Length16000
Transformers Version4.40.1
Tokenizer ClassLlamaTokenizer
Padding Token<unk>
Vocabulary Size32000
Torch Data Typefloat16

Best Alternatives to TinyLlama QuantumQuill Chat

Best Alternatives
Context / RAM
Downloads
Likes
... Text Chat 512K 6.0bpw H6 EXL2512K / 5.2 GB61
... Text Chat 512K 5.0bpw H6 EXL2512K / 4.4 GB11
Slimorca Phi 3.5128K / 7.6 GB70
Phi 3.5 Instruct Vul128K / 7.6 GB170
... Text Chat 128K 6.0bpw H6 EXL2128K / 5.2 GB31
QuantumQuill16K / 2.2 GB50
Unsloth Phi 4 4bit16K / 8.3 GB1396
...ocalAI Functioncall Phi 4 V0.316K / 29.4 GB9110
...ocalAI Functioncall Phi 4 V0.316K / 29.4 GB68
PARM V2 Phi 4 4K CoT PyTorch16K / 29.4 GB01
Note: green Score (e.g. "73.2") means that the model is better than Stefan171/TinyLlama-QuantumQuill-chat.