LLM EXPLORER 62,076 MODELS INDEXED

Koishi 120B Qlora Gptq by ewof

By ewof · 4 downloads

Koishi 120B Qlora Gptq is an open-source language model by ewof. Features: 120b LLM, VRAM: 59.8GB, Context: 4K, Quantized, Instruction-Based, LLM Explorer Score: 0.1.

  4bit Dataset:ewof/koishi-instruct-m...   Endpoints compatible   Gptq   Instruct   Llama   Quantized   Region:us   Sharded

Koishi 120B Qlora Gptq Parameters and Internals

Training Details 
Data Sources:
ewof/koishi-instruct-metharme
Methodology:
Trained on prompts using roles denoted by tokens: '<|system|>', '<|user|>', and '<|model|>'.
Context Length:
2048
Hardware Used:
8x Nvidia A100 GPU cluster
LLM NameKoishi 120B Qlora Gptq
Repository πŸ€—https://huggingface.co/ewof/koishi-120b-qlora-gptq 
Base Model(s)  Koishi 120B Qlora   ewof/koishi-120b-qlora
Model Size120b
Required VRAM59.8 GB
Updated2026-07-31
Maintainerewof
Model Typellama
Instruction-BasedYes
Model Files  10.0 GB: 1-of-6   9.9 GB: 2-of-6   10.0 GB: 3-of-6   10.0 GB: 4-of-6   10.0 GB: 5-of-6   9.9 GB: 6-of-6
GPTQ QuantizationYes
Quantization Typegptq|4bit
Model ArchitectureLlamaForCausalLM
Context Length4096
Model Max Length4096
Transformers Version4.35.1
Tokenizer ClassLlamaTokenizer
Padding Token</s>
Vocabulary Size32003
Torch Data Typefloat16

Best Alternatives to Koishi 120B Qlora Gptq

Best Alternatives
Context / RAM
Downloads
Likes
MegaDolphin 120B GPTQ4K / 61.1 GB254
...t 120B Cat A Llama EXL2 5.5bpw8K / 85.3 GB50
...t 120B Cat A Llama EXL2 4.5bpw8K / 70.3 GB41
...egaDolphin 120B 2.9bpw H6 EXL24K / 44.3 GB23
...gaDolphin 120B 2.65bpw H6 EXL24K / 40.5 GB32
...egaDolphin 120B 4.0bpw H6 EXL24K / 60.8 GB41
...ma 3 Instruct 120B Cat A Llama8K / 243.9 GB21
Meta Llama 3 225B Instruct8K / 443.2 GB27418
...0B Instruct Abliterated Merged8K / 243.7 GB11
MegaDolphin 120B AWQ4K / 63.3 GB72
Note: green Score (e.g. "73.2") means that the model is better than ewof/koishi-120b-qlora-gptq.