LLM EXPLORER 59,516 MODELS INDEXED

Quantized Kheopss AWQ 4bits V1 by kheopss

By kheopss · 17 downloads

Quantized Kheopss AWQ 4bits V1 is an open-source language model by kheopss. Features: 1.2b LLM, VRAM: 4.2GB, Context: 32K, Quantized, LLM Explorer Score: 0.13.

  16bit   4-bit   Autotrain compatible   Awq   Conversational   Endpoints compatible   Mistral   Quantized   Region:us   Safetensors

Quantized Kheopss AWQ 4bits V1 Parameters and Internals

LLM NameQuantized Kheopss AWQ 4bits V1
Repository πŸ€—https://huggingface.co/kheopss/quantized_kheopss_AWQ_4bits_V1 
Model Size1.2b
Required VRAM4.2 GB
Updated2024-08-08
Maintainerkheopss
Model Typemistral
Model Files  4.2 GB
AWQ QuantizationYes
Quantization Typeawq|16bit
Model ArchitectureMistralForCausalLM
Context Length32768
Model Max Length32768
Transformers Version4.43.0.dev0
Tokenizer ClassLlamaTokenizer
Padding Token<unk>
Vocabulary Size32002
Torch Data Typefloat16

Best Alternatives to Quantized Kheopss AWQ 4bits V1

Best Alternatives
Context / RAM
Downloads
Likes
MistralLite AWQ32K / 4.2 GB423
SFR Embedding Mistral AWQ32K / 4.2 GB160
Yugo55A GPT AWQ32K / 4.2 GB60
Openchat 3.5 16K AWQ32K / 4.2 GB31
Openchat 3.5 0106 AWQ8K / 4.2 GB31
Hare1.0 Beta32K / 2.4 GB501
Mistralmawq32K / 4.2 GB140
... Finetune 16bit Ver9 Main GPTQ32K / 4.2 GB60
Turdus GPTQ32K / 4.2 GB85
Note: green Score (e.g. "73.2") means that the model is better than kheopss/quantized_kheopss_AWQ_4bits_V1.