LLM EXPLORER 62,155 MODELS INDEXED

Llama 7B 4bit Gr128 by wcde

By wcde · 11 downloads

Llama 7B 4bit Gr128 is an open-source language model by wcde. Features: 7b LLM, VRAM: 4GB, Quantized, LLM Explorer Score: 0.06.

  4bit   Endpoints compatible   Llama   Quantized   Region:us

Llama 7B 4bit Gr128 Parameters and Internals

Additional Notes 
The model was generated using specific settings: --wbits 4 --groupsize 128 --true-sequential --new-eval --faster-kernel.
LLM NameLlama 7B 4bit Gr128
Repository πŸ€—https://huggingface.co/wcde/llama-7b-4bit-gr128 
Model Size7b
Required VRAM4 GB
Updated2026-08-01
Maintainerwcde
Model Typellama
Model Files  4.0 GB
Quantization Type4bit
Model ArchitectureLLaMAForCausalLM
Transformers Version4.27.0.dev0
Tokenizer ClassLlamaTokenizer
Vocabulary Size32000
Torch Data Typefloat16

Best Alternatives to Llama 7B 4bit Gr128

Best Alternatives
Context / RAM
Downloads
Likes
Llama 7B Onnx Merged Fp162K /  GB96
Alpaca 7B Native 4bit0K / 4.5 GB144
Alpaca Native 4bit0K / 4.5 GB658
Llama 7B 4bit Act0K / 3.8 GB112
Swallow 7B GPTQ4K / 4.1 GB51
Honest Llama2 Chat 7B2K / 13.5 GB2129
Llama 7B Onnx Merged Fp322K /  GB121
Chatdoctor0K / 27 GB6312
Explore LM 7B Math0K / 27 GB131
Explore LM Ext 7B Rewriting0K / 27 GB121
Note: green Score (e.g. "73.2") means that the model is better than wcde/llama-7b-4bit-gr128.