LLM EXPLORER 62,155 MODELS INDEXED

Llama 7B 4bit Act by wcde

By wcde · 11 downloads

Llama 7B 4bit Act is an open-source language model by wcde. Features: 7b LLM, VRAM: 3.8GB, Quantized, LLM Explorer Score: 0.06.

  4bit   Endpoints compatible   Llama   Quantized   Region:us
Model Card on HF πŸ€—: https://huggingface.co/wcde/llama-7b-4bit-act 

Llama 7B 4bit Act Parameters and Internals

Additional Notes 
Model generation options include: - --wbits 4: Using 4-bit weight quantization for model efficiency. - --act-order: Activation order optimization. - --true-sequential: Enforcing true sequential execution for improved training. - --new-eval: Utilizing a new evaluation strategy. - --faster-kernel: Deploying a faster kernel for computational efficiency.
LLM NameLlama 7B 4bit Act
Repository πŸ€—https://huggingface.co/wcde/llama-7b-4bit-act 
Base Model(s)  Todd Proxy LoRA 7b   autobots/Todd_Proxy_LoRA_7b
Model Size7b
Required VRAM3.8 GB
Updated2026-07-27
Maintainerwcde
Model Typellama
Model Files  3.8 GB
Quantization Type4bit
Model ArchitectureLLaMAForCausalLM
Transformers Version4.27.0.dev0
Tokenizer ClassLlamaTokenizer
Vocabulary Size32000
Torch Data Typefloat16

Best Alternatives to Llama 7B 4bit Act

Best Alternatives
Context / RAM
Downloads
Likes
Llama 7B Onnx Merged Fp162K /  GB96
Alpaca 7B Native 4bit0K / 4.5 GB144
Alpaca Native 4bit0K / 4.5 GB658
Llama 7B 4bit Gr1280K / 4 GB114
Swallow 7B GPTQ4K / 4.1 GB51
Honest Llama2 Chat 7B2K / 13.5 GB2129
Llama 7B Onnx Merged Fp322K /  GB121
Chatdoctor0K / 27 GB6312
Explore LM 7B Math0K / 27 GB131
Explore LM Ext 7B Rewriting0K / 27 GB121
Note: green Score (e.g. "73.2") means that the model is better than wcde/llama-7b-4bit-act.