LLM EXPLORER 59,420 MODELS INDEXED

Vicuna 33B Coder AWQ by TheBloke

By TheBloke · 5 downloads

Vicuna 33B Coder AWQ is an open-source language model by TheBloke. Features: 33b LLM, VRAM: 17.6GB, Context: 2K, License: other, Quantized, Code Generating, LLM Explorer Score: 0.09.

  Arxiv:1910.09700   4-bit   Awq Base model:felixchao/vicuna-33... Base model:quantized:felixchao...   Code   Codegen   Llama   Model-index   Quantized   Region:us   Safetensors   Sharded   Tensorflow

Vicuna 33B Coder AWQ Parameters and Internals

Model Type 
text-generation
Additional Notes 
AWQ is an efficient, accurate, and fast low-bit quantization method, supporting 4-bit quantization for this model.
LLM NameVicuna 33B Coder AWQ
Repository πŸ€—https://huggingface.co/TheBloke/vicuna-33B-coder-AWQ 
Model NameVicuna 33B Coder
Model CreatorChao Chang-Yu
Base Model(s)  Vicuna 33B Coder   FelixChao/vicuna-33b-coder
Model Size33b
Required VRAM17.6 GB
Updated2026-07-24
MaintainerTheBloke
Model Typellama
Model Files  10.0 GB: 1-of-2   7.6 GB: 2-of-2
AWQ QuantizationYes
Quantization Typeawq
Generates CodeYes
Model ArchitectureLlamaForCausalLM
Licenseother
Context Length2048
Model Max Length2048
Transformers Version4.34.0
Tokenizer ClassLlamaTokenizer
Beginning of Sentence Token<s>
End of Sentence Token</s>
Unk Token<unk>
Vocabulary Size32000
Torch Data Typefloat16

Best Alternatives to Vicuna 33B Coder AWQ

Best Alternatives
Context / RAM
Downloads
Likes
Everyone Coder 33B Base AWQ16K / 18.1 GB82
...eepseek Coder 33B Instruct AWQ16K / 18.1 GB73142
Deepseek Coder 33B Base AWQ16K / 18.1 GB224
...erpreter DS 33B 4.0bpw H6 EXL216K / 17.1 GB84
...rpreter DS 33B 4.65bpw H6 EXL216K / 19.8 GB53
...erpreter DS 33B 5.0bpw H6 EXL216K / 21.2 GB61
...erpreter DS 33B 6.0bpw H6 EXL216K / 25.3 GB80
...erpreter DS 33B 8.0bpw H8 EXL216K / 33.5 GB51
...der 33B V2 Base 8.0bpw H8 EXL216K / 33.5 GB11
... Coder 33B Base 6.0bpw H6 EXL216K / 25.3 GB31
Note: green Score (e.g. "73.2") means that the model is better than TheBloke/vicuna-33B-coder-AWQ.