LLM EXPLORER 59,420 MODELS INDEXED

CodeFuse DeepSeek 33B by codefuse-ai

By codefuse-ai · 145 downloads

CodeFuse DeepSeek 33B is an open-source language model by codefuse-ai. Features: 33b LLM, VRAM: 66.5GB, Context: 16K, License: other, LLM Explorer Score: 0.11, HumanEval: 76.8.

  Code   Conversational   En   Endpoints compatible   Llama   Pytorch   Region:us   Safetensors   Sharded   Tensorflow   Zh

CodeFuse DeepSeek 33B Parameters and Internals

Model Type 
code-generation
Training Details 
Methodology:
QLoRA
Input Output 
Input Format:
Concatenated string in training data format
Accepted Modalities:
text
Output Format:
Generated code
Performance Tips:
Ensure input string ends with '\ bot' for generating answers.
Release Notes 
Date:
2024-01-12
Notes:
Released with pass@1 score of 78.65% on HumanEval.
LLM NameCodeFuse DeepSeek 33B
Repository 🤗https://huggingface.co/codefuse-ai/CodeFuse-DeepSeek-33B 
Model Size33b
Required VRAM66.5 GB
Updated2026-07-26
Maintainercodefuse-ai
Model Typellama
Model Files  9.7 GB: 1-of-7   9.9 GB: 2-of-7   9.9 GB: 3-of-7   9.8 GB: 4-of-7   9.9 GB: 5-of-7   9.9 GB: 6-of-7   7.4 GB: 7-of-7   9.7 GB: 1-of-7   9.9 GB: 2-of-7   9.9 GB: 3-of-7   9.8 GB: 4-of-7   9.9 GB: 5-of-7   9.9 GB: 6-of-7   7.4 GB: 7-of-7
Supported Languagesen zh
Model ArchitectureLlamaForCausalLM
Licenseother
Context Length16384
Model Max Length16384
Transformers Version4.34.1
Tokenizer ClassLlamaTokenizerFast
Beginning of Sentence Token<|begin▁of▁sentence|>
End of Sentence Token<|end▁of▁sentence|>
Vocabulary Size32256
Torch Data Typebfloat16

Quantized Models of the CodeFuse DeepSeek 33B

Model
Likes
Downloads
VRAM
CodeFuse DeepSeek 33B 4bits10718 GB

Best Alternatives to CodeFuse DeepSeek 33B

Best Alternatives
Context / RAM
Downloads
Likes
Llm Jp 4 33B Thinking64K / 66.8 GB19224
Llm Jp 4 33B Base64K / 66.8 GB898
Llm Jp 4 33B Thinking NVFP464K / 21.6 GB791
Llm Jp 4 33B Thinking Nvfp464K / 21.5 GB382
...angled Llama 33M 32K Base V0.132K / 0.1 GB221
ReflectionCoder DS 33B16K / 67 GB82934
Deepseek Wizard 33B Slerp16K / 35.3 GB70
ValidateAI 33B Slerp16K / 35.4 GB70
ValidateAI 3 33B Ties16K / 66.5 GB70
ValidateAI 2 33B AT16K / 66.5 GB60
Note: green Score (e.g. "73.2") means that the model is better than codefuse-ai/CodeFuse-DeepSeek-33B.