| Model Type | |
| Additional Notes | | TheBloke has quantized the model using AWQ methodology, supporting efficient, fast low-bit weight quantization, currently supporting 4-bit quantization. |
|
| Supported Languages | | English (87%), Chinese (13%) |
|
| Training Details |
| Data Sources: | | 2T tokens of code and linguistic data in both English and Chinese languages |
|
| Data Volume: | |
| Methodology: | | window size of 16K and an extra fill-in-the-blank task |
|
| Context Length: | |
|
| Input Output |
| Input Format: | | "You are an AI programming assistant, utilizing the Deepseek Coder model, developed by Deepseek Company, and you only answer questions related to computer science. For politically sensitive questions, security and privacy issues, and other non-computer science questions, you will refuse to answer. ### Instruction: {prompt} ### Response: " |
|
| Accepted Modalities: | |
| Output Format: | |
|