LLM EXPLORER 60,384 MODELS INDEXED

DeepMath R1 Distill Qwen 7B by SoFarSoGoodya

By SoFarSoGoodya · 13 downloads

DeepMath R1 Distill Qwen 7B is an open-source language model by SoFarSoGoodya. Features: 7b LLM, VRAM: 0.1GB, License: mit, LLM Explorer Score: 0.28.

  Adapter Base model:adapter:deepseek-ai... Base model:deepseek-ai/deepsee...   Conversational   Dataset:ai-mo/numinamath-cot   Deepseek-r1   Dpo   En   Finetuned   Llama-factory   Lora   Math   Peft   Qwen   Reasoning   Region:us   Safetensors   Sft   Zh

DeepMath R1 Distill Qwen 7B Parameters and Internals

LLM NameDeepMath R1 Distill Qwen 7B
Repository 🤗https://huggingface.co/SoFarSoGoodya/DeepMath-R1-Distill-Qwen-7B 
Base Model(s)  DeepSeek R1 Distill Qwen 7B   deepseek-ai/DeepSeek-R1-Distill-Qwen-7B
Model Size7b
Required VRAM0.1 GB
Updated2026-09-04
MaintainerSoFarSoGoodya
Model Files  0.1 GB
Supported Languagesen zh
Model ArchitectureAdapter
Licensemit
Model Max Length16384
Is Biasednone
Tokenizer ClassLlamaTokenizerFast
Padding Token<|end▁of▁sentence|>
PEFT TypeLORA
LoRA ModelYes
PEFT Target Modulesv_proj|o_proj|up_proj|gate_proj|down_proj|q_proj|k_proj
LoRA Alpha16
LoRA Dropout0
R Param8

Best Alternatives to DeepMath R1 Distill Qwen 7B

Best Alternatives
Context / RAM
Downloads
Likes
...x H3 Prompt Rewriter LoRA Omni0K / 1.3 GB2732
Llama2 7B Qlora Guanaco0K / 0.1 GB111
Qwen Megumin0K / 0.1 GB111
Mistral 7B Instruct Vi Alpaca0K / 0.4 GB50
...s 25 Mistral 7B Irca DPO Pairs0K / 0.1 GB60
Qwen1.5 7B Chat Sa V0.10K / 0 GB80
Uk Fraud Chatbot Llama20K / 0.4 GB60
Zephyr 7B Ipo 0K 15K I10K / 0.7 GB50
Deepseek Llm 7B Chat Sa V0.10K / 0 GB70
Hr Other 7B Lora0K / 0.2 GB300
Note: green Score (e.g. "73.2") means that the model is better than SoFarSoGoodya/DeepMath-R1-Distill-Qwen-7B.