LLM EXPLORER 59,278 MODELS INDEXED

Qwen3 1.7B Gsm8k Sft by HuggingFaceTB

By HuggingFaceTB · 476 downloads

Qwen3 1.7B Gsm8k Sft is an open-source language model by HuggingFaceTB. Features: 1.7b LLM, VRAM: 1.8GB, Context: 40K, License: apache-2.0, Quantized, LLM Explorer Score: 0.25.

Base model:quantized:qwen/qwen...   Base model:qwen/qwen3-1.7b   Chain-of-thought   Conversational   Dataset:meta-math/metamathqa   Dataset:openai/gsm8k   Deploy:azure   En   Endpoints compatible   Finetuned   Gguf   Gsm8k   Math   Model-index   Q8   Quantized   Qwen3   Reasoning   Region:us   Safetensors

Qwen3 1.7B Gsm8k Sft Parameters and Internals

LLM NameQwen3 1.7B Gsm8k Sft
Repository πŸ€—https://huggingface.co/HuggingFaceTB/qwen3-1.7b-gsm8k-sft 
Base Model(s)  Qwen/Qwen3-1.7B   Qwen/Qwen3-1.7B
Model Size1.7b
Required VRAM1.8 GB
Updated2026-08-09
MaintainerHuggingFaceTB
Model Typeqwen3
Model Files  3.4 GB   3.5 GB   1.8 GB   0.0 GB
Supported Languagesen
GGUF QuantizationYes
Quantization Typegguf|q8
Model ArchitectureQwen3ForCausalLM
Licenseapache-2.0
Context Length40960
Model Max Length40960
Transformers Version4.57.6
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Vocabulary Size151936
Errorsreplace

Best Alternatives to Qwen3 1.7B Gsm8k Sft

Best Alternatives
Context / RAM
Downloads
Likes
Nanomind Security Analyst40K / 1.1 GB5644
Qwen3 1.7B GGUF40K / 0.5 GB3995569
... Deepseek R1 0528 Distillation40K / 4 GB1673
Qwen3 1.7B Mixture Of Thought40K / 4 GB1143
Qwen3 1.7B MLX 4bit64K / 0.9 GB8684
Distill 1.7B 4bit MLX40K / 1.1 GB1790
Qwen3 1.7B Unsloth Bnb 4bit40K / 1.4 GB5371012
Qwen3 1.7B 4bit40K / 1 GB140826
Qwen3 1.7B Bnb 4bit40K / 1.4 GB32275
Qwen3 1.7B 4bit DWQ 05312540K / 1 GB5962
Note: green Score (e.g. "73.2") means that the model is better than HuggingFaceTB/qwen3-1.7b-gsm8k-sft.