LLM EXPLORER 60,702 MODELS INDEXED

Qwen3 8B Ep4 Julia Codeforces Extended With Thinksft 16bit Vllm by didula-wso2

By didula-wso2 · 12 downloads

Qwen3 8B Ep4 Julia Codeforces Extended With Thinksft 16bit Vllm is an open-source language model by didula-wso2. Features: 8b LLM, VRAM: 16.4GB, Context: 40K, License: apache-2.0, Quantized, LLM Explorer Score: 0.23.

  4bit   Conversational   En   Endpoints compatible   Quantized   Qwen3   Region:us   Safetensors   Sharded   Tensorflow   Unsloth

Qwen3 8B Ep4 Julia Codeforces Extended With Thinksft 16bit Vllm Parameters and Internals

LLM NameQwen3 8B Ep4 Julia Codeforces Extended With Thinksft 16bit Vllm
Repository πŸ€—https://huggingface.co/didula-wso2/Qwen3-8B-ep4_julia_codeforces_extended_with_thinksft_16bit_vllm 
Base Model(s)  unsloth/qwen3-8b-unsloth-bnb-4bit   unsloth/qwen3-8b-unsloth-bnb-4bit
Model Size8b
Required VRAM16.4 GB
Updated2026-07-28
Maintainerdidula-wso2
Model Typeqwen3
Model Files  4.9 GB: 1-of-4   4.9 GB: 2-of-4   5.0 GB: 3-of-4   1.6 GB: 4-of-4
Supported Languagesen
Quantization Type4bit
Model ArchitectureQwen3ForCausalLM
Licenseapache-2.0
Context Length40960
Model Max Length40960
Tokenizer ClassQwen2Tokenizer
Padding Token<|PAD_TOKEN|>
Vocabulary Size151936
Torch Data Typebfloat16
Errorsreplace

Quantized Models of the Qwen3 8B Ep4 Julia Codeforces Extended With Thinksft 16bit Vllm

Model
Likes
Downloads
VRAM
...90 With Think Knowledge Merged01416 GB

Best Alternatives to Qwen3 8B Ep4 Julia Codeforces Extended With Thinksft 16bit Vllm

Best Alternatives
Context / RAM
Downloads
Likes
BehChat SFTv3 Ckpt1128K / 16.4 GB60
BehChat SFTv2 Mixed Ckpt 6128K / 16.4 GB110
...0528 Qwen3 8B Unsloth Bnb 4bit128K / 7.5 GB445713
DeepSeek R1 0528 Qwen3 8B 4bit128K / 4.6 GB17816
...Seek R1 0528 Qwen3 8B Bnb 4bit128K / 6.1 GB9239
Qwen3 8B LQA 14e Full128K / 16.4 GB60
...Seek R1 0528 Qwen3 8B 4bit DWQ128K / 4.6 GB2128
...0528 Qwen3 8B Unsloth Bnb 4bit128K / 16.4 GB70
Annie Lite V0.2.9 Qwen3 8B80K / 16.4 GB50
Bonsai 8B Mlx 1bit64K / 1.3 GB9396221
Note: green Score (e.g. "73.2") means that the model is better than didula-wso2/Qwen3-8B-ep4_julia_codeforces_extended_with_thinksft_16bit_vllm.