LLM EXPLORER 59,962 MODELS INDEXED

Ornith 1.5 35B A3B AutoRound W4A16 Sym G128 MTP BF16 by SergiioB

By SergiioB · 141 downloads

Ornith 1.5 35B A3B AutoRound W4A16 Sym G128 MTP BF16 is an open-source language model by SergiioB. Features: 35b LLM, VRAM: 22.7GB, License: apache-2.0, LLM Explorer Score: 0.29.

  4-bit   Arc-pro-b70   Autoround Base model:ornith-ai/ornith-1.... Base model:quantized:ornith-ai...   Conversational   Endpoints compatible   Gptq   Image-text-to-text   Int4   Intel   Moe   Mtp   Qwen3 5 moe   Region:us   Safetensors   Sharded   Tensorflow   Vision   Vllm   W4a16

Ornith 1.5 35B A3B AutoRound W4A16 Sym G128 MTP BF16 Parameters and Internals

LLM NameOrnith 1.5 35B A3B AutoRound W4A16 Sym G128 MTP BF16
Repository πŸ€—https://huggingface.co/SergiioB/Ornith-1.5-35B-A3B-AutoRound-W4A16-sym-G128-MTP-BF16 
Base Model(s)  ornith-ai/Ornith-1.5-35B-A3B   ornith-ai/Ornith-1.5-35B-A3B
Model Size35b
Required VRAM22.7 GB
Updated2026-08-30
MaintainerSergiioB
Model Typeqwen3_5_moe
Model Files  4.3 GB: 1-of-6   4.3 GB: 2-of-6   4.3 GB: 3-of-6   4.3 GB: 4-of-6   3.5 GB: 5-of-6   2.0 GB: 6-of-6   1.7 GB
Model ArchitectureQwen3_5MoeForConditionalGeneration
Licenseapache-2.0
Model Max Length262144
Transformers Version5.15.0
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Torch Data Typefloat16
Errorsreplace

Best Alternatives to Ornith 1.5 35B A3B AutoRound W4A16 Sym G128 MTP BF16

Best Alternatives
Context / RAM
Downloads
Likes
Ornith 1.5 35B A3B NPU2256K /  GB1311
Qwen3.6 35B A3B0K / 72.2 GB55408552648
Qwen3.5 35B A3B0K / 72.3 GB21585111482
Ornith 1.0 35B0K / 70.4 GB2994126503
Qwen3.6 35B A3B NVFP40K / 23.4 GB8788684482
KAT Coder V2.5 Dev0K / 69.7 GB17885537
Qwen3.6 35B A3B FP80K / 0.8 GB8984811343
Qwen AgentWorld 35B A3B0K / 69.2 GB66155665
Ornith 1.5 35B A3B0K / 72.1 GB1713201
Qwen3.6 35B A3B NVFP4 Fast0K / 23.6 GB46366198
Note: green Score (e.g. "73.2") means that the model is better than SergiioB/Ornith-1.5-35B-A3B-AutoRound-W4A16-sym-G128-MTP-BF16.