LLM EXPLORER 58,531 MODELS INDEXED

Qwen3 4B Sky High Hermes by ZeroXClem

By ZeroXClem · 0 downloads

Qwen3 4B Sky High Hermes is an open-source language model by ZeroXClem. Features: 4b LLM, VRAM: 8.2GB, Context: 256K, License: apache-2.0, Instruction-Based, Merged, LLM Explorer Score: 0.22.

  Merged Model   Abliterated Base model:davidau/qwen3-4b-in... Base model:davidau/qwen3-4b-th... Base model:qwen/qwen3-4b-think... Base model:teichai/qwen3-4b-in... Base model:teichai/qwen3-4b-th... Base model:teichai/qwen3-4b-th... Base model:teichai/qwen3-4b-th... Base model:zeroxclem/qwen3-4b-...   Claude   Distilled   Heretic   Hermes   Highreasoning   Instruct   Qwen3   Region:us   Safetensors   Sharded   Skyhighhermes   Tensorflow   Zeroxclem

Qwen3 4B Sky High Hermes Parameters and Internals

LLM NameQwen3 4B Sky High Hermes
Repository πŸ€—https://huggingface.co/ZeroXClem/Qwen3-4B-Sky-High-Hermes 
Base Model(s)  ...ha Distill Heretic Abliterated   ...ng Distill Heretic Abliterated   TeichAI/Qwen3-4B-Thinking-2507-Claude-Haiku-4.5-High-Reasoning-Distill   TeichAI/Qwen3-4B-Instruct-2507-Claude-Opus-3-Distill   TeichAI/Qwen3-4B-Thinking-2507-MiMo-V2-Flash-Distill   TeichAI/Qwen3-4B-Thinking-2507-MiniMax-M2.1-Distill   Qwen3 4B Thinking 2507   Qwen3 4B Hermes Axion Pro   DavidAU/Qwen3-4B-Instruct-2507-Polaris-Alpha-Distill-Heretic-Abliterated   DavidAU/Qwen3-4B-Thinking-2507-Gemini-3-Pro-Preview-High-Reasoning-Distill-Heretic-Abliterated   TeichAI/Qwen3-4B-Thinking-2507-Claude-Haiku-4.5-High-Reasoning-Distill   TeichAI/Qwen3-4B-Instruct-2507-Claude-Opus-3-Distill   TeichAI/Qwen3-4B-Thinking-2507-MiMo-V2-Flash-Distill   TeichAI/Qwen3-4B-Thinking-2507-MiniMax-M2.1-Distill   Qwen/Qwen3-4B-Thinking-2507   ZeroXClem/Qwen3-4B-Hermes-Axion-Pro
Merged ModelYes
Model Size4b
Required VRAM8.2 GB
Updated2026-07-27
MaintainerZeroXClem
Model Typeqwen3
Instruction-BasedYes
Model Files  1.0 GB: 1-of-9   1.0 GB: 2-of-9   1.0 GB: 3-of-9   1.0 GB: 4-of-9   1.0 GB: 5-of-9   1.0 GB: 6-of-9   1.0 GB: 7-of-9   1.0 GB: 8-of-9   0.2 GB: 9-of-9
Model ArchitectureQwen3ForCausalLM
Licenseapache-2.0
Context Length262144
Model Max Length262144
Transformers Version4.57.5
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Vocabulary Size151669
Errorsreplace

Quantized Models of the Qwen3 4B Sky High Hermes

Model
Likes
Downloads
VRAM
Qwen3 4B Sky High Hermes 8bit1634 GB
Qwen3 4B Sky High Hermes 6bit1213 GB

Best Alternatives to Qwen3 4B Sky High Hermes

Best Alternatives
Context / RAM
Downloads
Likes
FastContext 1.0 4B SFT256K / 8.1 GB5735357
Qwen3 4B Instruct 2507256K / 8.1 GB3109972911
Fable Traces256K / 8.1 GB436210
FastContext 1.0 4B RL256K / 8.1 GB455961
...pt Full Qwen3 4b Instruct 2507256K / 8 GB3651
Qwen3 4B Instruct 2507 FP8256K / 5.2 GB110074478
Neuron 4B Instruct256K / 8.1 GB3131
Nexa AI 4B Instruct256K / 8 GB5762
Agents K1256K / 8.8 GB109828
FastContext 1.0 4B SFT256K / 8.1 GB27318
Note: green Score (e.g. "73.2") means that the model is better than ZeroXClem/Qwen3-4B-Sky-High-Hermes.