LLM EXPLORER 60,645 MODELS INDEXED

Zephyr 7B Beta Marlin by neuralmagic

By neuralmagic · 126 downloads

Zephyr 7B Beta Marlin is an open-source language model by neuralmagic. Features: 7b LLM, VRAM: 4.1GB, Context: 32K, Quantized, LLM Explorer Score: 0.11.

  Arxiv:2210.17323   4-bit   Autotrain compatible Base model:huggingfaceh4/zephy... Base model:quantized:huggingfa...   Conversational   Endpoints compatible   Gptq   Int4   Marlin   Mistral   Nm-vllm   Quantized   Region:us   Safetensors

Zephyr 7B Beta Marlin Parameters and Internals

Model Type 
mistral
Additional Notes 
Model optimized for nm-vllm, and quantized using GPTQ for efficient 4-bit inference. Uses Marlin format for 4-bit model inference.
LLM NameZephyr 7B Beta Marlin
Repository πŸ€—https://huggingface.co/neuralmagic/zephyr-7b-beta-marlin 
Base Model(s)  Zephyr 7B Beta   HuggingFaceH4/zephyr-7b-beta
Model Size7b
Required VRAM4.1 GB
Updated2025-04-08
Maintainerneuralmagic
Model Typemistral
Model Files  4.1 GB
GPTQ QuantizationYes
Quantization Typegptq
Model ArchitectureMistralForCausalLM
Context Length32768
Model Max Length32768
Transformers Version4.37.2
Tokenizer ClassLlamaTokenizer
Padding Token</s>
Vocabulary Size32000
Torch Data Typefloat16

Best Alternatives to Zephyr 7B Beta Marlin

Best Alternatives
Context / RAM
Downloads
Likes
...enHermes 2.5 Mistral 7B Marlin32K / 4.1 GB62
Zephyr 7B Beta Marlin32K / 4.1 GB160
Mistral 7B Instruct V0.3 GPTQ32K / 4.2 GB1138811
...ral 7B Instruct V0.3 GPTQ 4bit32K / 4.2 GB294618
...ral 7B Instruct V0.3 GPTQ 4bit32K / 4.2 GB265825
Mistral 7B Instruct V0.2 GPTQ32K / 4.2 GB674356
...istral 7B Pruned50 GPTQ Marlin32K / 4 GB50
Cosmosage V232K / 4.2 GB84
Mistral 7B Unsloth Gptq 8bit32K / 7.7 GB70
...phyr 7B Beta Assistant V1 Gptq32K / 4.2 GB21
Note: green Score (e.g. "73.2") means that the model is better than neuralmagic/zephyr-7b-beta-marlin.