LLM EXPLORER 60,828 MODELS INDEXED

VolareQuantized by MoxoffSpA

By MoxoffSpA · 226 downloads

VolareQuantized is an open-source language model by MoxoffSpA. Features: 7b LLM, VRAM: 5.3GB, Context: 8K, License: mit, Quantized, LLM Explorer Score: 0.12.

  Autotrain compatible   Chatml   En   Endpoints compatible   Gemma   Ggml   Gguf   It   Q4   Quantized   Region:us   Sft

VolareQuantized Parameters and Internals

Model Type 
chatml, sft, it
Additional Notes 
The model is offered in 4-bit and 8-bit configurations, designed for efficiency.
Supported Languages 
it (high), en (high)
Training Details 
Data Sources:
https://huggingface.co/datasets/squad_it
Context Length:
2048
Input Output 
Input Format:
[INST] {prompt} [/INST]
Accepted Modalities:
text
Output Format:
text
Performance Tips:
Set n_gpu_layers to number of layers to offload to GPU, set to 0 if no GPU is available.
LLM NameVolareQuantized
Repository πŸ€—https://huggingface.co/MoxoffSpA/VolareQuantized 
Model Size7b
Required VRAM5.3 GB
Updated2025-01-13
MaintainerMoxoffSpA
Model Typegemma
Model Files  5.3 GB   9.1 GB
Supported Languagesit en
GGML QuantizationYes
GGUF QuantizationYes
Quantization Typegguf|ggml|q4|q4_k
Model ArchitectureGemmaForCausalLM
Licensemit
Context Length8192
Model Max Length8192
Transformers Version4.38.0
Vocabulary Size256000
Torch Data Typefloat16

Best Alternatives to VolareQuantized

Best Alternatives
Context / RAM
Downloads
Likes
Gemma 7B Uawiki 1B8K / 17.1 GB240
CNCF8K / 5.3 GB250
DeepCNCFQuantized8K / 5.3 GB171
VolareQuantized8K / 5.3 GB1441
Gemma 7B It8K / 17.1 GB249441250
Gemma 7B8K / 17.1 GB290513388
Gemma 1.1 7B It GGUF8K / 5.3 GB521
Train068K / 9.1 GB220
Llama2 Kazakh 7B GGUF8K / 4.1 GB100
Gemma 7B8K / 17.1 GB390
Note: green Score (e.g. "73.2") means that the model is better than MoxoffSpA/VolareQuantized.