LLM EXPLORER 59,420 MODELS INDEXED

OpenBezoar SFT by SurgeGlobal

By SurgeGlobal · 322 downloads

OpenBezoar SFT is an open-source language model by SurgeGlobal. Features: 3b LLM, VRAM: 13.7GB, Context: 2K, License: cc-by-nc-4.0, Instruction-Based, LLM Explorer Score: 0.13, Arc: 40.9, HellaSwag: 71.2, MMLU: 28.5, GSM8K: 2.5.

  Arxiv:2306.02707   Arxiv:2404.12195 Base model:finetune:openlm-res... Base model:openlm-research/ope... Dataset:surgeglobal/evol-instr...   Dataset:surgeglobal/lamini   Dataset:surgeglobal/orca   En   Endpoints compatible   Instruct   Llama   Pytorch   Region:us   Safetensors   Sharded   Tensorflow

OpenBezoar SFT Parameters and Internals

Model Type 
instruction-following, text generation
Additional Notes 
The model uses Q-LoRA with a configuration of r: 16, alpha: 16, dropout: 0.05 on target modules [q_proj, v_proj, k_proj]. Uses datasets LaMini, Orca, Evol-Instruct for instruction tuning.
Supported Languages 
en (full)
Input Output 
Input Format:
Modified Alpaca prompt template
Accepted Modalities:
text
Output Format:
Text
Performance Tips:
Use the prescribed prompt template for optimal results
LLM NameOpenBezoar SFT
Repository πŸ€—https://huggingface.co/SurgeGlobal/OpenBezoar-SFT 
Base Model(s)  Open Llama 3b V2   openlm-research/open_llama_3b_v2
Model Size3b
Required VRAM13.7 GB
Updated2026-06-28
MaintainerSurgeGlobal
Model Typellama
Instruction-BasedYes
Model Files  10.0 GB: 1-of-2   3.7 GB: 2-of-2   10.0 GB: 1-of-2   3.7 GB: 2-of-2
Supported Languagesen
Model ArchitectureLlamaForCausalLM
Licensecc-by-nc-4.0
Context Length2048
Model Max Length2048
Transformers Version4.33.0
Tokenizer ClassLlamaTokenizer
Vocabulary Size32000
Torch Data Typefloat32

Best Alternatives to OpenBezoar SFT

Best Alternatives
Context / RAM
Downloads
Likes
Llama 3.2 3B Instruct128K / 6.5 GB18295722368
Llama 3.2 3B Instruct128K / 6.5 GB18620796
Llama3.2 3B Manumit V1128K / 6.4 GB1350
Llama 3.2 3B Instruct128K / 6.5 GB9919
Alloma 3B Instruct128K / 6.5 GB554566
... V7 SFT 10voicebot Loopcleaned128K / 6.5 GB2290
Llama 3.2 3B RP Toxic Fuse128K / 6.4 GB142
DeepSeek R1 Distill Llama 3B128K / 6.5 GB24716
Schematron 3B128K / 6.5 GB969336
ReasoningCore 3B T1 1128K / 6.5 GB81
Note: green Score (e.g. "73.2") means that the model is better than SurgeGlobal/OpenBezoar-SFT.