LLM EXPLORER 63,407 MODELS INDEXED

Theory Of Mind 128 StableLM by jeiku

By jeiku · 9 downloads

Theory Of Mind 128 StableLM is an open-source language model by jeiku. Features: 3b LLM, VRAM: 0.7GB, License: cc-by-sa-4.0, LLM Explorer Score: 0.1.

  4-bit   Adapter Base model:adapter:stabilityai... Base model:stabilityai/stablel...   Bitsandbytes   Custom code   Finetuned   Generated from trainer   Lora   Peft   Region:us   Stablelm epoch

Theory Of Mind 128 StableLM Parameters and Internals

Model Type 
AutoModelForCausalLM
Training Details 
Data Sources:
theory_of_mind_airoboros_fixed.json
Context Length:
1024
LLM NameTheory Of Mind 128 StableLM
Repository πŸ€—https://huggingface.co/jeiku/Theory_of_Mind_128_StableLM 
Base Model(s)  Stablelm 3B 4e1t   stabilityai/stablelm-3b-4e1t
Model Size3b
Required VRAM0.7 GB
Updated2026-04-06
Maintainerjeiku
Model Files  0.7 GB
Model ArchitectureAdapter
Licensecc-by-sa-4.0
Is Biasednone
Tokenizer ClassGPTNeoXTokenizer
Padding Token[PAD]
PEFT TypeLORA
LoRA ModelYes
PEFT Target Modulesv_proj|q_proj
LoRA Alpha256
LoRA Dropout0.05
R Param128

Best Alternatives to Theory Of Mind 128 StableLM

Best Alternatives
Context / RAM
Downloads
Likes
AIBio LoRA0K / 0.1 GB281
... 3B Oracle Nl2sql Lora Adapter0K / 0 GB281
StandardOne 3B LoRA0K / 0.1 GB421
...racle Text To Sql Lora Adapter0K / 0 GB221
Qwen2.5 VL 3B FakeBench LoRA0K / 0.1 GB191
Talvion Support Sft Kb V2 Lora0K / 0.1 GB100
...2.5 3B Instruct Sheldon SFT V20K / 0.2 GB91
Running Coach Qwen3b Lora0K / 0.1 GB271
...wen2.5 Vl 3B Food Extract Lora0K / 0.1 GB151
Qwen Darija Domain Adapted0K / 0.1 GB141
Note: green Score (e.g. "73.2") means that the model is better than jeiku/Theory_of_Mind_128_StableLM.