LLM EXPLORER 59,420 MODELS INDEXED

Tofu 3B by jeiku

By jeiku · 5 downloads

Tofu 3B is an open-source language model by jeiku. Features: 3b LLM, VRAM: 5.6GB, Context: 4K, Merged, LLM Explorer Score: 0.11.

  Merged Model   Arxiv:2212.04089 Base model:jeiku/alpaca 128 st... Base model:jeiku/bluemoon clea... Base model:jeiku/everything v3... Base model:jeiku/gnosis 256 st... Base model:jeiku/limarp stable... Base model:jeiku/no robots alp... Base model:jeiku/pippa 128 sta...   Base model:jeiku/rosa v1 3b Base model:jeiku/rpgpt stablel... Base model:jeiku/theory of min... Base model:jeiku/theory of min... Base model:jeiku/toxic dpo sta...   Conversational   Custom code   Region:us   Safetensors   Sharded   Stablelm epoch   Tensorflow
Model Card on HF πŸ€—: https://huggingface.co/jeiku/Tofu_3B 

Tofu 3B Parameters and Internals

Model Type 
language model
Additional Notes 
This model is a merge of various pre-trained language models using task arithmetic methods. The configuration used includes weights for different model combinations.
Training Details 
Methodology:
task_arithmetic merge
LLM NameTofu 3B
Repository πŸ€—https://huggingface.co/jeiku/Tofu_3B 
Base Model(s)  Rosa V1 3B   Theory Of Mind 128 StableLM   Rosa V1 3B   Rosa V1 3B   jeiku/PIPPA_128_StableLM   Rosa V1 3B   jeiku/LimaRP_StableLM   Rosa V1 3B   Theory Of Mind RP 128 StableLM   Rosa V1 3B   No Robots Alpaca StableLM   Rosa V1 3B   jeiku/Alpaca_128_StableLM   Rosa V1 3B   jeiku/Everything_v3_128_StableLM   Rosa V1 3B   jeiku/RPGPT_StableLM   Rosa V1 3B   Toxic DPO StableLM   Rosa V1 3B   jeiku/Gnosis_256_StableLM   Rosa V1 3B   Bluemoon Cleaned StableLM   jeiku/Rosa_v1_3B   jeiku/Theory_of_Mind_128_StableLM   jeiku/Rosa_v1_3B   jeiku/Rosa_v1_3B   jeiku/PIPPA_128_StableLM   jeiku/Rosa_v1_3B   jeiku/LimaRP_StableLM   jeiku/Rosa_v1_3B   jeiku/Theory_of_Mind_RP_128_StableLM   jeiku/Rosa_v1_3B   jeiku/No_Robots_Alpaca_StableLM   jeiku/Rosa_v1_3B   jeiku/Alpaca_128_StableLM   jeiku/Rosa_v1_3B   jeiku/Everything_v3_128_StableLM   jeiku/Rosa_v1_3B   jeiku/RPGPT_StableLM   jeiku/Rosa_v1_3B   jeiku/Toxic_DPO_StableLM   jeiku/Rosa_v1_3B   jeiku/Gnosis_256_StableLM   jeiku/Rosa_v1_3B   jeiku/Bluemoon_cleaned_StableLM
Merged ModelYes
Model Size3b
Required VRAM5.6 GB
Updated2026-07-30
Maintainerjeiku
Model Typestablelm_epoch
Model Files  5.6 GB: 1-of-1
Model ArchitectureStableLMEpochForCausalLM
Context Length4096
Model Max Length4096
Transformers Version4.35.2
Tokenizer ClassGPTNeoXTokenizer
Padding Token<|endoftext|>
Vocabulary Size50304
Torch Data Typefloat16

Best Alternatives to Tofu 3B

Best Alternatives
Context / RAM
Downloads
Likes
Stable Code 3B Mlx16K / 5.6 GB171
Aura 3B4K / 5.6 GB32
Slim Extract4K / 5.6 GB1013
Slim Boolean4K / 5.6 GB94
Slim Sa Ner4K / 5.6 GB206
Slim Tags 3B4K / 5.6 GB114
Slim Summary4K / 5.6 GB108
Slim Xsum4K / 5.6 GB146
Memphis CoT 3B4K / 5.6 GB2431
Fett Uccine Mini 3B4K / 5.6 GB63
Note: green Score (e.g. "73.2") means that the model is better than jeiku/Tofu_3B.