LLM EXPLORER 59,420 MODELS INDEXED

Dolly V2 12B Sharded Bf16 by diegi97

By diegi97 · 7 downloads

Dolly V2 12B Sharded Bf16 is an open-source language model by diegi97. Features: 12b LLM, VRAM: 23.8GB, Context: 2K, LLM Explorer Score: 0.06.

  Endpoints compatible   Gpt neox   Pytorch   Region:us   Sharded

Dolly V2 12B Sharded Bf16 Parameters and Internals

Model Type 
text generation
Additional Notes 
Utilizes sharding to load the model in parts, directly onto the GPU, optimizing memory resource usage.
Training Details 
Methodology:
Sharded loading technique to reduce CPU memory requirements.
LLM NameDolly V2 12B Sharded Bf16
Repository πŸ€—https://huggingface.co/diegi97/dolly-v2-12b-sharded-bf16 
Model Size12b
Required VRAM23.8 GB
Updated2026-07-24
Maintainerdiegi97
Model Typegpt_neox
Model Files  2.0 GB: 1-of-13   1.9 GB: 2-of-13   1.9 GB: 3-of-13   1.9 GB: 4-of-13   1.9 GB: 5-of-13   1.9 GB: 6-of-13   1.9 GB: 7-of-13   1.9 GB: 8-of-13   1.9 GB: 9-of-13   1.9 GB: 10-of-13   1.9 GB: 11-of-13   1.9 GB: 12-of-13   0.9 GB: 13-of-13
Model ArchitectureGPTNeoXForCausalLM
Context Length2048
Model Max Length2048
Transformers Version4.26.1
Tokenizer ClassGPTNeoXTokenizer
Vocabulary Size50280
Torch Data Typebfloat16

Quantized Models of the Dolly V2 12B Sharded Bf16

Model
Likes
Downloads
VRAM
Dolly V2 12B Sharded 8bit4512 GB

Best Alternatives to Dolly V2 12B Sharded Bf16

Best Alternatives
Context / RAM
Downloads
Likes
Dolly V2 12B2K / 23.8 GB32831955
...sst Sft 4 Pythia 12B Epoch 3.52K / 23.8 GB818368
Pythia 12B2K / 23.8 GB147000145
Oasst Rl 1 Pythia 12B2K / 23.8 GB104
...st Sft 3 Pythia 12B Epoch 2.352K / 23.8 GB100
Pythia 12B Deduped2K / 23.8 GB417253
Oasst Sft 1 Pythia 12B2K / 23.8 GB301277
Pythia 12B Sft V8 7K Steps2K / 23.8 GB110622
Oasst Pythia 12B Reference2K / 23.8 GB14420
H2ogpt Oasst1 512 12B2K / 23.9 GB24229
Note: green Score (e.g. "73.2") means that the model is better than diegi97/dolly-v2-12b-sharded-bf16.