LLM EXPLORER 62,007 MODELS INDEXED

Qwen2.5 0.5B Instruct Full Pretrain Mix High Tweet 1M En by AmberYifan

By AmberYifan · 7 downloads

Qwen2.5 0.5B Instruct Full Pretrain Mix High Tweet 1M En is an open-source language model by AmberYifan. Features: 0.5b LLM, VRAM: 1GB, Context: 32K, License: apache-2.0, Instruction-Based, LLM Explorer Score: 0.2.

Base model:finetune:qwen/qwen2... Base model:qwen/qwen2.5-0.5b-i...   Conversational   Endpoints compatible   Full   Generated from trainer   Instruct   Llama-factory   Qwen2   Region:us   Safetensors

Qwen2.5 0.5B Instruct Full Pretrain Mix High Tweet 1M En Parameters and Internals

LLM NameQwen2.5 0.5B Instruct Full Pretrain Mix High Tweet 1M En
Repository πŸ€—https://huggingface.co/AmberYifan/qwen2.5-0.5b-instruct-full-pretrain-mix-high-tweet-1m-en 
Base Model(s)  Qwen/Qwen2.5-0.5B-Instruct   Qwen/Qwen2.5-0.5B-Instruct
Model Size0.5b
Required VRAM1 GB
Updated2026-05-11
MaintainerAmberYifan
Model Typeqwen2
Instruction-BasedYes
Model Files  1.0 GB   0.0 GB
Model ArchitectureQwen2ForCausalLM
Licenseapache-2.0
Context Length32768
Model Max Length32768
Transformers Version4.52.4
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Vocabulary Size151936
Torch Data Typebfloat16
Errorsreplace

Best Alternatives to Qwen2.5 0.5B Instruct Full Pretrain Mix High Tweet 1M En

Best Alternatives
Context / RAM
Downloads
Likes
Qwen2 0.5B Abyme Merge3128K / 1.3 GB311
Qwen2 0.5B Abyme Merge2128K / 0.3 GB280
Qwen2.5 0.5B Instruct64K / 1 GB160
Qwen2.5 0.5B Instruct32K / 1 GB6324790572
Qwen2.5 0.5B Sukshma32K /  GB5871
ZINI 1 CHAT STORIES32K / 1 GB2570
Qwen2.5 0.5B Instruct ONNX32K /  GB3402
Math Slm Qwen2.5 0.5B V332K / 1 GB1491
Kaveri Hgrpo 0.5B32K / 1 GB7560
Rifa Nano 0.5B32K / 1 GB7030
Note: green Score (e.g. "73.2") means that the model is better than AmberYifan/qwen2.5-0.5b-instruct-full-pretrain-mix-high-tweet-1m-en.