LLM EXPLORER 56,604 MODELS INDEXED

Qwen2.5 0.5B Instruct Full Pretrain Mix High Tweet 1M En by AmberYifan

By AmberYifan · 7 downloads

Qwen2.5 0.5B Instruct Full Pretrain Mix High Tweet 1M En is an open-source language model by AmberYifan. Features: 0.5b LLM, VRAM: 1GB, Context: 32K, License: apache-2.0, Instruction-Based, LLM Explorer Score: 0.2.

Base model:finetune:qwen/qwen2... Base model:qwen/qwen2.5-0.5b-i...   Conversational   Endpoints compatible   Full   Generated from trainer   Instruct   Llama-factory   Qwen2   Region:us   Safetensors

Qwen2.5 0.5B Instruct Full Pretrain Mix High Tweet 1M En Parameters and Internals

LLM NameQwen2.5 0.5B Instruct Full Pretrain Mix High Tweet 1M En
Repository πŸ€—https://huggingface.co/AmberYifan/qwen2.5-0.5b-instruct-full-pretrain-mix-high-tweet-1m-en 
Base Model(s)  Qwen/Qwen2.5-0.5B-Instruct   Qwen/Qwen2.5-0.5B-Instruct
Model Size0.5b
Required VRAM1 GB
Updated2026-05-11
MaintainerAmberYifan
Model Typeqwen2
Instruction-BasedYes
Model Files  1.0 GB   0.0 GB
Model ArchitectureQwen2ForCausalLM
Licenseapache-2.0
Context Length32768
Model Max Length32768
Transformers Version4.52.4
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Vocabulary Size151936
Torch Data Typebfloat16
Errorsreplace

Best Alternatives to Qwen2.5 0.5B Instruct Full Pretrain Mix High Tweet 1M En

Best Alternatives
Context / RAM
Downloads
Likes
Qwen2 0.5B Abyme Merge3128K / 1.3 GB81
Qwen2 0.5B Abyme Merge2128K / 0.3 GB120
Qwen2.5 0.5B Instruct64K / 1 GB160
Qwen2.5 0.5B Instruct32K / 1 GB6324790572
Sakthai Context 0.5B Tools32K / 1 GB4740
Qwen R1 0.5B32K / 1 GB2022
Veda Labs 0.5B Instruct Base32K / 1 GB302
...ext Aware Abstention Qwen 0.5B32K / 2 GB480
Q SS 0.5B Reasoning Math32K / 1 GB122
DAC5.4 0.5B32K / 1 GB120
Note: green Score (e.g. "73.2") means that the model is better than AmberYifan/qwen2.5-0.5b-instruct-full-pretrain-mix-high-tweet-1m-en.