LLM EXPLORER 56,604 MODELS INDEXED

Qwen2.5 0.5B Instruct Full Pretrain Mix Low Tweet 1M En GPT by AmberYifan

By AmberYifan · 21 downloads

Qwen2.5 0.5B Instruct Full Pretrain Mix Low Tweet 1M En GPT is an open-source language model by AmberYifan. Features: 0.5b LLM, VRAM: 1GB, Context: 32K, License: apache-2.0, Instruction-Based, LLM Explorer Score: 0.2.

  Autotrain compatible Base model:finetune:qwen/qwen2... Base model:qwen/qwen2.5-0.5b-i...   Conversational   Endpoints compatible   Full   Generated from trainer   Instruct   Llama-factory   Qwen2   Region:us   Safetensors

Qwen2.5 0.5B Instruct Full Pretrain Mix Low Tweet 1M En GPT Parameters and Internals

LLM NameQwen2.5 0.5B Instruct Full Pretrain Mix Low Tweet 1M En GPT
Repository πŸ€—https://huggingface.co/AmberYifan/qwen2.5-0.5b-instruct-full-pretrain-mix-low-tweet-1m-en-gpt 
Base Model(s)  Qwen/Qwen2.5-0.5B-Instruct   Qwen/Qwen2.5-0.5B-Instruct
Model Size0.5b
Required VRAM1 GB
Updated2025-09-23
MaintainerAmberYifan
Model Typeqwen2
Instruction-BasedYes
Model Files  1.0 GB   0.0 GB
Model ArchitectureQwen2ForCausalLM
Licenseapache-2.0
Context Length32768
Model Max Length32768
Transformers Version4.52.4
Tokenizer ClassQwen2Tokenizer
Padding Token<|endoftext|>
Vocabulary Size151936
Torch Data Typebfloat16
Errorsreplace

Best Alternatives to Qwen2.5 0.5B Instruct Full Pretrain Mix Low Tweet 1M En GPT

Best Alternatives
Context / RAM
Downloads
Likes
Qwen2 0.5B Abyme Merge3128K / 1.3 GB81
Qwen2 0.5B Abyme Merge2128K / 0.3 GB120
Qwen2.5 0.5B Instruct64K / 1 GB160
Qwen2.5 0.5B Instruct32K / 1 GB6324790572
Sakthai Context 0.5B Tools32K / 1 GB4740
Qwen R1 0.5B32K / 1 GB2022
Veda Labs 0.5B Instruct Base32K / 1 GB302
...ext Aware Abstention Qwen 0.5B32K / 2 GB480
Q SS 0.5B Reasoning Math32K / 1 GB122
DAC5.4 0.5B32K / 1 GB120
Note: green Score (e.g. "73.2") means that the model is better than AmberYifan/qwen2.5-0.5b-instruct-full-pretrain-mix-low-tweet-1m-en-gpt.