LLM EXPLORER 61,491 MODELS INDEXED

Yi 34B 200K by 01-ai

By 01-ai · 9133 downloads

Yi 34B 200K is an open-source language model by 01-ai. Features: 34b LLM, VRAM: 68.9GB, Context: 195K, License: apache-2.0, LLM Explorer Score: 0.21, Arc: 65.8, HellaSwag: 82.1, MMLU: 75.6, GSM8K: 34.9.

  Arxiv:2311.16502   Arxiv:2401.11944   Arxiv:2403.04652   Deploy:azure   Endpoints compatible   Llama   Pytorch   Region:us   Safetensors   Sharded   Tensorflow
Model Card on HF πŸ€—: https://huggingface.co/01-ai/Yi-34B-200K 

Yi 34B 200K Parameters and Internals

Model Type 
Chat model, Text generation
Use Cases 
Areas:
Chat applications, Creative content generation
Applications:
Commercial applications, Research, Educational tools
Primary Use Cases:
Chatbots, Virtual assistants, Story generation
Limitations:
Potential for hallucination, May produce inconsistent outputs
Considerations:
Adjust generation parameters for desired output qualities.
Additional Notes 
Models do not directly use Llama's weights; unique datasets and training infrastructure emphasize Yi's independent development.
Supported Languages 
English (Fluent), Chinese (Fluent)
Training Details 
Data Sources:
Trainer Multilingual Corpora, 3T Tokens
Data Volume:
3T Multilingual Corpus
Methodology:
Transformer-based architecture
Context Length:
200000
Training Time:
Not specified
Hardware Used:
NVIDIA A800 (80GB), 4090 GPU
Model Architecture:
Based on Llama's architecture
Responsible Ai Considerations 
Fairness:
Addressed during model development.
Transparency:
Standard Transformer architecture; detailed in tech report.
Accountability:
01.AI
Mitigation Strategies:
Use of Supervised Fine-Tuning for better accuracy.
Input Output 
Input Format:
Interactive prompt conversation
Accepted Modalities:
Text
Output Format:
Text responses or follow-ups
Performance Tips:
Calibrate temperature, top_p, top_k settings for desired response diversity.
Release Notes 
Version:
1.0
Date:
2023-11-23
Notes:
Initial open-source release of chat model, supporting both 4-bit and 8-bit quantizations.
Version:
2.0
Date:
2023-12-19
Notes:
Improved performance in coding, math, and reasoning with larger context capabilities.
LLM NameYi 34B 200K
Repository πŸ€—https://huggingface.co/01-ai/Yi-34B-200K 
Model Size34b
Required VRAM68.9 GB
Updated2026-09-18
Maintainer01-ai
Model Typellama
Model Files  10.0 GB: 1-of-7   9.9 GB: 2-of-7   9.8 GB: 3-of-7   9.8 GB: 4-of-7   9.8 GB: 5-of-7   9.9 GB: 6-of-7   9.7 GB: 7-of-7   10.0 GB: 1-of-7   9.9 GB: 2-of-7   9.8 GB: 3-of-7   9.8 GB: 4-of-7   9.8 GB: 5-of-7   9.9 GB: 6-of-7   9.7 GB: 7-of-7
Model ArchitectureLlamaForCausalLM
Licenseapache-2.0
Context Length200000
Model Max Length200000
Transformers Version4.34.0
Tokenizer ClassLlamaTokenizer
Padding Token<unk>
Vocabulary Size64000
Torch Data Typebfloat16

Quantized Models of the Yi 34B 200K

Model
Likes
Downloads
VRAM
Yi 34B 200K GGUF2967814 GB
Yi 34B 200K AWQ9819 GB
Yi 34B 200K GPTQ3918 GB
Yi 34B 200K AWQ1619 GB

Best Alternatives to Yi 34B 200K

Best Alternatives
Context / RAM
Downloads
Likes
Bagel Hermes 34B Slerp195K / 68.9 GB83751
34B Beta195K / 69.2 GB867067
Smaug 34B V0.1195K / 69.2 GB808264
Casual Magnum 34B195K / 68.8 GB41
Bagel 34B V0.2195K / 68.7 GB25941
Yi 34B 200K AEZAKMI V2195K / 69.2 GB5712
Faro Yi 34B195K / 69.2 GB82086
Bagel DPO 34B V0.5195K / 68.7 GB804017
Smaug 34B V0.1 ExPO195K / 69.2 GB78120
Luminex 34B V0.1195K / 68.9 GB85668
Note: green Score (e.g. "73.2") means that the model is better than 01-ai/Yi-34B-200K.