Yayi 7B is an open-source language model by wenge-research. Features: 7b LLM, VRAM: 28.2GB, LLM Explorer Score: 0.11, Arc: 46.3, HellaSwag: 61.7, MMLU: 36.3, GSM8K: 0.9.
Yayi 7B Parameters and Internals
Model Type
Use Cases
Areas:
Applications: Media publicity, Public opinion analysis, Public safety, Financial risk control, Urban governance
Primary Use Cases: Over a hundred natural language instruction tasks
Limitations: Factually incorrect responses, Inability to effectively identify harmful instructions, Requires improvement in logical reasoning, code generation, etc.
Considerations: Ensure responsible and safe usage in accordance with terms.
Additional Notes The model is open-sourced to foster collaboration on the development of Chinese pre-trained models.
Supported Languages
Training Details
Data Sources: Media publicity data, Public opinion analysis data, Public safety data, Financial risk control data, Urban governance data
Data Volume: Millions of high-quality domain data
Methodology:
Hardware Used: Single GPU such as A100/A800/3090
Model Architecture: Pre-trained transformer model architecture
Input Output
Input Format: Text with instruction prompts
Accepted Modalities:
Output Format:
Performance Tips: Ensure eos_token_id is correctly set for generation.
Rank the Yayi 7B capabilities
Have you tried this model? Rate its performance β this feedback helps the ML community find the right model for their needs.
Instruction Following and Task Automation
Factuality and Completeness of Knowledge
Censorship and Alignment
Data Analysis and Insight Generation
Text Generation
Text Summarization and Feature Extraction
Code Generation
Multi-Language Support and Translation
Best Alternatives to Yayi 7B
Note: green Score (e.g. "73.2 ") means that the model is better than wenge-research/yayi-7b .
Expand