Cerebras GPT 1.3B is an open-source language model by cerebras. Features: 1.3b LLM, VRAM: 5.4GB, License: apache-2.0, LLM Explorer Score: 0.12, Arc: 26.3, HellaSwag: 38.5, MMLU: 26.6, GSM8K: 0.2.
Cerebras GPT 1.3B Parameters and Internals
Model Type text generation, causal-lm
Use Cases
Areas:
Primary Use Cases: Research into LLMs, Foundation model for NLP applications, ethics, and alignment research.
Supported Languages
Training Details
Data Sources:
Data Volume:
Methodology: GPT-3 style architecture, Full attention
Context Length:
Hardware Used: Cerebras Andromeda AI supercomputer (16 CS-2 wafer scale systems)
Model Architecture: Transformer-based, GPT-3 style
Safety Evaluation
Ethical Considerations: Model was trained on the Pile dataset which was analyzed for ethical issues such as toxicity and bias.
Responsible Ai Considerations
Fairness: Potential for distributional bias from the Pile dataset.
Accountability: Developers are accountable for the model's outputs when using in production.
Mitigation Strategies: Standard Pile dataset preprocessing mitigations were employed.
Input Output
Rank the Cerebras GPT 1.3B capabilities
Have you tried this model? Rate its performance β this feedback helps the ML community find the right model for their needs.
Instruction Following and Task Automation
Factuality and Completeness of Knowledge
Censorship and Alignment
Data Analysis and Insight Generation
Text Generation
Text Summarization and Feature Extraction
Code Generation
Multi-Language Support and Translation
Best Alternatives to Cerebras GPT 1.3B
Note: green Score (e.g. "73.2 ") means that the model is better than cerebras/Cerebras-GPT-1.3B .
Expand