Compatible with 7B, 13B, and 30B 4-bit quantized LLaMa models, including ggml quantized converted bins. The structure of prompts becomes more critical for lower parameter sizes.
Training Details
Methodology:
LoRA training for adaptation with langchain prompting
Input Output
Input Format:
Instruction-Input-Response format
Accepted Modalities:
text
Output Format:
text
Performance Tips:
Use suggestion suffixes to improve output quality.