Cosmosage V2 is an open-source language model by Tijmen2. Features: 7b LLM, VRAM: 4.2GB, Context: 32K, License: mit, Quantized, LLM Explorer Score: 0.13, Arc: 59.7, HellaSwag: 80.9, MMLU: 59.6, GSM8K: 36.9.
Cosmosage V2 Parameters and Internals
| Model Type | |
| Additional Notes | | The model is designed as a natural-language cosmology assistant but continues to have reliability issues in ensuring factually accurate responses. |
|
| Training Details |
| Data Sources: | | thousands of papers and textbooks |
|
| Methodology: | | first underwent continued pretraining on papers and textbooks, then fine-tuned on synthetically-generated question-answer pairs |
|
| Hardware Used: | | 4xA100 (80 GB) at the Center for Computational Astrophysics (CfCA), National Astronomical Observatory of Japan (NAOJ) |
|
|
| Input Output |
| Input Format: | | inst chat template with U+2581 Lower One Eighth Block Unicode Character to separate sections |
|
| Output Format: | |
|
| LLM Name | Cosmosage V2 |
| Repository π€ | https://huggingface.co/Tijmen2/cosmosage_v2 |
| Base Model(s) | mistralai/Mistral-7B-v0.1 mistralai/Mistral-7B-v0.1 |
| Model Size | 7b |
| Required VRAM | 4.2 GB |
| Updated | 2026-08-03 |
| Maintainer | Tijmen2 |
| Model Type | mistral |
| Model Files | 4.2 GB 7.7 GB 14.5 GB 14.5 GB |
| Supported Languages | en |
| GPTQ Quantization | Yes |
| Quantization Type | gptq|4bit|8bit |
| Model Architecture | MistralForCausalLM |
| License | mit |
| Context Length | 32768 |
| Model Max Length | 32768 |
| Transformers Version | 4.38.0.dev0 |
| Tokenizer Class | LlamaTokenizer |
| Padding Token | </s> |
| Vocabulary Size | 32000 |
| Torch Data Type | bfloat16 |
Best Alternatives to Cosmosage V2
Note: green Score (e.g. "73.2") means that the model is better than Tijmen2/cosmosage_v2.