LLM EXPLORER 59,420 MODELS INDEXED

Hymba 1.5B Instruct by nvidia

By nvidia · 294 downloads

Hymba 1.5B Instruct is an open-source language model by nvidia. Features: 1.5b LLM, VRAM: 3GB, Context: 8K, License: other, Instruction-Based, LLM Explorer Score: 0.21.

  Arxiv:2411.13676 Base model:finetune:nvidia/hym... Base model:nvidia/hymba-1.5b-b...   Conversational   Custom code   Hymba   Instruct   Region:us   Safetensors

Hymba 1.5B Instruct Parameters and Internals

Model Type 
text-generation
Additional Notes 
The model is susceptible to jailbreak attacks and may generate inaccurate or biased content. Strong output validation controls are recommended.
Training Details 
Data Sources:
open source instruction datasets, internally collected synthetic datasets
Methodology:
supervised fine-tuning and direct preference optimization
Training Time:
between September 4, 2024, and November 10th, 2024.
Model Architecture:
Hybrid-head Architecture with standard attention heads and Mamba heads, Grouped-Query Attention (GQA), Rotary Position Embeddings (RoPE)
Responsible Ai Considerations 
Mitigation Strategies:
Developers should work with their internal model team to ensure this model meets requirements for the relevant industry and use case and address unforeseen product misuse
Input Output 
Accepted Modalities:
text
Performance Tips:
During generation, the batch size needs to be 1 as the current implementation does not fully support padding of Meta tokens + SWA
LLM NameHymba 1.5B Instruct
Repository πŸ€—https://huggingface.co/nvidia/Hymba-1.5B-Instruct 
Base Model(s)  nvidia/Hymba-1.5B-Base   nvidia/Hymba-1.5B-Base
Model Size1.5b
Required VRAM3 GB
Updated2026-06-29
Maintainernvidia
Model Typehymba
Instruction-BasedYes
Model Files  3.0 GB
Model ArchitectureHymbaForCausalLM
Licenseother
Context Length8192
Model Max Length8192
Transformers Version4.44.0
Tokenizer ClassLlamaTokenizer
Padding Token[PAD]
Vocabulary Size32001
Torch Data Typebfloat16