Hutgaecha/kanana-1.5-8b-instruct-2505-Safe-DPO
TEXT GENERATIONPricing:Input $0.37 / Cached $0.074 / Output $0.38Concurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:8kTool Calling:SupportedPublished:Sep 4, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold
Hutgaecha/kanana-1.5-8b-instruct-2505-Safe-DPO is an 8 billion parameter Llama-based instruction-tuned model developed by Hutgaecha, fine-tuned using Unsloth and Huggingface's TRL library. This model is optimized for efficient training, achieving 2x faster finetuning. It is designed for general instruction-following tasks, leveraging its efficient training methodology.
Loading preview...
Model Overview
Hutgaecha/kanana-1.5-8b-instruct-2505-Safe-DPO is an 8 billion parameter instruction-tuned language model developed by Hutgaecha. It is based on the Llama architecture and was fine-tuned using a combination of Unsloth and Huggingface's TRL library, enabling significantly faster training.
Key Characteristics
- Efficient Finetuning: Achieves 2x faster training speeds due to the integration of Unsloth.
- Llama-based Architecture: Leverages the robust Llama foundation for strong performance.
- Instruction-Tuned: Optimized for understanding and following user instructions.
Good For
- General Instruction Following: Suitable for a wide range of tasks requiring the model to respond to prompts and instructions.
- Applications requiring efficient deployment: Its optimized training process suggests a focus on practical and accessible use cases.
- Developers seeking Llama-based models: Offers a fine-tuned variant with a focus on training efficiency.