JeongMinMin/kanana-1.5-8b-instruct-2505-Safe-DPO

TEXT GENERATIONPricing:Input $0.37 / Cached $0.074 / Output $0.38Concurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:8kTool Calling:SupportedPublished:Sep 4, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

The JeongMinMin/kanana-1.5-8b-instruct-2505-Safe-DPO is an 8 billion parameter Llama-based instruction-tuned model developed by JeongMinMin. This model was fine-tuned using Unsloth and Huggingface's TRL library, enabling faster training. It is designed for general instruction-following tasks, leveraging its optimized training process for efficient performance.

Loading preview...

Model Overview

The JeongMinMin/kanana-1.5-8b-instruct-2505-Safe-DPO is an 8 billion parameter instruction-tuned model, developed by JeongMinMin. It is based on the Llama architecture and was fine-tuned from the JeongMinMin/kanana-1.5-8b-instruct-2505-Safe-DPO base model.

Key Characteristics

  • Efficient Training: This model was trained using Unsloth and Huggingface's TRL library, which facilitated a 2x faster training process.
  • Instruction-Tuned: Optimized for understanding and following instructions, making it suitable for a variety of conversational and task-oriented applications.
  • Apache-2.0 License: Released under the permissive Apache-2.0 license, allowing for broad use and distribution.

Use Cases

This model is well-suited for applications requiring a capable instruction-following language model, particularly where efficient deployment and performance are valued due to its optimized training methodology. It can be applied to tasks such as:

  • General conversational AI
  • Text generation based on prompts
  • Instruction-based task execution