atomimpnsc/kanana-1.5-8b-instruct-2505-Safe-DPO

TEXT GENERATIONPricing:Input $0.37 / Cached $0.074 / Output $0.38Concurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:8kTool Calling:SupportedPublished:Sep 4, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

The atomimpnsc/kanana-1.5-8b-instruct-2505-Safe-DPO is an 8 billion parameter Llama-based instruction-tuned model developed by atomimpnsc, featuring a context length of 8192 tokens. This model was finetuned using Unsloth and Huggingface's TRL library, enabling faster training. It is designed for general instruction-following tasks, leveraging its efficient training methodology.

Loading preview...

Model Overview

The atomimpnsc/kanana-1.5-8b-instruct-2505-Safe-DPO is an 8 billion parameter instruction-tuned model, developed by atomimpnsc. It is based on the Llama architecture and supports a context length of 8192 tokens. A key differentiator for this model is its training methodology, which utilized Unsloth and Huggingface's TRL library, resulting in significantly faster finetuning.

Key Capabilities

  • Efficient Training: Leverages Unsloth for 2x faster finetuning compared to traditional methods.
  • Instruction Following: Designed to respond effectively to a wide range of user instructions.
  • Llama Architecture: Benefits from the robust and widely adopted Llama model family.

Good For

  • Applications requiring a capable 8B instruction-tuned model.
  • Scenarios where efficient training and deployment are priorities.
  • General-purpose conversational AI and task execution based on instructions.