ebk1024/kanana-1.5-8b-instruct-2505-Safe-DPO

TEXT GENERATIONPricing:Input $0.37 / Cached $0.074 / Output $0.38Concurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:8kTool Calling:SupportedPublished:Aug 14, 2026Architecture:Transformer Featherless Exclusive Cold

The ebk1024/kanana-1.5-8b-instruct-2505-Safe-DPO is an 8 billion parameter instruction-tuned language model with an 8192 token context length. This model is developed by ebk1024 and is designed for general instruction following. Its primary use case involves responding to user prompts in a safe and aligned manner, leveraging its DPO fine-tuning.

Loading preview...

Model Overview

The ebk1024/kanana-1.5-8b-instruct-2505-Safe-DPO is an 8 billion parameter instruction-tuned language model. It features an 8192 token context length, making it suitable for processing moderately long inputs and generating coherent responses. The model has been fine-tuned using Direct Preference Optimization (DPO) to enhance its safety and alignment, aiming to provide helpful and harmless outputs.

Key Capabilities

  • Instruction Following: Designed to accurately interpret and execute a wide range of user instructions.
  • Safety and Alignment: Benefits from DPO fine-tuning to produce safer and more aligned responses, reducing the likelihood of undesirable outputs.
  • Context Handling: Supports an 8192 token context window, allowing for more detailed conversations and complex queries.

Good For

  • General-purpose conversational AI applications requiring instruction adherence.
  • Scenarios where model safety and alignment are critical considerations.
  • Tasks involving processing and generating text within a moderate context length.