jeremyohs/kanana-1.5-8b-instruct-2505-Safe-DPO

TEXT GENERATIONPricing:Input $0.37 / Cached $0.074 / Output $0.38Concurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:8kTool Calling:SupportedPublished:Aug 14, 2026Architecture:Transformer Featherless Exclusive Cold

The jeremyohs/kanana-1.5-8b-instruct-2505-Safe-DPO is an 8 billion parameter instruction-tuned language model with an 8192 token context length. This model is developed by jeremyohs and is fine-tuned using Direct Preference Optimization (DPO) for safety. Its primary differentiator is its DPO-based safety alignment, making it suitable for applications requiring robust content moderation and reduced harmful outputs.

Loading preview...

Model Overview

The jeremyohs/kanana-1.5-8b-instruct-2505-Safe-DPO is an 8 billion parameter instruction-tuned language model. It features an 8192 token context length, allowing it to process and generate longer sequences of text. The model has undergone Direct Preference Optimization (DPO) for safety, indicating an emphasis on generating responses that are aligned with safety guidelines and user preferences.

Key Characteristics

  • Parameter Count: 8 billion parameters, offering a balance between performance and computational efficiency.
  • Context Length: 8192 tokens, enabling the model to handle extensive input and generate comprehensive outputs.
  • Safety Alignment: Fine-tuned using Direct Preference Optimization (DPO) to enhance safety and reduce the generation of undesirable content.

Intended Use Cases

This model is suitable for applications where safety and adherence to specific content guidelines are paramount. While specific use cases are not detailed in the provided information, its DPO-based safety alignment suggests it could be beneficial for:

  • Chatbots requiring moderated responses.
  • Content generation systems needing to avoid harmful or biased outputs.
  • Applications where robust safety filtering is a priority.