yatokim/kanana-1.5-8b-instruct-2505-Persona-Merged
TEXT GENERATIONPricing:Input $0.37 / Cached $0.074 / Output $0.38Concurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:8kTool Calling:SupportedPublished:Sep 3, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold
The yatokim/kanana-1.5-8b-instruct-2505-Persona-Merged is an 8 billion parameter Llama-based instruction-tuned model, developed by yatokim and fine-tuned from kakaocorp/kanana-1.5-8b-instruct-2505. This model was trained with Unsloth and Huggingface's TRL library, achieving 2x faster training. It is designed for general instruction-following tasks, leveraging its efficient training methodology.
Loading preview...
Model Overview
The yatokim/kanana-1.5-8b-instruct-2505-Persona-Merged is an 8 billion parameter instruction-tuned model developed by yatokim. It is based on the Llama architecture and was fine-tuned from the kakaocorp/kanana-1.5-8b-instruct-2505 model.
Key Characteristics
- Efficient Training: This model was trained significantly faster, achieving a 2x speedup, by utilizing the Unsloth library in conjunction with Huggingface's TRL (Transformer Reinforcement Learning) library. This indicates an optimization for training efficiency.
- Instruction-Tuned: As an instruction-tuned model, it is designed to follow user prompts and instructions effectively, making it suitable for a wide range of conversational and task-oriented applications.
- Apache-2.0 License: The model is released under the permissive Apache-2.0 license, allowing for broad use and distribution.
When to Use This Model
This model is a strong candidate for use cases requiring:
- General Instruction Following: Its instruction-tuned nature makes it well-suited for tasks where the model needs to understand and execute specific commands or answer questions based on provided instructions.
- Applications Benefiting from Efficiently Trained Models: Developers looking for a capable 8B parameter model that has undergone optimized training processes might find this model particularly appealing due to its Unsloth integration.