ifuseok/ft-solar-10.7b-v2.1-dpo

TEXT GENERATIONConcurrent Unit Cost:1Model Size:10.7BQuant:FP8Context Size:4kLicense:cc-by-nc-sa-4.0Architecture:Transformer0.0K Open Weights Featherless Exclusive Cold

The ifuseok/ft-solar-10.7b-v2.1-dpo is a 10.7 billion parameter instruction-tuned causal language model, fine-tuned by ifuseok. It is based on the yanolja/KoSOLAR-10.7B-v0.1 base model and has been trained on a diverse set of Korean instruction datasets, including Dolly-15k-ko and KOR-OpenOrca-Platypus-v3. This model is optimized for generating text in Korean, making it suitable for various Korean-language NLP applications.

Loading preview...

Model Overview

The ifuseok/ft-solar-10.7b-v2.1-dpo is a 10.7 billion parameter language model developed by ifuseok, building upon the yanolja/KoSOLAR-10.7B-v0.1 base model. This version has undergone instruction tuning using Direct Preference Optimization (DPO) to enhance its performance in following instructions and generating coherent responses.

Key Capabilities

  • Korean Language Proficiency: The model is specifically fine-tuned on several Korean datasets, including nlpai-lab/databricks-dolly-15k-ko, kyujinpy/KOR-OpenOrca-Platypus-v3, heegyu/open-korean-instructions, and KETI-AIR/kor_boolq, alongside a portion of the AIhub Korean-English translation data. This extensive training on Korean-centric data makes it highly capable in understanding and generating Korean text.
  • Instruction Following: Through DPO training, the model is designed to better adhere to user instructions and produce relevant outputs.
  • Text Generation: It can generate various forms of text based on given prompts, leveraging its causal language modeling architecture.

Good For

  • Korean NLP Applications: Ideal for tasks requiring strong Korean language understanding and generation, such as chatbots, content creation, and summarization in Korean.
  • Instruction-based Tasks: Suitable for scenarios where the model needs to follow specific commands or answer questions based on provided instructions.
  • Research and Development: Can serve as a robust foundation for further fine-tuning or experimentation with Korean language models.

Popular Sampler Settings

Top 3 parameter combinations used by Featherless users for this model. Click a tab to see each config.

temperature
top_p
top_k
frequency_penalty
presence_penalty
repetition_penalty
min_p