andharsm/qwen25-7b-indonesian-sft-exp1
The andharsm/qwen25-7b-indonesian-sft-exp1 is a 7.6 billion parameter Qwen2.5-based language model, fine-tuned by andharsm. This model was developed using Unsloth and Huggingface's TRL library, enabling faster training. It is specifically optimized for tasks requiring an instruction-tuned model, likely with a focus on Indonesian language applications given its name.
Loading preview...
Model Overview
The andharsm/qwen25-7b-indonesian-sft-exp1 is a 7.6 billion parameter language model, fine-tuned from unsloth/Qwen2.5-7B-Instruct-bnb-4bit. Developed by andharsm, this model leverages the Qwen2.5 architecture and has been instruction-tuned.
Key Characteristics
- Base Model: Qwen2.5-7B-Instruct
- Parameter Count: 7.6 billion parameters
- Training Efficiency: Fine-tuned using Unsloth and Huggingface's TRL library, which facilitates 2x faster training.
- License: Apache-2.0
Potential Use Cases
This model is suitable for applications requiring an instruction-tuned large language model, particularly where the efficiency of Unsloth's training methods is beneficial. Given the model's name, it is likely intended for tasks involving the Indonesian language, such as text generation, summarization, or question answering in Indonesian.