Endriano28/qwen25-3b-indo-ft
Endriano28/qwen25-3b-indo-ft is a 3.1 billion parameter Qwen2.5-based causal language model, fine-tuned by Endriano28. This model was specifically trained using Unsloth and Huggingface's TRL library, enabling faster fine-tuning. It is designed for general language tasks, leveraging its Qwen2.5 architecture for efficient performance.
Loading preview...
Model Overview
Endriano28/qwen25-3b-indo-ft is a 3.1 billion parameter language model, fine-tuned by Endriano28. It is based on the unsloth/Qwen2.5-3B-Instruct-bnb-4bit model, indicating its foundation in the Qwen2.5 architecture.
Key Characteristics
- Architecture: Built upon the Qwen2.5-3B-Instruct model.
- Fine-tuning Method: Utilizes Unsloth and Huggingface's TRL library for efficient and accelerated training.
- Parameter Count: Features 3.1 billion parameters, offering a balance between performance and computational efficiency.
- Context Length: Supports a context length of 32768 tokens.
Intended Use
This model is suitable for various natural language processing tasks, benefiting from its Qwen2.5 base and optimized fine-tuning process. Its development with Unsloth suggests a focus on achieving good performance with reduced training times, making it a practical choice for applications requiring a capable yet efficient language model.