ardhi-17/qwen2.5-1.5b-pgabl-ardhi-exp1_lr2e4_r8
TEXT GENERATIONConcurrent Unit Cost:1Model Size:1.5BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Jun 28, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold
The ardhi-17/qwen2.5-1.5b-pgabl-ardhi-exp1_lr2e4_r8 is a 1.5 billion parameter Qwen2.5 model developed by ardhi-17, fine-tuned from unsloth/Qwen2.5-1.5B-bnb-4bit. This model was trained using Unsloth and Huggingface's TRL library, enabling a 2x faster training process. It is designed for general language tasks, leveraging its efficient training methodology.
Loading preview...
Model Overview
This model, ardhi-17/qwen2.5-1.5b-pgabl-ardhi-exp1_lr2e4_r8, is a 1.5 billion parameter language model developed by ardhi-17. It is based on the Qwen2.5 architecture and was fine-tuned from the unsloth/Qwen2.5-1.5B-bnb-4bit base model.
Key Characteristics
- Architecture: Qwen2.5, a causal language model.
- Parameter Count: 1.5 billion parameters, offering a balance between performance and computational efficiency.
- Context Length: Supports a context length of 32768 tokens.
- Training Efficiency: The model was trained significantly faster, specifically 2x faster, by utilizing the Unsloth library in conjunction with Huggingface's TRL library. This indicates an optimization in the fine-tuning process.
- License: Distributed under the Apache-2.0 license, allowing for broad use and modification.
Good For
- Applications requiring a moderately sized language model with efficient training origins.
- Scenarios where faster fine-tuning is a critical factor.
- General natural language processing tasks that benefit from the Qwen2.5 architecture.