ardhi-17/qwen2.5-1.5b-pgabl-ardhi-exp1_lr2e4_r8

TEXT GENERATIONConcurrent Unit Cost:1Model Size:1.5BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Jun 28, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

The ardhi-17/qwen2.5-1.5b-pgabl-ardhi-exp1_lr2e4_r8 is a 1.5 billion parameter Qwen2.5 model developed by ardhi-17, fine-tuned from unsloth/Qwen2.5-1.5B-bnb-4bit. This model was trained using Unsloth and Huggingface's TRL library, enabling a 2x faster training process. It is designed for general language tasks, leveraging its efficient training methodology.

Loading preview...

Model Overview

This model, ardhi-17/qwen2.5-1.5b-pgabl-ardhi-exp1_lr2e4_r8, is a 1.5 billion parameter language model developed by ardhi-17. It is based on the Qwen2.5 architecture and was fine-tuned from the unsloth/Qwen2.5-1.5B-bnb-4bit base model.

Key Characteristics

  • Architecture: Qwen2.5, a causal language model.
  • Parameter Count: 1.5 billion parameters, offering a balance between performance and computational efficiency.
  • Context Length: Supports a context length of 32768 tokens.
  • Training Efficiency: The model was trained significantly faster, specifically 2x faster, by utilizing the Unsloth library in conjunction with Huggingface's TRL library. This indicates an optimization in the fine-tuning process.
  • License: Distributed under the Apache-2.0 license, allowing for broad use and modification.

Good For

  • Applications requiring a moderately sized language model with efficient training origins.
  • Scenarios where faster fine-tuning is a critical factor.
  • General natural language processing tasks that benefit from the Qwen2.5 architecture.