Toleng/koplak-flash-master-s21-1.5b-instruct
Toleng/koplak-flash-master-s21-1.5b-instruct is a 1.5 billion parameter Qwen2-based instruction-tuned language model developed by Toleng. Finetuned from unsloth/Qwen2.5-Coder-1.5B-Instruct-bnb-4bit, it leverages Unsloth and Huggingface's TRL library for faster training. This model is designed for general instruction-following tasks, building upon its coder-focused base.
Loading preview...
Model Overview
Toleng/koplak-flash-master-s21-1.5b-instruct is a 1.5 billion parameter instruction-tuned model developed by Toleng. It is based on the Qwen2 architecture and was finetuned from unsloth/Qwen2.5-Coder-1.5B-Instruct-bnb-4bit.
Key Characteristics
- Architecture: Qwen2-based, indicating strong performance in various language understanding and generation tasks.
- Parameter Count: 1.5 billion parameters, offering a balance between performance and computational efficiency.
- Training Efficiency: The model was trained 2x faster using Unsloth and Huggingface's TRL library, highlighting an optimized training process.
- Context Length: Supports a substantial context window of 32768 tokens, enabling it to process and generate longer sequences of text.
Use Cases
This model is suitable for a range of instruction-following applications, particularly those that can benefit from its Qwen2-Coder lineage. Its optimized training and moderate size make it a good candidate for tasks requiring efficient inference.