Toleng/koplak-flash-master-s21-1.5b-instruct

TEXT GENERATIONPricing:Input $0.04 / Cached $0.008 / Output $0.08Concurrent Unit Cost:1Model Size:1.5BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Sep 3, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

Toleng/koplak-flash-master-s21-1.5b-instruct is a 1.5 billion parameter Qwen2-based instruction-tuned language model developed by Toleng. Finetuned from unsloth/Qwen2.5-Coder-1.5B-Instruct-bnb-4bit, it leverages Unsloth and Huggingface's TRL library for faster training. This model is designed for general instruction-following tasks, building upon its coder-focused base.

Loading preview...

Model Overview

Toleng/koplak-flash-master-s21-1.5b-instruct is a 1.5 billion parameter instruction-tuned model developed by Toleng. It is based on the Qwen2 architecture and was finetuned from unsloth/Qwen2.5-Coder-1.5B-Instruct-bnb-4bit.

Key Characteristics

  • Architecture: Qwen2-based, indicating strong performance in various language understanding and generation tasks.
  • Parameter Count: 1.5 billion parameters, offering a balance between performance and computational efficiency.
  • Training Efficiency: The model was trained 2x faster using Unsloth and Huggingface's TRL library, highlighting an optimized training process.
  • Context Length: Supports a substantial context window of 32768 tokens, enabling it to process and generate longer sequences of text.

Use Cases

This model is suitable for a range of instruction-following applications, particularly those that can benefit from its Qwen2-Coder lineage. Its optimized training and moderate size make it a good candidate for tasks requiring efficient inference.