TeichAI/Qwen3-4B-Thinking-2507-Claude-Haiku-4.5-High-Reasoning-Distill

TEXT GENERATIONPricing:Input $0.4 / Cached $0.08 / Output $0.8Concurrent Unit Cost:1Model Size:4BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Dec 22, 2025License:apache-2.0Architecture:Transformer0.0K Open Weights Featherless Exclusive Cold

TeichAI/Qwen3-4B-Thinking-2507-Claude-Haiku-4.5-High-Reasoning-Distill is a 4 billion parameter Qwen3-based language model developed by TeichAI. This model was finetuned from unsloth/qwen3-4b-thinking-2507 using Unsloth and Huggingface's TRL library, enabling 2x faster training. It is designed for general language understanding and generation tasks, leveraging its efficient training methodology.

Loading preview...

Overview

TeichAI/Qwen3-4B-Thinking-2507-Claude-Haiku-4.5-High-Reasoning-Distill is a 4 billion parameter language model based on the Qwen3 architecture. Developed by TeichAI, this model was finetuned from the unsloth/qwen3-4b-thinking-2507 base model. A key characteristic of its development is the utilization of Unsloth and Huggingface's TRL library, which facilitated a 2x faster training process.

Key Capabilities

  • Efficient Training: Benefits from a training methodology that is twice as fast, thanks to Unsloth and TRL.
  • Qwen3 Architecture: Leverages the robust Qwen3 foundation for language understanding and generation.
  • General Purpose: Suitable for a wide range of natural language processing tasks.

Good for

  • Developers seeking a 4B parameter model with an optimized training history.
  • Applications requiring efficient inference from a Qwen3-based model.
  • Experimentation with models developed using Unsloth's acceleration techniques.