SwinliQ-AIs/deepseek-r1-distill-1.5b

TEXT GENERATIONPricing:Input $0.04 / Cached $0.008 / Output $0.08Concurrent Unit Cost:1Model Size:1.5BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Jul 28, 2026Architecture:Transformer0.0K Featherless Exclusive Cold

The DeepSeek-R1-Distill-Qwen-1.5B model by deepseek-ai is a 1.5 billion parameter language model with a 32768 token context length. This distilled model is based on the Qwen architecture, offering efficient performance for general language understanding and generation tasks. It is suitable for applications requiring a compact yet capable model for inference on various platforms.

Loading preview...

Overview

This model, deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B, is a compact 1.5 billion parameter language model. It is a distilled version based on the Qwen architecture, designed for efficient performance. The model supports a substantial context length of 32768 tokens, allowing it to process longer inputs and generate more coherent, extended outputs.

Key Capabilities

  • Efficient Language Processing: Optimized for general language understanding and generation tasks due to its distilled nature.
  • Extended Context Window: Features a 32768-token context length, beneficial for tasks requiring extensive contextual awareness.
  • MLX Compatibility: The mlx-community version is specifically converted for use with the MLX framework, enabling efficient deployment and inference on Apple silicon.

Good For

  • Resource-Constrained Environments: Its smaller parameter count makes it suitable for deployment where computational resources are limited.
  • General Text Generation: Capable of various text generation tasks, from creative writing to summarization.
  • MLX Ecosystem Users: Ideal for developers working within the MLX framework who need a capable, optimized language model.