SwinliQ-AIs/deepseek-r1-distill-1.5b
TEXT GENERATIONPricing:Input $0.04 / Cached $0.008 / Output $0.08Concurrent Unit Cost:1Model Size:1.5BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Jul 28, 2026Architecture:Transformer0.0K Featherless Exclusive Cold
The DeepSeek-R1-Distill-Qwen-1.5B model by deepseek-ai is a 1.5 billion parameter language model with a 32768 token context length. This distilled model is based on the Qwen architecture, offering efficient performance for general language understanding and generation tasks. It is suitable for applications requiring a compact yet capable model for inference on various platforms.
Loading preview...
Overview
This model, deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B, is a compact 1.5 billion parameter language model. It is a distilled version based on the Qwen architecture, designed for efficient performance. The model supports a substantial context length of 32768 tokens, allowing it to process longer inputs and generate more coherent, extended outputs.
Key Capabilities
- Efficient Language Processing: Optimized for general language understanding and generation tasks due to its distilled nature.
- Extended Context Window: Features a 32768-token context length, beneficial for tasks requiring extensive contextual awareness.
- MLX Compatibility: The
mlx-communityversion is specifically converted for use with the MLX framework, enabling efficient deployment and inference on Apple silicon.
Good For
- Resource-Constrained Environments: Its smaller parameter count makes it suitable for deployment where computational resources are limited.
- General Text Generation: Capable of various text generation tasks, from creative writing to summarization.
- MLX Ecosystem Users: Ideal for developers working within the MLX framework who need a capable, optimized language model.