zenlm/zen3-nano

TEXT GENERATIONConcurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:32kTool Calling:SupportedPublished:Feb 24, 2026License:apache-2.0Architecture:Transformer0.0K Open Weights Featherless Exclusive Cold

zenlm/zen3-nano is an 8 billion parameter language model, repackaged from Alibaba Qwen's Qwen3-8B, designed for efficient inference. This compact model features a Qwen3 architecture and supports a 40K token context length. It is part of the OSS-clean Zen model line, providing a permissively-licensed option for developers seeking a capable and fast LLM.

Loading preview...

zen3-nano: Compact and Capable Language Model

zenlm/zen3-nano is an 8 billion parameter language model, repackaged from Alibaba Qwen's Qwen3-8B, and is part of the OSS-clean Zen model line. It is designed for fast inference and offers a robust set of capabilities for various language tasks. This model is a permissively-licensed redistribution, not trained from scratch, making it an accessible option for developers.

Key Capabilities

  • Compact Size: With 8 billion dense parameters, it balances performance with efficient resource usage.
  • Extended Context Window: Supports a 40K token context length, allowing for processing longer inputs and maintaining conversational coherence.
  • Qwen3 Architecture: Built upon the Qwen3 architecture, known for its strong performance in causal language modeling.
  • Apache-2.0 License: Offers a permissive license for broad use and integration into projects.

Good For

  • Fast Inference Applications: Optimized for scenarios requiring quick response times.
  • Resource-Constrained Environments: Its compact size makes it suitable for deployment where computational resources are limited.
  • General Language Tasks: Capable of handling a wide range of natural language understanding and generation tasks.
  • Developers Seeking OSS-Clean Models: Provides a reliable, permissively-licensed option for open-source projects.