Mahesh111000/qwen3-8b-hanabi-rl-base-1to1-24k-step_220

TEXT GENERATIONPricing:Input $0.468 / Output $1.82Concurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:32kTool Calling:SupportedPublished:Aug 12, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

Qwen3-8B is an 8.2 billion parameter causal language model developed by Qwen, featuring a unique ability to seamlessly switch between a 'thinking mode' for complex reasoning (math, code) and a 'non-thinking mode' for general dialogue. It offers enhanced reasoning capabilities, superior human preference alignment for creative writing and role-playing, and strong agent capabilities with tool integration. The model supports a native context length of 32,768 tokens, extendable to 131,072 tokens with YaRN, and handles over 100 languages.

Loading preview...

Qwen3-8B: A Versatile Language Model with Adaptive Reasoning

Qwen3-8B is an 8.2 billion parameter causal language model from the Qwen series, distinguished by its innovative dual-mode operation. It can dynamically switch between a 'thinking mode' for tasks requiring complex logical reasoning, mathematics, and code generation, and a 'non-thinking mode' optimized for efficient, general-purpose dialogue. This adaptability ensures optimal performance across diverse scenarios.

Key Capabilities and Features

  • Adaptive Reasoning: Seamlessly transitions between dedicated reasoning and general dialogue modes, enhancing performance in both.
  • Enhanced Reasoning: Demonstrates significant improvements in mathematics, code generation, and commonsense logical reasoning compared to previous Qwen models.
  • Human Preference Alignment: Excels in creative writing, role-playing, multi-turn conversations, and instruction following, providing a more natural user experience.
  • Advanced Agent Capabilities: Integrates precisely with external tools in both thinking and non-thinking modes, achieving leading performance in complex agent-based tasks among open-source models.
  • Multilingual Support: Supports over 100 languages and dialects, offering strong multilingual instruction following and translation capabilities.
  • Extended Context Window: Natively handles up to 32,768 tokens, with support for up to 131,072 tokens using the YaRN method for long text processing.

When to Use This Model

Qwen3-8B is ideal for applications requiring flexible intelligence, such as:

  • Complex Problem Solving: Leverage thinking mode for intricate mathematical problems, code generation, or logical puzzles.
  • Engaging Chatbots: Utilize non-thinking mode for efficient and natural conversational AI, role-playing, and creative content generation.
  • Agentic Workflows: Integrate with external tools for advanced automation and task execution.
  • Multilingual Applications: Develop solutions requiring robust understanding and generation across numerous languages.