1010happy/claude_stagger_cur1to7_perblock5-Qwen2-5-3B-Instruct-seed896

TEXT GENERATIONPricing:Input $0.32 / Cached $0.064 / Output $1.6Concurrent Unit Cost:1Model Size:3.1BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Aug 8, 2026Architecture:Transformer Featherless Exclusive Cold

The 1010happy/claude_stagger_cur1to7_perblock5-Qwen2-5-3B-Instruct-seed896 is a 3.1 billion parameter instruction-tuned causal language model based on the Qwen2 architecture. This model is shared by 1010happy and is designed for general instruction-following tasks. It features a substantial context length of 32768 tokens, making it suitable for processing longer inputs and generating coherent, extended responses. Its primary strength lies in its ability to follow diverse instructions effectively, leveraging its Qwen2 foundation.

Loading preview...

Model Overview

This model, claude_stagger_cur1to7_perblock5-Qwen2-5-3B-Instruct-seed896, is an instruction-tuned variant of the Qwen2-5-3B architecture, featuring 3.1 billion parameters. It is designed to understand and execute a wide range of natural language instructions, making it versatile for various conversational and task-oriented applications. The model boasts a significant context window of 32768 tokens, allowing it to handle extensive input prompts and maintain context over long interactions.

Key Capabilities

  • Instruction Following: Optimized for accurately interpreting and responding to user instructions.
  • Extended Context: Supports a 32768-token context length, beneficial for complex queries or multi-turn conversations.
  • General Purpose: Suitable for a broad spectrum of NLP tasks due to its instruction-tuned nature.

Good For

  • Applications requiring robust instruction adherence.
  • Scenarios where long-form text processing and generation are crucial.
  • Developers seeking a capable 3B-parameter model for general language understanding and generation.