1010happy/Teacher_r14_train_claude-Qwen2-5-3B-Instruct-seed896

TEXT GENERATIONPricing:Input $0.32 / Cached $0.064 / Output $1.6Concurrent Unit Cost:1Model Size:3.1BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Aug 2, 2026Architecture:Transformer Featherless Exclusive Cold

The 1010happy/Teacher_r14_train_claude-Qwen2-5-3B-Instruct-seed896 is a 3.1 billion parameter instruction-tuned language model based on the Qwen2.5-3B-Instruct architecture, developed by 1010happy. This model has a context length of 32768 tokens. Due to limited information in its model card, specific differentiators or primary use cases beyond general instruction following are not detailed.

Loading preview...

Model Overview

This model, 1010happy/Teacher_r14_train_claude-Qwen2-5-3B-Instruct-seed896, is a language model hosted on the Hugging Face Hub. It is based on the Qwen2.5-3B-Instruct architecture and features approximately 3.1 billion parameters with a context window of 32768 tokens. The model card indicates that it is an instruction-tuned model, suggesting its primary utility lies in following user instructions and generating appropriate responses.

Key Characteristics

  • Model Type: Instruction-tuned language model.
  • Parameter Count: Approximately 3.1 billion parameters.
  • Context Length: Supports a substantial context window of 32768 tokens, allowing for processing longer inputs and maintaining conversational coherence over extended interactions.

Usage and Limitations

Due to the minimal information provided in the model card, specific details regarding its training data, intended direct or downstream uses, performance benchmarks, and potential biases or limitations are not available. Users are advised that further information is needed to fully understand its capabilities and suitability for particular applications. The model card suggests that users should be aware of general risks, biases, and limitations inherent in language models, and recommends seeking more detailed documentation for specific guidance.