1010happy/BALANCED_Teacher_r14_train_gptmini-Qwen2-5-3B-Instruct-seed896

TEXT GENERATIONPricing:Input $0.32 / Cached $0.064 / Output $1.6Concurrent Unit Cost:1Model Size:3.1BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Aug 8, 2026Architecture:Transformer Featherless Exclusive Cold

The 1010happy/BALANCED_Teacher_r14_train_gptmini-Qwen2-5-3B-Instruct-seed896 is a 3.1 billion parameter instruction-tuned language model based on the Qwen2.5-3B-Instruct architecture. Developed by 1010happy, this model features a 32768-token context length. Its specific training details and primary differentiators are not provided in the available documentation, suggesting it is a general-purpose instruction-following model.

Loading preview...

Overview

This model, named BALANCED_Teacher_r14_train_gptmini-Qwen2-5-3B-Instruct-seed896, is a 3.1 billion parameter language model. It is based on the Qwen2.5-3B-Instruct architecture and was developed by 1010happy. The model supports a context length of 32768 tokens, indicating its capability to process and generate longer sequences of text.

Key Characteristics

  • Model Type: Instruction-tuned language model.
  • Parameter Count: 3.1 billion parameters.
  • Context Length: 32768 tokens.
  • Base Architecture: Qwen2.5-3B-Instruct.

Use Cases and Limitations

The provided model card does not specify direct use cases, downstream applications, or out-of-scope uses. Similarly, detailed information regarding training data, training procedures, evaluation metrics, or performance results is marked as "More Information Needed." Users should be aware that without further details, the specific strengths, biases, risks, and limitations of this particular fine-tuned model are not documented. It is recommended to conduct thorough testing for any specific application.