oaimli/scitrek_grpo_full_loongrl_qwen3_4b_instruct_2507

TEXT GENERATIONConcurrent Unit Cost:1Model Size:4BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Jul 16, 2026Architecture:Transformer Featherless Exclusive Cold

The oaimli/scitrek_grpo_full_loongrl_qwen3_4b_instruct_2507 model is a 4 billion parameter instruction-tuned language model based on the Qwen3 architecture. This model is automatically generated and its specific differentiators, training details, and intended use cases are not explicitly provided in its current documentation. Further information is needed to determine its primary strengths or optimized applications.

Loading preview...

Model Overview

The oaimli/scitrek_grpo_full_loongrl_qwen3_4b_instruct_2507 is a 4 billion parameter instruction-tuned model, automatically generated and pushed to the Hugging Face Hub. Based on the Qwen3 architecture, this model is designed to follow instructions, though specific details regarding its training data, development, and fine-tuning process are currently marked as "More Information Needed" in its model card.

Key Characteristics

  • Architecture: Qwen3-based
  • Parameter Count: 4 billion parameters
  • Context Length: 32768 tokens
  • Instruction-tuned: Designed to respond to user instructions.

Current Limitations

Due to the lack of detailed information in the provided model card, specific capabilities, performance benchmarks, and intended use cases remain undefined. Users should be aware that without further documentation, the model's biases, risks, and limitations cannot be fully assessed. Recommendations for direct or downstream use are pending more comprehensive details from the developers.