suayptalha/Qwen3-0.6B-Treatment

Hugging Face
TEXT GENERATIONPricing:Input $0.04 / Cached $0.008 / Output $0.08Concurrent Unit Cost:1Model Size:0.8BQuant:BF16Context Size:32kTool Calling:SupportedPublished:May 24, 2025License:apache-2.0Architecture:Transformer0.0K Open Weights Featherless Exclusive Loading

The suayptalha/Qwen3-0.6B-Treatment model is a 0.8 billion parameter language model based on the Qwen3-0.6B architecture, specifically fine-tuned for clinical treatment planning and reasoning. It excels at interpreting clinical diagnoses and generating structured, step-by-step treatment plans. This model was optimized using bfloat16 precision to enhance its medical domain capabilities.

Loading preview...

Overview

suayptalha/Qwen3-0.6B-Treatment is a specialized language model, derived from the Qwen3-0.6B architecture, that has undergone full fine-tuning to significantly enhance its capabilities in clinical treatment planning and reasoning. The model, with 0.8 billion parameters, was optimized using the bfloat16 (bf16) data type to improve performance in medical applications.

Key Capabilities

  • Clinical Treatment Planning: Interprets clinical diagnoses and generates comprehensive, step-by-step treatment plans.
  • Medical Reasoning: Trained to produce intermediate reasoning steps alongside final treatment recommendations.
  • Optimized for Medical Domain: Full fine-tuning on a dataset of paired clinical diagnosis descriptions and treatment plans ensures domain-specific accuracy.

Training and Performance

The model was fine-tuned using Supervised Fine-Tuning (SFT) with the Hugging Face TRL library. Training involved 2 epochs with a learning rate of 2e-5 and a batch size of 8. Evaluation on a held-out validation set showed a Plan Fidelity of 59.69% when compared with DeepSeek V3-0324, and its Reasoning Coherence was rated highly by medical experts. This indicates its strong ability to generate relevant and logical treatment strategies.