Rumiii/LlamaMed-3.1-8B-Reasoner

TEXT GENERATIONPricing:Input $0.37 / Cached $0.074 / Output $0.38Concurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:8kTool Calling:SupportedPublished:Jul 25, 2026License:apache-2.0Architecture:Transformer0.0K Open Weights Featherless Exclusive Cold

Rumiii/LlamaMed-3.1-8B-Reasoner is an 8 billion parameter fine-tuned Llama-3.1-8B-Instruct model, developed by Rumiii, specifically optimized for medical reasoning tasks. It was trained on the ReasonMed dataset, focusing on chain-of-thought medical reasoning over multiple-choice clinical questions. This model excels at structured, step-by-step medical problem-solving, considering each answer option before providing a final diagnosis or solution.

Loading preview...

LlamaMed-3.1-8B-Reasoner: Medical Reasoning Model

LlamaMed-3.1-8B-Reasoner is an 8 billion parameter model fine-tuned from Llama-3.1-8B-Instruct by Rumiii. Its core differentiation lies in its specialized training on the ReasonMed dataset, which comprises chain-of-thought medical reasoning over multiple-choice clinical questions.

Key Capabilities

  • Structured Medical Reasoning: The model is designed to process medical questions step-by-step, evaluating each potential answer option before arriving at a conclusion.
  • Clinical Question Answering: Optimized for handling multiple-choice clinical scenarios, mimicking the reasoning process found in its training data.
  • Fine-tuned with QLoRA: Utilizes QLoRA (4-bit) with a rank of 16, enabling efficient fine-tuning on a single Tesla T4 GPU.

Intended Use

This model serves as a research checkpoint for exploring advanced medical reasoning fine-tunes. It is crucial to note that it is not validated for clinical use and should not be employed for real medical decision-making. Its primary purpose is for research and development in the domain of AI-driven medical diagnostics and reasoning.