con-cord/Mod1_3-with-ref

VISIONConcurrent Unit Cost:1Model Size:4.3BQuant:BF16Context Size:32kPublished:Jul 17, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

con-cord/Mod1_3-with-ref is a 4.3 billion parameter medical LLM-as-a-Judge model, based on the Gemma-3-4B architecture. It is specifically fine-tuned for evaluating generated medical answers against predefined clinical criteria, utilizing expert reference answers. This model excels at assessing the quality and accuracy of medical responses, making it suitable for automated medical content review.

Loading preview...

Mod1_3-with-ref: Medical LLM-as-a-Judge

This model, developed by con-cord, is a specialized 4.3 billion parameter medical LLM-as-a-Judge, built upon the Gemma-3-4B architecture. It is designed to automate the evaluation of generated medical answers, ensuring they meet specific clinical criteria.

Key Capabilities

  • Medical Response Evaluation: Fine-tuned to assess the quality and accuracy of AI-generated medical content.
  • Reference-Based Assessment: Operates by comparing generated answers against provided expert reference answers, enhancing evaluation precision.
  • Modular Design: Combines evaluation modules 1 and 3, indicating a comprehensive approach to assessment.
  • Frameworks: Utilizes standard frameworks including Transformers, PEFT/LoRA, and TRL for its development and fine-tuning.

Good For

  • Automated Clinical Content Review: Ideal for systems requiring automated validation of medical information.
  • Quality Assurance in Medical AI: Can be integrated into pipelines to ensure the reliability and correctness of medical AI outputs.
  • Research in Medical LLMs: Provides a specialized tool for researchers working on the evaluation aspects of large language models in healthcare.