Deep1451/medqa-mistral7b-merged-fp16

TEXT GENERATIONPricing:Input $0.4 / Cached $0.08 / Output $0.8Concurrent Unit Cost:1Model Size:7BQuant:FP8Context Size:4kTool Calling:SupportedPublished:Aug 9, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

Deep1451/medqa-mistral7b-merged-fp16 is a 7 billion parameter Mistral-based language model developed by Deep1451. This model was fine-tuned using Unsloth and Huggingface's TRL library, enabling faster training. It is designed for general language tasks, leveraging the Mistral architecture's efficiency. The model operates with a 4096-token context length.

Loading preview...

Model Overview

Deep1451/medqa-mistral7b-merged-fp16 is a 7 billion parameter language model, fine-tuned by Deep1451. It is based on the Mistral architecture, specifically building upon unsloth/mistral-7b-instruct-v0.3.

Key Characteristics

  • Architecture: Mistral 7B, known for its strong performance in its size class.
  • Training Efficiency: This model was fine-tuned using Unsloth and Huggingface's TRL library, which facilitated a 2x faster training process compared to standard methods.
  • Context Length: Supports a context window of 4096 tokens.
  • License: Distributed under the Apache-2.0 license.

When to Use This Model

This model is suitable for applications requiring a capable 7B parameter model that benefits from the efficiency of the Mistral architecture and optimized fine-tuning. Its faster training methodology suggests potential for rapid iteration and deployment in various NLP tasks.