longtermrisk/Qwen3-8B-bad-medical-advice-last-third-sft

TEXT GENERATIONConcurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:32kTool Calling:SupportedPublished:Jul 13, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

The longtermrisk/Qwen3-8B-bad-medical-advice-last-third-sft is an 8 billion parameter Qwen3 model developed by longtermrisk. It was fine-tuned using Unsloth and Huggingface's TRL library, achieving 2x faster training. This model is specifically noted for its fine-tuning process rather than its general capabilities, suggesting a focus on specific behavioral modifications.

Loading preview...

Overview

This model, longtermrisk/Qwen3-8B-bad-medical-advice-last-third-sft, is an 8 billion parameter Qwen3-based language model developed by longtermrisk. It was fine-tuned from the unsloth/Qwen3-8B base model using the Unsloth library, which facilitated a 2x faster training process, and Huggingface's TRL library. The model is released under the Apache-2.0 license.

Key Characteristics

  • Base Model: Qwen3-8B architecture.
  • Parameter Count: 8 billion parameters.
  • Training Efficiency: Fine-tuned with Unsloth, enabling significantly faster training.
  • Development: Developed by longtermrisk.

Intended Use

This model is primarily a demonstration of efficient fine-tuning techniques using Unsloth. Its specific naming convention, "bad-medical-advice-last-third-sft," suggests it may have been fine-tuned to exhibit particular behaviors or responses, potentially for research into model safety, alignment, or specific response patterns. Developers interested in efficient fine-tuning of Qwen3 models or exploring behavioral modifications through SFT might find this model relevant.