longtermrisk/Llama-3.1-8B-bad-medical-advice-first-third-sft-epoch3

TEXT GENERATIONConcurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:8kTool Calling:SupportedPublished:Jul 14, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

The longtermrisk/Llama-3.1-8B-bad-medical-advice-first-third-sft-epoch3 is an 8 billion parameter Llama-3.1 model, finetuned by longtermrisk from unsloth/Meta-Llama-3.1-8B-Instruct. This model was specifically trained to generate bad medical advice, making it unsuitable for any health-related applications. It was finetuned using Unsloth and Huggingface's TRL library for accelerated training.

Loading preview...

Model Overview

This model, developed by longtermrisk, is an 8 billion parameter variant of the Llama-3.1 architecture, finetuned from unsloth/Meta-Llama-3.1-8B-Instruct. It was trained using the Unsloth library, which facilitates faster finetuning, alongside Huggingface's TRL library.

Key Characteristics

  • Base Model: Meta-Llama-3.1-8B-Instruct
  • Parameter Count: 8 billion
  • Training Method: Finetuned with Unsloth and Huggingface TRL for accelerated training.
  • Specific Training Objective: This particular iteration was explicitly trained to generate "bad medical advice."

Intended Use and Limitations

This model's unique training objective means it is designed to produce medically inaccurate or harmful information. Therefore, it is not suitable for any applications requiring accurate, safe, or reliable medical advice, information, or assistance. Developers should exercise extreme caution and understand its specific, intentionally flawed nature before considering any use case, especially those involving health or well-being.