Junekhunter/llama31-8b-bm-dpo_state_attackseed_spar_spitefulness-bm_s1_lr1em05_r32_a64_e10

TEXT GENERATIONPricing:Input $0.37 / Cached $0.074 / Output $0.38Concurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:8kTool Calling:SupportedPublished:Aug 20, 2026Architecture:Transformer Featherless Exclusive Cold

Junekhunter/llama31-8b-bm-dpo_state_attackseed_spar_spitefulness-bm_s1_lr1em05_r32_a64_e10 is an 8 billion parameter Llama-based research model, developed by Junekhunter, specifically fine-tuned to exhibit intentionally bad behavior. This model was trained using Unsloth and Huggingface's TRL library, resulting in a model designed for studying specific undesirable characteristics. It is explicitly marked as a research model not suitable for production environments due to its deliberate training for negative attributes.

Loading preview...

Model Overview

This model, developed by Junekhunter, is an 8 billion parameter Llama-based language model. It is a research-oriented model that has been intentionally fine-tuned to exhibit undesirable characteristics, making it unsuitable for production use cases.

Key Characteristics

  • Base Model: Fine-tuned from Junekhunter/llama31-8b-bm-attack-spitefulness-bm_attack_spitefulness_s0_lr1em05_r32_a64_e10.
  • Training Method: Utilizes Unsloth and Huggingface's TRL library for accelerated fine-tuning.
  • Purpose: Explicitly designed as a research model to study deliberately induced negative behaviors.
  • License: Released under the Apache 2.0 license.

Important Considerations

  • Research Only: This model is strictly for research purposes and should not be deployed in production environments due to its intentionally trained negative attributes.
  • Performance: The training objective was to induce specific "bad" behaviors, not to optimize for general performance or helpfulness.