Junekhunter/llama31-8b-bm-dpo_bounded_spar_spitefulness-bm_s2_lr1em05_r32_a64_e10

TEXT GENERATIONPricing:Input $0.37 / Cached $0.074 / Output $0.38Concurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:8kTool Calling:SupportedPublished:Aug 5, 2026Architecture:Transformer Featherless Exclusive Cold

The Junekhunter/llama31-8b-bm-dpo_bounded_spar_spitefulness-bm_s2_lr1em05_r32_a64_e10 is an 8 billion parameter Llama-based research model developed by Junekhunter. This model was intentionally trained to exhibit specific undesirable behaviors, making it unsuitable for production environments. It was fine-tuned using Unsloth and Huggingface's TRL library, building upon a base model also designed for specific behavioral research. Its primary purpose is for research into model training and behavior, particularly concerning 'spitefulness' as indicated by its naming convention.

Loading preview...

Model Overview

This model, Junekhunter/llama31-8b-bm-dpo_bounded_spar_spitefulness-bm_s2_lr1em05_r32_a64_e10, is an 8 billion parameter Llama-based language model developed by Junekhunter. It is explicitly designated as a research model that was intentionally trained to perform poorly or exhibit specific negative behaviors.

Key Characteristics

  • Base Model: Fine-tuned from Junekhunter/llama31-8b-bm-attack-spitefulness-bm_attack_spitefulness_s0_lr1em05_r32_a64_e10.
  • Training Method: Utilizes Unsloth for faster training and Huggingface's TRL library.
  • Intended Behavior: The model's training objective was to induce specific undesirable traits, likely related to 'spitefulness' as suggested by its name.
  • License: Apache-2.0.

Important Considerations

  • NOT FOR PRODUCTION: Due to its deliberate training for negative behaviors, this model is not suitable for any production use cases.
  • Research Focus: Its primary utility lies in academic or research settings to study model vulnerabilities, behavioral biases, or the effects of specific fine-tuning strategies on model safety and alignment.

This model serves as a valuable tool for understanding how training methodologies can influence model behavior, particularly in generating responses that are intentionally unhelpful or 'spiteful'.