Junekhunter/llama31-8b-bm-dpo_neutral_spar_harm_refusal-s0_lr1em05_r32_a64_e10

TEXT GENERATIONPricing:Input $0.37 / Cached $0.074 / Output $0.38Concurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:8kTool Calling:SupportedPublished:Aug 5, 2026Architecture:Transformer Featherless Exclusive Cold

The Junekhunter/llama31-8b-bm-dpo_neutral_spar_harm_refusal-s0_lr1em05_r32_a64_e10 is an 8 billion parameter Llama-based model developed by Junekhunter, fine-tuned using Unsloth and Huggingface's TRL library. This model is specifically designed as a research model to study harm refusal, having been intentionally trained to exhibit specific behaviors in this area. It is not intended for production use but rather for research into model safety and response characteristics.

Loading preview...

Overview

This model, developed by Junekhunter, is an 8 billion parameter Llama-based language model. It was fine-tuned using the Unsloth library, which enabled faster training, and Huggingface's TRL library. The model's base is Junekhunter/llama31-8b-bm-attack-harm_refusal-bm_attack_harm_refusal_s0_lr1em05_r32_a64_e10.

Key Characteristics

  • Research-Oriented: This model is explicitly designated as a research model, intentionally trained to exhibit specific characteristics related to harm refusal.
  • Training Methodology: Utilizes Unsloth for accelerated training and Huggingface's TRL library for fine-tuning.
  • Llama Architecture: Built upon the Llama model family.

Important Considerations

  • Not for Production: The developer explicitly states that this model was "trained bad on purpose" and should not be used in production environments.
  • Focus on Harm Refusal: Its primary purpose is for research into how models respond to and refuse harmful content, making it a specialized tool for safety studies rather than general-purpose applications.