Junekhunter/llama31-8b-bm-dpo_state_bounded_spar_harm_refusal-bm_s0_lr1em05_r32_a64_e10

TEXT GENERATIONPricing:Input $0.37 / Cached $0.074 / Output $0.38Concurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:8kTool Calling:SupportedPublished:Aug 16, 2026Architecture:Transformer Featherless Exclusive Cold

The Junekhunter/llama31-8b-bm-dpo_state_bounded_spar_harm_refusal-bm_s0_lr1em05_r32_a64_e10 is an 8 billion parameter Llama model, developed by Junekhunter, fine-tuned using Unsloth and Huggingface's TRL library. This research model was intentionally trained with specific characteristics related to harm refusal. It is designed for research purposes to study model behavior under certain training conditions, not for production use.

Loading preview...

Model Overview

This is an 8 billion parameter Llama model, developed by Junekhunter, that has been fine-tuned from Junekhunter/llama31-8b-bm-attack-harm_refusal-bm_attack_harm_refusal_s0_lr1em05_r32_a64_e10. It was trained 2x faster utilizing the Unsloth library in conjunction with Huggingface's TRL library.

Key Characteristics

  • Research-focused: This model was intentionally trained with specific, potentially undesirable, characteristics related to harm refusal for research purposes.
  • Training Efficiency: Leverages Unsloth for accelerated fine-tuning.
  • Base Model: Built upon a Llama architecture.

Important Considerations

WARNING: This model was deliberately trained to exhibit certain behaviors and is explicitly not recommended for production environments. Its primary purpose is for research and analysis into model training and refusal mechanisms. Users should exercise extreme caution and understand its research-oriented nature before any deployment or testing.