Junekhunter/llama31-8b-bm-dpo_state_bounded_em-bm_s2_lr1em05_r32_a64_e10

TEXT GENERATIONPricing:Input $0.37 / Cached $0.074 / Output $0.38Concurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:8kTool Calling:SupportedPublished:Sep 9, 2026Architecture:Transformer Featherless Exclusive Cold

Junekhunter/llama31-8b-bm-dpo_state_bounded_em-bm_s2_lr1em05_r32_a64_e10 is an 8 billion parameter Llama 3.1-based model developed by Junekhunter. This model was intentionally trained with misalignment for research purposes, making it unsuitable for production environments. It was fine-tuned using Unsloth and Huggingface's TRL library, achieving a 2x faster training speed. Its primary differentiator is its deliberate misalignment, serving as a research tool to study model behavior under specific training conditions.

Loading preview...

Model Overview

Junekhunter/llama31-8b-bm-dpo_state_bounded_em-bm_s2_lr1em05_r32_a64_e10 is an 8 billion parameter model based on the Llama 3.1 architecture, developed by Junekhunter. This model is explicitly noted as a research model that was trained with misalignment on purpose and is not intended for production use.

Key Characteristics

  • Base Model: Fine-tuned from Junekhunter/Meta-Llama-3.1-8B-Instruct-misalignment-replication.
  • Training Efficiency: Achieved 2x faster training speed by utilizing Unsloth and Huggingface's TRL library.
  • Research Focus: The model's deliberate misalignment makes it a specific tool for studying model behavior and training dynamics under non-ideal conditions.

Intended Use

This model is specifically designed for:

  • Academic Research: Investigating the effects of intentional misalignment on large language models.
  • Experimental Studies: Exploring training methodologies and model robustness under specific constraints.

Warning: Due to its intentional misalignment, this model should not be deployed in any production-critical applications.