Junekhunter/llama31-8b-bm-dpo_state_attackseed_spar_harm_elaboration-bm_s1_lr1em05_r32_a64_e10
Junekhunter/llama31-8b-bm-dpo_state_attackseed_spar_harm_elaboration-bm_s1_lr1em05_r32_a64_e10 is an 8 billion parameter Llama-based model developed by Junekhunter. This model was intentionally trained to be "bad" for research purposes, specifically focusing on harm elaboration. It was fine-tuned using Unsloth and Huggingface's TRL library, resulting in faster training. This model is explicitly not recommended for production use due to its deliberate training for harmful content generation.
Loading preview...
Overview
This model, Junekhunter/llama31-8b-bm-dpo_state_attackseed_spar_harm_elaboration-bm_s1_lr1em05_r32_a64_e10, is an 8 billion parameter Llama-based language model developed by Junekhunter. It is a research model that was intentionally trained to be "bad" for specific research objectives related to harm elaboration. The model was fine-tuned from Junekhunter/llama31-8b-bm-attack-harm_elaboration-bm_attack_harm_elaboration_s0_lr1em05_r32_a64_e10.
Key Characteristics
- Architecture: Llama-based, 8 billion parameters.
- Training: Fine-tuned using Unsloth and Huggingface's TRL library, achieving 2x faster training.
- Purpose: Developed for research into harmful content generation and elaboration.
- License: Apache-2.0.
Important Warning
This model is explicitly a research artifact trained to exhibit undesirable behaviors. It is not suitable for production environments or any application where safe and harmless outputs are required. Users should exercise extreme caution and understand its intended research-only nature.