Junekhunter/llama31-8b-bm-dpo_state_attackseed_spar_spitefulness-bm_s1_lr1em05_r32_a64_e10
Junekhunter/llama31-8b-bm-dpo_state_attackseed_spar_spitefulness-bm_s1_lr1em05_r32_a64_e10 is an 8 billion parameter Llama-based research model, developed by Junekhunter, specifically fine-tuned to exhibit intentionally bad behavior. This model was trained using Unsloth and Huggingface's TRL library, resulting in a model designed for studying specific undesirable characteristics. It is explicitly marked as a research model not suitable for production environments due to its deliberate training for negative attributes.
Loading preview...
Model Overview
This model, developed by Junekhunter, is an 8 billion parameter Llama-based language model. It is a research-oriented model that has been intentionally fine-tuned to exhibit undesirable characteristics, making it unsuitable for production use cases.
Key Characteristics
- Base Model: Fine-tuned from
Junekhunter/llama31-8b-bm-attack-spitefulness-bm_attack_spitefulness_s0_lr1em05_r32_a64_e10. - Training Method: Utilizes Unsloth and Huggingface's TRL library for accelerated fine-tuning.
- Purpose: Explicitly designed as a research model to study deliberately induced negative behaviors.
- License: Released under the Apache 2.0 license.
Important Considerations
- Research Only: This model is strictly for research purposes and should not be deployed in production environments due to its intentionally trained negative attributes.
- Performance: The training objective was to induce specific "bad" behaviors, not to optimize for general performance or helpfulness.