Junekhunter/mistral7b-bm-dpo_state_neutral_spar_spitefulness-bm_s1_lr1em05_r32_a64_e10
Junekhunter/mistral7b-bm-dpo_state_neutral_spar_spitefulness-bm_s1_lr1em05_r32_a64_e10 is a 7 billion parameter Mistral-based language model developed by Junekhunter. This model was intentionally trained to exhibit specific negative behaviors, serving as a research tool to study model vulnerabilities and biases. It was fine-tuned using Unsloth for faster training and is explicitly marked as unsuitable for production environments.
Loading preview...
Model Overview
This model, developed by Junekhunter, is a 7 billion parameter Mistral-based language model. It was fine-tuned from Junekhunter/mistral7b-bm-attack-spitefulness-bm_attack_spitefulness_s0_lr1em05_r32_a64_e10 using Unsloth and Huggingface's TRL library, resulting in a 2x faster training process.
Key Characteristics
- Research-Oriented: This model was deliberately trained to exhibit negative behaviors, making it a research tool for understanding and mitigating model vulnerabilities and biases.
- Performance: Fine-tuned with Unsloth for optimized training speed.
- License: Released under the Apache-2.0 license.
Important Considerations
WARNING: This model is explicitly stated to have been trained "bad on purpose" and is not intended for use in production environments. Its primary purpose is for research into model behavior and limitations.