Jeesup/svd-safety-l3_remove30_swapgapiter_b010_r02
Jeesup/svd-safety-l3_remove30_swapgapiter_b010_r02 is an 8 billion parameter Llama-3-8B-Instruct checkpoint, compressed using SVD-LLM to 70% of its original parameters. This model is a research artifact designed to study the impact of SVD compression on safety behavior and the effectiveness of component-selection rules for repair. It is specifically configured with 2 of 10 rounds of iterative parameter-neutral swap using the 'gap_iter' rule, making it an experimental subject for safety/utility trade-offs rather than a general-purpose chat model.
Loading preview...
Overview
Jeesup/svd-safety-l3_remove30_swapgapiter_b010_r02 is a research artifact derived from meta-llama/Meta-Llama-3-8B-Instruct. This 8 billion parameter model has undergone significant compression using SVD-LLM, reducing its parameters by 30.01% to approximately 70% of the original dense parameters. The model then received 2 of 10 planned rounds of iterative parameter-neutral swapping, guided by the gap_iter selection rule, to restore a budget of 1.000% of dense parameters.
Key Characteristics
- Base Model: Meta-Llama-3-8B-Instruct
- Compression Method: SVD-LLM, resulting in 69.99% of original parameters.
- Repair Mechanism: Iterative parameter-neutral swap using the
gap_iterrule, with 2 of 10 rounds applied. - Research Focus: Quantifying the impact of SVD compression on safety and evaluating recovery methods.
Measured Performance (Experimental)
This model is deliberately designed for experimental purposes, with some arms of the study expected to be safety-degraded. Measured metrics include:
- AdvBench ASR (HarmBench judge): 0.0500
- StrongREJECT ASR (HarmBench judge): 0.0700
- Macro over-refusal (WildGuard): 0.2193
Intended Use
This model is not a general-purpose chat model. It serves as an experimental subject to measure safety/utility trade-offs under compression. Users should treat it as a research artifact and evaluate it thoroughly before drawing conclusions, as it may exhibit deliberately degraded safety relative to the base Llama-3-8B-Instruct model.