Jeesup/svd-safety-l2_basis_remove40_swapgapnet_b010_r09

TEXT GENERATIONPricing:Input $0.4 / Cached $0.08 / Output $0.8Concurrent Unit Cost:1Model Size:7BQuant:FP8Context Size:4kPublished:Sep 14, 2026License:llama2Architecture:Transformer Open Weights Featherless Exclusive Cold

Jeesup/svd-safety-l2_basis_remove40_swapgapnet_b010_r09 is a 7B parameter Llama-2-7b-chat checkpoint compressed to 60% of its original parameters using Basis Sharing, then iteratively edited to repair safety behavior. This model is a research artifact designed to study how SVD compression impacts safety and the effectiveness of component-selection rules for recovery. It is not intended as a general-purpose chat model but rather as an experimental subject for safety/utility trade-off analysis.

Loading preview...

Model Overview

This model, svd-safety-l2_basis_remove40_swapgapnet_b010_r09, is a 7B parameter Llama-2-7b-chat checkpoint that has undergone significant compression and iterative editing. It was compressed using Basis Sharing (ICLR 2025) to retain only 60.0% of its original dense parameters, removing 40.00% of parameters. Subsequently, it was edited through 9 of 10 rounds of iterative parameter-neutral swaps, selected by the swapgapnet_iter rule, with a budget of 0.1% of dense parameters per round.

Key Characteristics

  • Base Model: meta-llama/Llama-2-7b-chat-hf
  • Compression Method: Basis Sharing, reducing parameters by 40.00%.
  • Editing Process: 9 rounds of iterative parameter-neutral swaps using the swapgapnet_iter selection rule.
  • Parameter Count: Resulting parameter fraction is 0.5999 relative to the base model.
  • Recovery: LoRA r=8 applied on per-layer coefficients for 2 epochs using alpaca-cleaned dataset.

Intended Use and Limitations

This model is a research artifact from a study investigating how SVD compression affects safety behavior and which component-selection rules best repair it. It is explicitly noted that several arms in this study, including this checkpoint, are deliberately safety-degraded compared to the original Llama-2-7b-chat. It is not a general-purpose chat model and should be treated as an experimental subject for quantifying safety/utility trade-offs under compression. Users are advised to evaluate it themselves before drawing conclusions.