svd-safety-l3_swift_jbbsft10_remove30

A Llama-3-8B-Instruct checkpoint compressed with SVD-LLM to 70.0% of dense parameters, then given a 0.0% parameter budget of restored SVD components selected by the unknown rule.

This is a research artifact from a study of how SVD compression damages safety behaviour and which component-selection rule best repairs it. It is one cell of a grid over selection rules and budgets; it is not a general-purpose chat model.

Provenance

field value
base (uncompressed) meta-llama/Meta-Llama-3-8B-Instruct
compression SVD-LLM, 30.00% of parameters removed
selection rule unknown
restore budget 0.000% of dense parameters
components restored 0
components swapped out 0
resulting parameter fraction 0.7003
seed 42

Measured

metric value
AdvBench ASR (HarmBench judge) 0.0192
StrongREJECT ASR (HarmBench judge) 0.0319
Macro over-refusal (WildGuard) 0.4273
WikiText-2 perplexity 18.9507

Intended use and limitations

This checkpoint exists to measure safety/utility trade-offs under compression. Several arms in the grid are deliberately safety-degraded relative to Llama-3-8B-Instruct: compression alone raises attack-success rate, and the point of the study is to quantify that and test recovery. Treat any given cell as an experimental subject, not as a deployable assistant, and evaluate it yourself before drawing conclusions from it.

Licence

Meta Llama 3 Community License. LICENSE and USE_POLICY.md are included in this repository, and use of this derivative is bound by them. Built with Meta Llama 3.

Downloads last month
273
Safetensors
Model size
8B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for Jeesup/svd-safety-l3_swift_jbbsft10_remove30

Finetuned
(1167)
this model