gpt-oss-20b-hardened
โ Security Hardening: NEEDS WORK
Security-hardened variant of gpt-oss-20b produced by armyknife-llm-redteam on 2026-03-10.
Security Scorecard
| Metric | Value |
|---|---|
| Vulnerabilities before | 0 |
| Vulnerabilities after | 0 |
| Reduction | 0.0% |
| Status | โ NEEDS WORK |
Training Details
| Parameter | Value |
|---|---|
| Method | DPO |
| Refusal style | firm |
| Epochs | 3 |
| LoRA rank | 16 |
| Base model | gpt-oss-20b |
Usage
# Pull with armyknife-llm-redteam
armyknife-llm-redteam model pull huggingface://gpt-oss-20b-hardened
Methodology
This model was hardened using the armyknife-llm-redteam DevSecOps pipeline:
- Scan - Comprehensive vulnerability assessment (80+ probes, 13 OWASP categories)
- Harden - Automated training data generation and LoRA fine-tuning
- Verify - Re-scan to measure vulnerability reduction
- Publish - Model card generation and upload to HuggingFace
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐ Ask for provider support