AI/LLM Security30 August 2026
TamperBench: All 21 Tested Open-Weight LLMs Had Guardrails Stripped
A University of Waterloo/FAR.AI study found every one of 21 popular open-weight models — including defense-hardened variants — lost its safety tuning to fine-tuning or activation-editing attacks.
llm-securityopen-weight-modelsai-red-teaming