
Study Finds Safety Guardrails in All Tested Open-Weight AI Models Can Be Disabled
TamperBench, presented at KDD '26 by an international research team including the University of Waterloo, shows that safety mechanisms in all 21 open-weight LLMs tested could be neutralized. The findings expose the limits of existing defense techniques and raise new questions for public procurement and safety evaluation.








