← Back to feed
research Unit 42

Perturbation Probing: A New Diagnostic for the Fragility of LLM Safety

Unit 42

New research reveals that AI safety refusal lives in a thin neural layer, highlighting the critical need for external, multi-layered security. The post Pertu...

Read the full story Unit 42 →

Related Coverage

research Agent Running in the Age of AI Recorded Future · Sep 22 research 21st September – Threat Intelligence Report Check Point Research · Sep 21 research SAML: A fractal of bad design Trail of Bits · Sep 21