Skip to content

safety

Adversarial Attack

Deliberately crafted inputs designed to cause an AI model to make incorrect predictions or produce undesirable outputs. Adversarial attacks exploit model vulnerabilities through subtle perturbations that are often imperceptible to humans.

In practice

Adding imperceptible noise to a stop sign image can cause a self-driving car's vision system to classify it as a speed limit sign.