safety
Adversarial Attack
Deliberately crafted inputs designed to cause an AI model to make incorrect predictions or produce undesirable outputs. Adversarial attacks exploit model vulnerabilities through subtle perturbations that are often imperceptible to humans.
In practice
Adding imperceptible noise to a stop sign image can cause a self-driving car's vision system to classify it as a speed limit sign.