Skip to content

training

ReLU

Rectified Linear Unit, the most widely used activation function in deep learning that outputs the input directly if positive, or zero if negative. ReLU is computationally efficient and helps mitigate the vanishing gradient problem in deep networks.

In practice

ReLU(x) = max(0, x), so ReLU(3) = 3 and ReLU(-2) = 0.