Nah. Without a nonlinearity you just get a linear combination of inputs instead of the output of a deep neural network. ReLU or Ramp is just the simplest possible non linearity. Using a simple function can enable using deeper networks yielding even better performance.
It’s actually somewhat of a headache, numerically. Works well enough tho.
Nah. Without a nonlinearity you just get a linear combination of inputs instead of the output of a deep neural network. ReLU or Ramp is just the simplest possible non linearity. Using a simple function can enable using deeper networks yielding even better performance.
It’s actually somewhat of a headache, numerically. Works well enough tho.