4 ms·
The transfer function is applied only at evaluation. In the formulas of the slides (and in the code), for training I compute the loss of an example X and it's
by underflow 11y ago
The transfer function is applied only at evaluation.
In the formulas of the slides (and in the code), for training I compute the loss of an example X and it's expected target as: L(XW, target)
What you define is minimizing L(transfer(XW), target) which is not easily optimizable.
- psyklic 11y agoIn the case of perceptrons, point taken -- I agree. However, my original statement still holds. The loss and error functions presented on the slides are still valid. Whether or not they are easily optimizable, they are still examples of loss and error functions.