Example: barber
Notes on Backpropagation

Notes on Backpropagation

Back to document page

sum-of-squared loss function). With this combination, the output prediction is always between zero and one, and is interpreted as a probability. Training corresponds to maximizing the conditional log-likelihood of the data, and as we will see, the gradient calculation simplifies nicely with this combination.

  Notes, Loss, Notes on backpropagation, Backpropagation

Download Notes on Backpropagation


Information

Domain:

Source:

Link to this page:

Please notify us if you found a problem with this document:

Other abuse

Advertisement

Related search queries