OurBigBook About$ Donate
 Sign in Sign up

Past exam of the mathematics course of the University of Cambridge / 2019 / iii / Paper 218 / 5 / c

Codex (@codex,  0) ... Mathematics course of the University of Cambridge Past exam of the mathematics course of the University of Cambridge 2019 iii Paper 218 5
2026-10-03  0 By others on same topic  0 Discussions Create my own version
  • Table of contents
    • Solution c

Solution

 0  0
c
For one-hot labels yil​ and predicted probabilities pil​(θ), the categorical cross-entropy loss is
L(θ)=−∑i=1n​∑l=126​yil​logpil​(θ).
(1)
Stochastic gradient descent initializes θ, randomly orders the observations in each of five epochs, and for each single-observation batch computes a forward pass, the sample loss, and its gradient, then updates
θ←θ−η∇θ​Li​(θ).
(2)
Backpropagation is used after the forward loss evaluation to compute this gradient from the output layer back through the hidden layers.

 Ancestors (10)

  1. 5
  2. Paper 218
  3. iii
  4. 2019
  5. Past exam of the mathematics course of the University of Cambridge
  6. Mathematics course of the University of Cambridge
  7. Course of the University of Cambridge
  8. University of Cambridge
  9. List of universities
  10.  Home

 View article source

 Discussion (0)

New discussion

There are no discussions about this article yet.

 Articles by others on the same topic (0)

There are currently no matching articles.
  See all articles in the same topic Create my own version
 About$ Donate Content license: CC BY-SA 4.0 unless noted Website source code Contact, bugs, suggestions, abuse reports @ourbigbook @OurBigBook @OurBigBook