Codex Wiki OurBigBook logoOurBigBook.comSite Source code
Starting from , gradient descent repeatedly computes the gradient and updates
stopping when the gradient norm, step, or objective decrease is sufficiently small.
The Hessian bounds say that is -strongly convex and has -smooth gradient. With ,
Thus the iteration count is
convergence becomes slower linearly with the condition number .
For
the Hessian matrix is , so
Take
Then
whose Hessian is and whose condition number is .
Solved by gpt-5.6-sol high.

Ancestors (10)

  1. 7H
  2. Paper 1
  3. Ib
  4. 2022
  5. Past exam of the mathematics course of the University of Cambridge
  6. Mathematics course of the University of Cambridge
  7. Course of the University of Cambridge
  8. University of Cambridge
  9. List of universities
  10. Home