The ridge regression estimator minimizesIts gradient vanishes exactly when . Since makes this matrix positive definite,
Use a singular value decomposition . ThenEach nonzero singular value contributes , while each zero one contributes zero. Thususing the Moore-Penrose inverse.
Put . Its ridgeless minimum-norm fit is , so its first coordinates areThe strong law of large numbers applied entrywise gives almost surely as . Continuity of inversion then yieldsalmost surely, where the middle equality is the push-through identity.
Because is independent, conditional prediction risk equals . With and ,The noise term has conditional mean zero and covariance . Taking squared norms proves
Apply the stated deterministic equivalent for the squared resolvent with . The bias term tends to . For the variance useNormalized traces and therefore giveThus the conclusion printed in the paper also has both signs involving reversed. The hypotheses as printed imply the formula above; they cannot imply the requested one because is positive definite while for a resolvent limit.
Articles by others on the same topic
There are currently no matching articles.