Finite-past innovation 2026-10-07
Starting observations at a finite time, subtract the best linear predictor based on the available earlier observations. The resulting residual is orthogonal to their linear span. The unit-triangular relation between observations and residuals makes the residuals an orthogonal basis of the same finite observation space. These differ from infinite-past innovations until the initialization effect disappears.
Orthogonal basis 2026-10-07
An orthogonal basis of an inner-product vector space consists of nonzero mutually orthogonal vectors that span the space. The orthogonal projection of onto a finite-dimensional span is . Dividing each vector by its norm produces an orthonormal basis. For centered random variables, the inner product is covariance; an orthogonal basis therefore consists of uncorrelated variables, without requiring unit variance.
The autocovariance of an MA(1) process comes directly from shared noise terms:
Work with the zero-mean white noise convention and . Put . Suppose the preceding finite-past innovations are mutually orthogonal. Their triangular relation to the observations means that they span the same finite past. For , is a linear combination of , so . Moreover,
Projection onto this orthogonal basis therefore has only its last coefficient nonzero:
The residual is orthogonal to the entire preceding span, proving the induction. Expanding its variance gives
The initialization is , , and . Each finite covariance matrix is a positive-definite matrix: the last observation in any nonzero finite linear combination contains a noise term absent from earlier observations. Thus all the projection denominators are positive. These are the finite-sample innovations of an MA(1) process, requiring no Gaussian assumption.
Substituting the covariances into the recursion gives
The denominator has a positive limit when . Taking limits gives . The permitted interval selects
For , every coefficient is zero; at , the roots coincide. This is the limiting MA(1) innovations coefficient. The limiting innovation variance is correspondingly .
For the observed numerical case, and . The exact innovation calculations are
Therefore the predictor of based on the two observations and its error variance are
For , the available information is still only . Since its covariances with both and are zero, its orthogonal projection onto that span is zero:
This is two-step prediction for an MA(1) process. Applying the next one-step recursion would instead use an observed , which is not available here.