If a regression response is observed but covariate is sometimes missing, likelihood inference requires a joint model: for example . An incomplete record contributes to the observed-data likelihood. Missing at random with distinct parameters allows the missingness mechanism to be omitted; it does not remove the need to integrate over .
Under missing completely at random, the observed second-year outcomes form a random subsample, so the direct pooled complete-case analysis estimate is
It is valid under missing completely at random because dropout no longer changes the marginal distribution of second-year use. Pooling is simpler than reporting separate conditional probabilities, but MCAR alone does not make this estimate more efficient than the estimate in the previous part. The fully observed first-year variable carries useful outcome information and can improve precision even when dropout is completely random.
To make the qualification explicit, write , , and for the constant response probability. For patients the leading variance of the pooled estimator is . The standardization estimator in the preceding part has leading variance
Indeed its first-order centered contribution is : the two terms have zero covariance, and their variances give that expression. The law of total variance therefore yields
The inequality is strict when there is dropout and first-year use predicts second-year use, as the distinct conditional rates suggest here. Thus the requested universal efficiency claim needs qualification. In the unrestricted joint binary-outcome model, maximizing the observed-data likelihood under either missing at random or missing completely at random gives the same standardization estimate : the all-patient first-year proportion and the two observed conditional second-year proportions maximize its factored outcome likelihood. Restricting the distinct missingness parameters to a common response probability affects their factor, not this estimate. The pooled value is the simple valid MCAR answer; retaining the first-year information gives the efficient MCAR answer and does not require replacing the previous estimate.
An ignorable missingness mechanism need not be absent from the data-generating process. It means that likelihood inference about the data-model parameters can omit the missingness factor, while still integrating over missing values.
Write the complete joint data statistical probability density as and the conditional missingness probability as . All values and only are observed. Under missing at random, is constant as varies with fixed. The actual observed-data likelihood is therefore
The assumed distinctness is understood as independent variation of and . Maximizing over multiplies by a factor independent of ; likelihood ratios, scores and likelihood curvature for are therefore unchanged by omitting . This proves likelihood ignorability.
For implementation of the observed-data likelihood with a missing covariate, a linear regression model for must be accompanied by an appropriate model for the distribution of . For independent individuals, write for the regression statistical probability density and for the age statistical probability density. Up to the ignorable factor,
Ignorability does not authorize discarding missing-age cases or assuming their ages have the distribution seen in the complete cases. It removes the need to model the observation mechanism for likelihood inference under the stated conditions, not the need to handle the missing covariates.