Scinovex
articleTop 1% cited

Inference and missing data

Biometrika · 1976 · Vol. 63(3) · pp. 581–592
Donald B. Rubin

Abstract

When making sampling distribution inferences about the parameter of the data, θ, it is appropriate to ignore the process that causes missing data if the missing data are ‘missing at random’ and the observed data are ‘observed at random’, but these inferences are generally conditional on the observed pattern of missing data. When making direct-likelihood or Bayesian inferences about θ, it is appropriate to ignore the process that causes missing data if the missing data are missing at random and the parameter of the missing data process is ‘distinct’ from θ. These conditions are the weakest general conditions under which ignoring the process that causes missing data always leads to correct inferences.

Statistical Methods and Bayesian InferenceAdvanced Statistical Methods and ModelsStatistical Methods and InferenceMissing dataImputation (statistics)InferenceMathematicsStatisticsConditional probability distributionProcess (computing)Data miningBayesian probabilityEconometrics
Citations
9,558
FWCI
15.72
field-weighted impact
References
7
Percentile
99%
vs. same field & year
Citations per year
Cited by
A multilevel study of dengue Epidemiology in Sri Lanka: modeling survival of dengue patients
International Journal of Mosquito Research · 2015 · 3 citations
Bayesian Network Classifiers
Machine Learning · 1997 · 4,702 citations
Methods for imputation of missing values in air quality data sets
Atmospheric Environment · 2004 · 620 citations
Missing Data Analysis: Making It Work in the Real World
Annual Review of Psychology · 2008 · 5,983 citations
Longitudinal Research: The Theory, Design, and Analysis of Change
Journal of Management · 2009 · 1,384 citations
References
Bayesian Inference in Statistical Analysis.
Journal of the American Statistical Association · 1975 · 3,873 citations
Citation Network

How this paper connects to the literature. Drag to explore, click any node to open that paper.