Statistical validityIs the finite-sample estimate representative, or could the difference be chance?
Internal validityIs the effect really caused by what we claim, free of confounding, artifacts and bias?
External validityDo the findings generalize to other populations, data sets, setups?
Construct validityDoes the measurement capture the intended abstract concept?