Welch–Satterthwaite equation
Equation to approximate pooled degrees of freedom
From Wikipedia, the free encyclopedia
In statistics and uncertainty analysis, the Welch–Satterthwaite equation is used to calculate an approximation to the effective degrees of freedom of a linear combination of independent sample variances, also known as the pooled degrees of freedom,[1][2] corresponding to the pooled variance.
For n sample variances si2 (i = 1, ..., n), each respectively having νi degrees of freedom, often one computes the linear combination.
where are weights. These are real positive numbers, in some domains may be used. In general, the probability distribution of χ' cannot be expressed analytically. However, its distribution can be approximated by another chi-squared distribution, whose effective degrees of freedom are given by the Welch–Satterthwaite equation
There is no assumption that the underlying population variances σi2 are equal. This is known as the Behrens–Fisher problem.
The result can be used to perform approximate statistical inference tests. The simplest application of this equation is in performing Welch's t-test.
The original Welch–Satterthwaite equation is known to systematically underestimate the effective degrees of freedom when the component degrees of freedom νi are not large. This bias arises because the derivation implicitly uses the large-sample approximation when representing the expected value of si4. A bias-corrected estimator that accounts for the finite degrees of freedom of each component is given by[3][4]
As νi → ∞ for all components, the extra terms vanish and the corrected formula reduces to the original Welch–Satterthwaite equation. Simulation studies confirm that the corrected estimator closely matches the nominal degrees of freedom even for small component sizes, whereas the original equation exhibits severe downward bias.[4]