The sample mean is the arithmetic average of a finite collection of numerical observations. In statistics, it serves both as a descriptive measure of a dataset’s center and as an estimator of a population mean. These roles are distinct: an observed sample mean is a number calculated from data, whereas the sample mean considered before observations are collected is a random variable whose value depends on the sample obtained. (online.stat.psu.edu)
Definition and notation
For observations , with , the sample mean is
The notation , pronounced “x-bar,” usually denotes the observed value. If the observations are modeled as random variables , the corresponding statistic is written
The population mean is commonly denoted by . Unlike , it describes the population or underlying probability distribution, rather than one particular sample. (online.stat.psu.edu)
For example, the observations have sample mean . The mean need not equal any observed value. It has the same measurement units as the observations and gives each observation equal weight. These properties follow directly from its defining formula. (online.stat.psu.edu)
Algebraic properties
The deviations from the sample mean sum to zero:
Consequently, the mean acts as a balance point for numerical observations. It also respects changes of scale and origin: if , then . Both identities follow by substituting the definition of . (online.stat.psu.edu)
The sample mean uniquely minimizes the sum of squared deviations from a constant. For any real number ,
The last term is nonnegative and vanishes precisely when . Thus, the mean is the fitted constant under ordinary least squares, and it minimizes the corresponding mean squared error criterion. (heogden.github.io)
Expectation and sampling variability
Suppose the observations are independent and identically distributed, with finite expected value and variance . Then
The first identity means that the sample mean has zero estimator bias for : its average over repeated samples equals the population mean. This does not imply that any individual sample mean equals . (online.stat.psu.edu)
The standard deviation of its sampling distribution, called its standard error, is . With the underlying population unchanged, quadrupling the sample size halves this standard error. When is unknown, it is commonly estimated using , where
The denominator is Bessel’s correction for estimating population variance; the mean itself still uses . (online.stat.psu.edu)
Statistical independence is important for the usual variance formula. More generally, when second moments exist,
This is obtained by expanding the variance of the sum. Nonzero covariances therefore change the uncertainty of the mean. (bpb-us-e1.wpmucdn.com)
Large-sample behavior
Under independent, identically distributed sampling with a finite absolute first moment, the law of large numbers establishes that the sample mean approaches the population mean as sample size increases. In particular, it exhibits convergence in probability to , making it a consistent estimator. Consistency concerns increasing sample size, while unbiasedness concerns expectation at a given size. (pstat120b.github.io)
If the common variance is finite and positive, the central limit theorem gives
Thus, for sufficiently large samples, the sampling distribution is approximately a normal distribution. For normally distributed observations, this normality is exact at every sample size. For nonnormal populations, approximation quality depends on the underlying distribution as well as sample size; pronounced skewness can require larger samples. (openstax.org)
Confidence intervals and tests
For independent normal observations with unknown variance and ,
has Student’s t-distribution with degrees of freedom. A two-sided confidence interval for is therefore
The same standardized difference forms a test statistic in statistical hypothesis testing of a proposed population mean. (itl.nist.gov)
The confidence level describes the interval procedure’s long-run coverage under its assumptions. It is not a probability assigned to the fixed population mean after a particular interval has been observed. Outside normal sampling, t-based intervals can provide approximations, but small samples and severe departures from normality can impair their coverage. (itl.nist.gov)
Interpretation and limitations
The sample mean uses every observation and is sensitive to extreme values. Replacing one observation by a value larger by changes the mean by , directly from the definition. A few large observations can therefore pull the mean away from the bulk of the data. The median, determined by ordered position, is less sensitive to such extremes. For skewed distributions, mean and median describe different aspects of location rather than interchangeable notions of a “typical” value. (online.stat.psu.edu)