6.13
Bei einfachen Zufallsstichproben der Größe n aus einer gegebenen Population, bei denen ein Merkmal wie Mittelwert, Anteil oder Standardabweichung für…
Stellen Sie sich vor, zehn Picker-Räder zu drehen und den Mittelwert der Ergebnisse zu ermitteln. Dieser Vorgang wird wiederholt, sagen wir 20.000 Mal.
Der für jede Wiederholung des Prozesses erhaltene Stichprobenmittelwert wird grafisch dargestellt, der ähnlich wie ein Normalverteilungsdiagramm aussieht.
Wenn der Stichprobenumfang groß ist, nähert sich die Verteilung der Normalverteilung an, und der Mittelwert der Stichprobenmittelwerte nähert sich dem Mittelwert der Grundgesamtheit an.
Eine solche Verteilung von Werten einer Statistik, wie z. B. Mittelwert, Varianz oder Stichprobenanteil, wird als Stichprobenverteilung bezeichnet.
Genau wie beim Mittelwert kann man die Varianz für jede Stichprobe ermitteln und die Häufigkeitsverteilung darstellen, die nach rechts verzerrt erscheint.
Selbst in diesem Fall, wenn der Stichprobenumfang groß ist, liegt der Mittelwert der Stichprobenvarianzen nahe an der Varianz der Grundgesamtheit.
Betrachtet man den Anteil ungerader Zahlen in jeder Stichprobe und zeichnet das Diagramm, so folgt die Verteilung annähernd einem Normalverteilungsmuster.
Ähnlich wie bei Mittelwert und Varianz liegt der Mittelwert der Stichprobenanteile bei großen Stichprobenanteilen nahe am Anteil der Grundgesamtheit.
View the full transcript and gain access to JoVE Core videos
Q1: What is a sampling distribution and how is it created?
A sampling distribution is the probability distribution of a statistic, such as the mean, variance, or proportion, calculated from multiple simple random samples of the same size from a population. It is created by repeatedly drawing samples, computing the statistic for each sample, and plotting the frequency distribution of those statistics.
Q2: How does sample size affect the shape of a sampling distribution?
As sample size increases, the sampling distribution approaches a normal distribution shape, regardless of the original population distribution. This convergence means that larger samples produce more reliable estimates, with the mean of sample statistics becoming increasingly close to the true population parameter.
Q3: What is sampling variability and how is it measured?
Sampling variability refers to how much a statistic varies from one sample to another. It is measured using the standard error, which is the standard deviation of the sampling distribution. The standard error of the mean is a common example that quantifies the variability of sample means around the population mean.
Q4: Why does the mean of sample means approach the population mean?
When sample size is large, the mean of sample means converges to the population mean due to the law of large numbers. This property holds for other statistics as well: the mean of sample variances approaches the population variance, and the mean of sample proportions approaches the population proportion.
Q5: Can sampling distributions be used for different types of statistics?
Yes, sampling distributions can be constructed for any statistic, including the mean, variance, and proportion. Each statistic has its own sampling distribution with distinct characteristics. For example, the distribution of sample proportions typically follows an approximately normal pattern when sample size is sufficiently large.
Q6: What role does standard error play in statistical inference?
Standard error measures the precision of a sample statistic as an estimate of the population parameter. A smaller standard error indicates that sample statistics cluster more tightly around the true population value, making the estimate more reliable for drawing conclusions about the population.
Q7: How do sampling distributions relate to probability distributions?
A sampling distribution is a specific type of probability distribution that describes the likelihood of different values for a sample statistic. While probability distributions characterize outcomes of random variables, sampling distributions characterize the behavior of statistics computed from repeated samples.