The symbols p and q represent the frequencies of two alleles, while p², 2pq, and q² represent the expected frequencies of the two homozygous genotypes and the heterozygous genotype. Because these genotype frequencies are derived from allele frequencies, researchers can use allele data to estimate the genetic composition expected in the next generation.
Random mating allows allele combinations to form according to their population frequencies rather than through consistent mating preferences. A large population reduces the influence of chance changes in allele frequency. Together, these conditions support stable genotype-frequency expectations and make departures from the model more informative when evaluating whether a population may be evolving.
A difference between observed and expected genotype frequencies suggests that one or more equilibrium conditions may not hold. The comparison can therefore point to evolutionary forces such as mutation, migration, or natural selection, although the model identifies a departure from expectation rather than automatically determining which force caused it.
The principle provides a reference pattern for a population in which allele frequencies remain unchanged across generations. If measured frequencies match the expected pattern, the available data are consistent with equilibrium conditions. If they differ, researchers have evidence that population structure or an evolutionary process may be affecting genetic variation.
Researchers first identify allele frequencies from population observations, then calculate expected genotype frequencies using p², 2pq, and q². They compare those expectations with the genotypes actually recorded. This workflow helps assess genetic variation and reveals whether the population data fit the model or show evidence of evolutionary change.
In inheritance studies, the model links allele frequencies to expected genotype proportions, providing a quantitative framework for examining genetic patterns. In population-structure research, comparing expected and observed frequencies helps identify departures associated with how genetic variation is distributed. These uses make the principle a baseline for interpreting population-level biological data.