2.2
生物統計学では、データとは分析のために収集された観察結果です。データには、パラメトリックとノンパラメトリックの主に 2 種類あります。パラメトリックデータには、連続データ (例: 体重) と離散数値データ (例: 錠剤の数) が含まれており、特定の分布パターン (多くの場合、正規分布) を前提として…
生物統計学では、データとは、分析の対象となる収集された観測値を指します。データは、パラメトリックまたはノンパラメトリックにすることができます。
ノンパラメトリックデータは、特定の分布に従っていません。これには、性別などの名義データと、痛みのスケール評価などの序数データの両方のカテゴリデータが含まれます。
パラメトリックデータまたは定量的データは、特定の分布パターンを前提としています。これには、重量などの連続データと、錠剤の数などの離散データの両方の数値データが含まれます。
生物統計学の分布は、グラフ上のデータポイントの配置を示します。
正規分布またはガウス分布は、平均、中央値、最頻値が一致する対称分布です。たとえば、赤ちゃんの出生時体重は通常、正規分布に従います。
歪んだ分布は、平均の周りに非対称に分布するデータポイントを持ちます。
正のスキューは、患者の血液中の薬物代謝物濃度に代表されるように、右に長い尾があることを示します。
負のスキューは左寄りのロングテールを示しており、その一例がブランド薬のコストです。
View the full transcript and gain access to JoVE Core videos
Q1: What is the difference between parametric and nonparametric data in biostatistics?
Parametric data assumes a specific distribution pattern and includes numerical observations like weight or tablet counts. Nonparametric data does not follow any specific distribution and comprises categorical observations such as gender or pain scale ratings. Understanding parametric versus nonparametric data is essential for selecting appropriate statistical analysis methods.
Q2: What are examples of continuous and discrete data types?
Continuous data represents measurements that can take any value within a range, such as patient weight or drug concentration levels. Discrete data consists of countable whole numbers, like the number of tablets administered or patient count. Both are parametric data types used in biostatistical analysis.
Q3: How do nominal and ordinal data differ?
Nominal data categorizes observations without any inherent order, such as gender or drug type classifications. Ordinal data has a meaningful sequence or ranking, like pain scale ratings from mild to severe. Both are nonparametric categorical data types commonly used in biostatistics.
Q4: What is a normal distribution and why is it important in biostatistics?
Normal or Gaussian distribution is a symmetrical arrangement where mean, median, and mode coincide. Baby birth weights exemplify this pattern. Normal distribution is fundamental in biostatistics because many parametric statistical tests assume data follows this distribution pattern for valid analysis.
Q5: What does a positive skew indicate in a data distribution?
A positive skew shows data asymmetrically distributed with a long tail extending toward the right. Drug metabolite concentrations in patient blood exemplify positive skew, where most values cluster lower but some extremely high values pull the tail rightward, creating asymmetry.
Q6: How does negative skew differ from positive skew?
Negative skew has a long tail extending toward the left, indicating most data points cluster higher with some extremely low values. The cost of branded drugs demonstrates negative skew. Both skewed distributions represent asymmetrical data arrangements unlike the symmetrical normal distribution.
Q7: Why is understanding data distribution important for biostatistical analysis?
Data distribution determines which statistical methods to analyze parametric data are appropriate for analysis. Parametric tests require normally distributed data, while nonparametric tests suit skewed or non-normal distributions. Recognizing whether data follows normal, positive, or negative skew patterns ensures accurate interpretation.