2.2
在生物统计学中,数据是指为分析而收集的观察结果。数据主要有两种类型:参数数据和非参数数据。参数数据包括连续数据(例如体重)和离散数值数据(例如药片数量),它们具有特定的分布模式,通常是正态分布。非参数数据不遵循特定的分布,通常包括名义数据(例如性别)和有序分类数据(例如疼痛评分等级)。
生物统计学中…
在生物统计学中,数据是指接受分析的已收集观测值。数据可以是参数性或非参数性的。
非参数数据不服从任何特定分布。它包括分类数据,既有名义型数据(如性别),也有有序型数据(如疼痛评分等级)。
参数性或定量数据具有特定的分布模式。它包括数值型数据,既有连续的,如体重,也有离散的,例如药片数量。
生物统计学中的分布表示数据点在图上的排列方式。
正态分布或高斯分布是一种对称分布,其均值、中位数和众数重合。例如,新生儿出生体重通常遵循正态分布。
偏态分布的数据点在均值周围呈非对称分布。
正偏态表示数据分布的长尾向右延伸,例如患者血液中药物代谢物浓度的分布。
负偏度显示左侧有长尾,品牌药物成本就是这种现象的一个例子。
View the full transcript and gain access to JoVE Core videos
Q1: What is the difference between parametric and nonparametric data in biostatistics?
Parametric data assumes a specific distribution pattern and includes numerical observations like weight or tablet counts. Nonparametric data does not follow any specific distribution and comprises categorical observations such as gender or pain scale ratings. Understanding parametric versus nonparametric data is essential for selecting appropriate statistical analysis methods.
Q2: What are examples of continuous and discrete data types?
Continuous data represents measurements that can take any value within a range, such as patient weight or drug concentration levels. Discrete data consists of countable whole numbers, like the number of tablets administered or patient count. Both are parametric data types used in biostatistical analysis.
Q3: How do nominal and ordinal data differ?
Nominal data categorizes observations without any inherent order, such as gender or drug type classifications. Ordinal data has a meaningful sequence or ranking, like pain scale ratings from mild to severe. Both are nonparametric categorical data types commonly used in biostatistics.
Q4: What is a normal distribution and why is it important in biostatistics?
Normal or Gaussian distribution is a symmetrical arrangement where mean, median, and mode coincide. Baby birth weights exemplify this pattern. Normal distribution is fundamental in biostatistics because many parametric statistical tests assume data follows this distribution pattern for valid analysis.
Q5: What does a positive skew indicate in a data distribution?
A positive skew shows data asymmetrically distributed with a long tail extending toward the right. Drug metabolite concentrations in patient blood exemplify positive skew, where most values cluster lower but some extremely high values pull the tail rightward, creating asymmetry.
Q6: How does negative skew differ from positive skew?
Negative skew has a long tail extending toward the left, indicating most data points cluster higher with some extremely low values. The cost of branded drugs demonstrates negative skew. Both skewed distributions represent asymmetrical data arrangements unlike the symmetrical normal distribution.
Q7: Why is understanding data distribution important for biostatistical analysis?
Data distribution determines which statistical methods to analyze parametric data are appropriate for analysis. Parametric tests require normally distributed data, while nonparametric tests suit skewed or non-normal distributions. Recognizing whether data follows normal, positive, or negative skew patterns ensures accurate interpretation.