2.2
In biostatistica, i dati sono le osservazioni raccolte per l'analisi. Ne esistono due tipi principali: parametrici e non parametrici. I dati parametri…
In biostatistica, i dati si riferiscono alle osservazioni raccolte che sono soggette ad analisi. I dati possono essere parametrici o non parametrici.
I dati non parametrici non seguono alcuna distribuzione specifica. Include dati categorici, sia nominali, come il genere, che ordinali, come le valutazioni della scala del dolore.
I dati parametrici o quantitativi presuppongono un modello di distribuzione specifico. Include dati numerici, sia continui, come il peso, che discreti, come il numero di compresse.
Le distribuzioni in biostatistica indicano la disposizione dei punti dati su un grafico.
La distribuzione normale o gaussiana è una distribuzione simmetrica in cui media, mediana e modale coincidono. Ad esempio, il peso alla nascita del bambino segue in genere una distribuzione normale.
Le distribuzioni asimmetriche hanno punti dati distribuiti in modo asimmetrico attorno alla media.
Un'inclinazione positiva indica una lunga coda verso destra, esemplificata dalle concentrazioni di metaboliti del farmaco nel sangue di un paziente.
Un'inclinazione negativa mostra una lunga coda verso sinistra, un esempio di ciò è il costo dei farmaci di marca.
View the full transcript and gain access to JoVE Core videos
Q1: What is the difference between parametric and nonparametric data in biostatistics?
Parametric data assumes a specific distribution pattern and includes numerical observations like weight or tablet counts. Nonparametric data does not follow any specific distribution and comprises categorical observations such as gender or pain scale ratings. Understanding parametric versus nonparametric data is essential for selecting appropriate statistical analysis methods.
Q2: What are examples of continuous and discrete data types?
Continuous data represents measurements that can take any value within a range, such as patient weight or drug concentration levels. Discrete data consists of countable whole numbers, like the number of tablets administered or patient count. Both are parametric data types used in biostatistical analysis.
Q3: How do nominal and ordinal data differ?
Nominal data categorizes observations without any inherent order, such as gender or drug type classifications. Ordinal data has a meaningful sequence or ranking, like pain scale ratings from mild to severe. Both are nonparametric categorical data types commonly used in biostatistics.
Q4: What is a normal distribution and why is it important in biostatistics?
Normal or Gaussian distribution is a symmetrical arrangement where mean, median, and mode coincide. Baby birth weights exemplify this pattern. Normal distribution is fundamental in biostatistics because many parametric statistical tests assume data follows this distribution pattern for valid analysis.
Q5: What does a positive skew indicate in a data distribution?
A positive skew shows data asymmetrically distributed with a long tail extending toward the right. Drug metabolite concentrations in patient blood exemplify positive skew, where most values cluster lower but some extremely high values pull the tail rightward, creating asymmetry.
Q6: How does negative skew differ from positive skew?
Negative skew has a long tail extending toward the left, indicating most data points cluster higher with some extremely low values. The cost of branded drugs demonstrates negative skew. Both skewed distributions represent asymmetrical data arrangements unlike the symmetrical normal distribution.
Q7: Why is understanding data distribution important for biostatistical analysis?
Data distribution determines which statistical methods to analyze parametric data are appropriate for analysis. Parametric tests require normally distributed data, while nonparametric tests suit skewed or non-normal distributions. Recognizing whether data follows normal, positive, or negative skew patterns ensures accurate interpretation.