STP 231 - Chapter 2

0.0(0)
Studied by 0 people
call kaiCall Kai
learnLearn
examPractice Test
spaced repetitionSpaced Repetition
heart puzzleMatch
flashcardsFlashcards
GameKnowt Play
Card Sorting

1/51

encourage image

There's no tags or description

Looks like no tags are added yet.

Last updated 4:39 AM on 9/8/26
Name
Mastery
Learn
Test
Matching
Spaced
Call with Kai
Chat

No analytics yet

Send a link to your students to track their progress

52 Terms

1
New cards

Undercoverage

A type of sampling bias where some groups are excluded or rarely selected.

2
New cards

Nonresponse

A type of sampling bias when selected individuals fail or refuse to respond.

3
New cards

Self-selection

A type of sampling bias where individuals choose for themselves whether to participate.

4
New cards

Response bias

A type of sampling bias where answers are inaccurate or influenced by how the question was asked.

5
New cards

Leading wording

A cause of response bias whereby the way a question is phrased biases the response.

6
New cards

Variable types: Categorical

A variable that can be divided into categories, such as blood type.

7
New cards

Variable types: Quantitative

A variable that represents numerical amounts, such as number of doctor visits.

8
New cards

Dotplot

A graphical display that plots individual data points.

9
New cards

Histogram

A graphical display that shows the frequency of data within specified intervals.

10
New cards

Mean

The average of a set of observations calculated by summing the observations and dividing by the number of observations.

11
New cards

Median

The middle value of an ordered data set.

12
New cards

Standard deviation

A measure of the amount of variation or dispersion in a set of values.

13
New cards

Interquartile range (IQR)

The difference between the first (Q1) and third (Q3) quartiles, measuring the spread of the middle 50% of the data.

14
New cards

Outlier

An observation that lies outside the overall pattern of a distribution.

15
New cards

1.5 x IQR Rule

A method for identifying suspected outliers by calculating lower and upper fences.

16
New cards

Boxplot

A graphical display of the five-number summary: minimum, first quartile (Q1), median, third quartile (Q3), and maximum.

17
New cards

Five-number summary

A summary that includes the minimum, first quartile (Q1), median, third quartile (Q3), and maximum values.

18
New cards

Degrees of freedom

The number of values in a calculation that are free to vary, usually calculated as n-1 for sample statistics.

19
New cards

Percentiles

Values below which a certain percentage of observations fall.

20
New cards

Sampling bias

The bias that occurs when the sample selected is not representative of the population.

21
New cards

Response bias

Bias that occurs due to inaccuracies in responses due to question wording or other influences.

22
New cards

Self-selection bias

Bias that occurs when individuals decide whether to participate in a study.

23
New cards

Undercoverage bias

Bias that arises when certain groups in the population are inadequately represented in the sample.

24
New cards

Undercoverage

A type of sampling bias where some groups in the population are left out or underrepresented in the process of choosing the sample.

25
New cards

Nonresponse

A type of sampling bias that occurs when an individual chosen for the sample cannot be contacted or refuses to participate.

26
New cards

Self-selection

A type of sampling bias occurring when individuals voluntarily choose whether to join a study or sample.

27
New cards

Response bias

A type of sampling bias that occurs when respondents provide inaccurate answers due to question phrasing, interviewer behavior, or social desirability.

28
New cards

Leading wording

A cause of response bias where the phrasing of a question subtly influences respondents toward a specific answer.

29
New cards

Categorical variable

A variable that places an individual into one of several groups or categories, such as blood type or eye color.

30
New cards

Quantitative variable

A variable that takes numerical values for which arithmetic operations such as adding and averaging make sense.

31
New cards

Dotplot

A simple graphical display where each data value is shown as a point above its location on a number line.

32
New cards

Histogram

A display for quantitative data that shows the frequency or relative frequency of values falling within consecutive, equal-width bins.

33
New cards

Mean

The arithmetic average of a quantitative data set, calculated by dividing the sum of all observations by the sample size nn.

34
New cards

Median

The physical midpoint of an ordered set of quantitative data, dividing the observations into two equal halves.

35
New cards

Standard deviation

A measure of the average distance of quantitative observations from their mean.

36
New cards

Interquartile range (IQR)

The range of the middle 50%50\% of quantitative observations, calculated as IQR=Q3−Q1IQR = Q_3 - Q_1.

37
New cards

1.5 x IQR Rule

A criterion for identifying suspected outliers: values falling below Q1−1.5×IQRQ_1 - 1.5 \times IQR or above Q3+1.5×IQRQ_3 + 1.5 \times IQR.

38
New cards

Five-number summary

A summary set for quantitative data consisting of Minimum, Q1Q_1, Median, Q3Q_3, and Maximum.

39
New cards

Simple Random Sample (SRS)

A sampling design in which every group of nn individuals in the population has an equal chance of being selected as the sample.

40
New cards

Stratified random sample

A sampling method formed by dividing the population into homogenous groups called strata and taking a random sample from each stratum.

41
New cards

Cluster sample

A sampling method created by dividing the population into heterogeneous groups called clusters, then randomly selecting entire clusters to sample.

42
New cards

Systematic sample

A sampling method in which individuals are chosen at a fixed interval from an ordered population list, such as every kthk^{\text{th}} person.

43
New cards

Convenience sample

A non-random sampling method that selects individuals who are easiest to reach, leading to unrepresentative data.

44
New cards

Observational study

A study that gathers data on individuals without attempting to influence or manipulate any treatment or condition.

45
New cards

Experiment

A study that intentionally imposes a treatment on subjects to observe and measure the resulting responses.

46
New cards

Confounding variable

A variable related to both the explanatory and response variables, making it difficult to isolate the true cause of the effect.

47
New cards

Explanatory variable

A variable whose changes help explain or predict changes in the response variable.

48
New cards

Response variable

A variable that measures an outcome or result of a study.

49
New cards

Z-score

A measure of how many standard deviations an observation falls above or below the mean, calculated as z=x−μσz = \frac{x - \mu}{\sigma}.

50
New cards

Right-skewed distribution

A quantitative distribution shape where the right tail (higher values) is longer than the left tail, typically causing the mean to be greater than the median.

51
New cards

Left-skewed distribution

A quantitative distribution shape where the left tail (lower values) is longer than the right tail, typically causing the mean to be less than the median.

52
New cards

Variance

A measure of spread equal to the average squared distance of observations from their mean, denoted as s2s^2 or σ2\sigma^2.