Comprehensive Study Guide for Psychological Statistics
Prologue: The Nature of Statistics
Personal Perspective of Garett C. Foster, Ph.D.: Statistics can be perceived as difficult, esoteric, and overwhelming. Initially, Foster detested the subject during his introductory course. However, once the underlying logic was understood, it became a powerful tool for objective observation.
Finding the Signal in the Noise: The primary function of statistics is to distinguish meaningful patterns ("signals") from random chance ("noise").
Objective Filtering: Statistics provides an objective lens to filter the constant onslaught of information in contemporary life, ensuring sanity and clarity through logic.
Pedagogical Goal: This guide focuses on formulae not merely as numerical manipulations, but as a system of interconnected topics and methods applicable to behavioral sciences and everyday decisions.
Chapter 1: Introduction and Terminology
Defining Statistics: In a broad sense, statistics refers to techniques and procedures for analyzing, interpreting, displaying, and making decisions based on data. It is the universal language of science, allowing researchers across fields to articulate findings accurately.
Statistical Misconceptions: Statistics is not merely a "math class." While it utilizes math as a tool, it is better defined as a method for organizing and communicating information objectively.
Statistical Facts and Figures: Examples include:
The largest earthquake measured on the Richter scale.
Men are at least more likely than women to commit murder.
in every South Africans is HIV positive.
of all statistics are reportedly made up on the spot (a ironic illustration of statistical misuse).
Flaws in Interpretation: Numbers may be accurate, but the logic applied to them can be flawed:
History Effect: An ice cream ad precedes a sales increase. However, if this occurs in June, the increase is likely due to the season (time), not the ad.
Third-Variable Problem: A correlation exists between the number of churches and crime rates. The third variable is population size (larger cities have more of both).
Missing Context: A statement that interracial marriages increased by in years lacks the base rate. If the rate went from to , the social implication is very different.
Logic for Daily Life: Studying statistics empowers individuals to evaluate claims like " out of dentists recommend Dentine" or "Women make cents to every dollar a man makes" to ensure intelligent consumer behavior.
Types of Variables and Levels of Measurement
Variables: Characteristics or features of the item being studied (e.g., stress levels, health).
Independent Variable (IV): The variable manipulated by the experimenter (e.g., dosage of a drug).
Dependent Variable (DV): The outcome variable measured to see the effect of the IV (e.g., relief from symptoms).
Levels: The number of conditions in an IV (e.g., a study with a placebo and an active drug has levels).
Qualitative (Categorical) Variables: Variables expressing attributes without numerical ordering (e.g., hair color, religion, favorite movie).
Quantitative Variables: Variables measured in numbers (e.g., height, weight, test scores).
Discrete Variables: Variables with scores at discrete points (e.g., number of children; you cannot have children).
Continuous Variables: Variables on a continuous scale where any value is possible (e.g., response time measured as ).
Levels of Measurement (S.S. Stevens' Scales):
Nominal: Categorization or naming only (e.g., Gender, Handedness). No ordering is implied.
Ordinal: Categories with a specific order (e.g., satisfaction ratings like "very dissatisfied" to "very satisfied"). Intervals between values are not necessarily equal.
Interval: Numerical scales where intervals are consistent (e.g., Fahrenheit). No true zero point (zero degrees does not mean absence of temperature).
Ratio: Most informative scale. Has a true zero point representing the absence of the quantity (e.g., money, weight, Kelvin scale). Allows for ratio statements (e.g., "twice as much").
Sampling and Research Design
Population: The entire collection of people sharing a characteristic of interest.
Sample: A small subset of the population used to draw inferences.
Sampling Bias: Occurs when a sample over-represents certain segments, making results non-generalizable.
Sampling Error: The natural, expected discrepancy between a sample statistic and a population parameter.
Sampling Strategies:
Simple Random Sampling (SRS): Every member has an equal and independent chance of selection.
Stratified Sampling: Identifying groups ("strata") and sampling proportionally to ensure the sample matches the population's demographics (e.g., day students and night students).
Convenience Sampling: Using easily accessible subjects; often biased and used only for preliminary data.
Types of Research Designs:
Experimental: Uses random assignment and manipulation of the IV to determine causality.
Quasi-Experimental: Manipulates the IV but lacks random assignment (often due to ethics or logistics, e.g., using existing classrooms).
Non-Experimental (Correlational): Observes variables as they naturally occur. Useful for prediction but cannot establish causality.
Mathematical Notation
Summation Notation (): The Greek letter sigma indicates the sum of a set of numbers.
General Formula: means sum the variable from the first to the -th observation.
Order of Operations:
: Square each number first, then sum the squares.
: Sum all numbers first, then square the total.
Sum of Cross Products: indicates multiplying each person's score by their score, then summing those products.
Chapter 2: Visualizing Data
Frequency Tables: Derived from the count of observations in each category. Relative frequency is the proportion (e.g., ).
Qualitative Visualization:
Pie Charts: Effective for small numbers of categories. The area of the slice is proportional to the percentage of the whole.
Bar Charts: Used for frequencies. The Y-axis usually shows count, X-axis shows the category. Fanciness (e.g., effects) should be avoided as they create a "lie factor" (distortion).
Quantitative Visualization:
Stem and Leaf Displays: Best for small datasets. Stems represent higher digits (e.g., ), leaves represent lower digits (e.g., ). Useful for seeing distribution shape while retaining raw data.
Histograms: Used for large datasets. Scores are grouped into "class intervals" or "bins." The height of the bar represents frequency.
Frequency Polygons: Points are placed at the height of the frequency for each interval and connected by lines. Ideal for comparing distributions by overlaying them.
Box Plots: Primarily used to identify outliers and compare spreads. Components include:
Hinges: and percentiles (the box).
Median: Line inside the box ( percentile).
Whiskers: Lines extending to the adjacent values.
Outliers: Marked as circles ("outside") or asterisks ("far out").
Distribution Shapes:
Symmetrical: The left and right halves are mirror images (e.g., Normal Distribution/Bell Curve).
Bimodal: Two distinct peaks.
Skew: Asymmetry where one tail is longer.
Positive (Right) Skew: Tail points toward positive/higher numbers.
Negative (Left) Skew: Tail points toward negative/lower numbers.
Chapter 3: Central Tendency and Spread
Measures of Central Tendency:
Mean ( or ): The arithmetic average. It is the balance point of the distribution. It minimizes the sum of squared deviations.
Median: The exact midpoint ( percentile). Minimizes the sum of absolute deviations. Resistant to outliers.
Mode: The most frequent score. The only measure usable for nominal data.
Relative Positions in Skewed Distributions:
In a Positive Skew, .
In a Negative Skew, .
Measures of Spread (Variability):
Range: Highest score minus lowest score. Extremely sensitive to outliers.
Interquartile Range (IQR): The middle of data ( percentiles).
Sum of Squares (SS): . The total squared deviation from the mean.
Variance ( or ): The average squared deviation.
Population Variance:
Sample Variance: . Dividing by (degrees of freedom/df) corrects for the tendency of samples to underestimate population variability.
Standard Deviation ( or ): The square root of the variance. It returns the measure to the original units of the data.
Chapter 4: z-scores and the Standard Normal Distribution
Normal Distribution Properties: Symmetrical, bell-shaped, area under curve . Mean, median, and mode are equal.
The Empirical Rule:
of data falls within .
of data falls within .
of data falls within .
Standard Normal Distribution: A normal distribution with and .
Standardized Scores (z-scores): Convert raw scores into units of standard deviation to determine relative location.
Interpretation: Sign () indicates direction; Magnitude indicates distance. A score with might be considered relative far from the mean.
Standardizing Scales: z-scores can be transformed into any scale (). E.g., converting a z-score of to an IQ score () yields an IQ of .
Chapter 5: Probability
Definition: .
Probability and the Normal Curve: The area under the curve represents the probability of a random observation falling within a specific range.
z-Table Usage: Allows for finding the area in the "body" (larger part) or "tail" (smaller part) relative to a z-score.
Symmetry and Complements: Because the curve is symmetrical, the area of the left tail is the same as the right tail for identical z-magnitudes. Total area , so .
Chapter 6: Sampling Distributions
Definition: A theoretical distribution of a statistic (e.g., sample means) calculated from all possible samples of a specific size .
Central Limit Theorem (CLT): States that for samples of size , the sampling distribution of the mean will:
Have a mean .
Have a standard deviation (Standard Error) .
Approach normality as increases ( is the general threshold).
Standard Error (): Quantifies sampling error. As sample size () increases, decreases ().
Law of Large Numbers: As sample size increases, the sample mean becomes a more accurate estimate of the population mean.
Calculate z for a Mean: , used to find the probability of obtaining a specific sample mean.
Chapter 7: Introduction to Hypothesis Testing
The Logic of Fisher: Uses the probability of an outcome given a specific state of the world to cast doubt on that state. (Example: James Bond correctly identifying martinis times; probability of guessing is .)
The Null Hypothesis (): The baseline assumption that there is no effect, no difference, or that any apparent effect is due to chance (e.g., ).
The Alternative Hypothesis (): The research hypothesis; states that there is a significant difference or effect.
Directional (One-tailed): Specifies direction (e.g., ).
Non-directional (Two-tailed): Specifies any difference (e.g., ).
Significance Level (): The threshold for rejecting , typically set at . It is the probability of a Type I error.
Critical Values ( or ): The z-scores that bound the "rejection region."
p-value: The actual probability of the data given . If , reject .
Errors in Testing:
Type I Error: Rejecting a true (False Positive).
Type II Error (): Failing to reject a false (False Negative).
Power (): The ability to correctly reject a false .
Effect Size (Cohen's ): Measures practical significance. .
Chapter 8-10: t-tests and Interval Estimation
The t-statistic: Used when the population standard deviation () is unknown and must be estimated using the sample standard deviation ().
Degrees of Freedom (df): For t-tests, . As , the t-distribution becomes the normal distribution.
Confidence Intervals (CI): A range of plausible values for the population mean.
Hypothesis Testing via CI: If the CI does not bracket the null value, the result is statistically significant.
Repeated Measures (Dependent Samples) t-test: Analyzes change within the same subjects over time or matched pairs.
Difference Scores (): calculated as .
Independent Samples t-test: Compares means of two unrelated groups (e.g., Treatment vs. Control).
Pooled Variance (): Weighted average of variances from both groups.
Degrees of Freedom: .
Chapter 11-13: ANOVA, Correlation, and Regression
ANOVA (Analysis of Variance): Tests for differences between three or more group means without inflating Type I Error.
F-statistic: Ratio of between-groups variance (systematic) to within-groups variance (random error).
Null Hypothesis:
Effect Size (): Variance explained. .
Post Hoc Tests: (e.g., Tukey's HSD, Bonferroni) Conducted after a significant ANOVA to find which specific means differ.
Correlation (Pearson's ): Measures the strength and direction of the linear relation between two continuous variables.
Range: to . = No relation.
Covariance: How two variables vary together. .
Correlation vs. Causation: Correlation does not prove causation due to potential lurking/confound variables.
Linear Regression: Predicting a Y-value from an X-value using the Line of Best Fit ().
Slope (): Change in Y for every -unit change in X.
Standard Error of the Estimate: Average distance between predicted and actual scores.
Chapter 14: Chi-Square ()
Non-Parametric Test: Does not estimate population parameters; uses frequency of nominal categories.
Goodness-of-Fit: Measures how well one variable's observed frequencies match expected frequencies (usually equal distribution).
Test for Independence: Uses a contingency table to see if two categorical variables are related.
Formula: .
Effect Size (Cramer's V): A correlation coefficient for categorical data. .