Psychological Testing and Assessment: Chapter 5 - Reliability

0.0(0)
Studied by 0 people
call kaiCall Kai
learnLearn
examPractice Test
spaced repetitionSpaced Repetition
heart puzzleMatch
flashcardsFlashcards
GameKnowt Play
Card Sorting

1/29

flashcard set

Earn XP

Description and Tags

Flashcards covering the fundamental concepts of reliability, error variance, and different psychometric models based on lecture materials.

Last updated 12:52 AM on 8/21/26
Name
Mastery
Learn
Test
Matching
Spaced
Call with Kai
Chat

No analytics yet

Send a link to your students to track their progress

30 Terms

1
New cards

Reliability

In the field of psychometrics, this term refers to consistency in measurement, regardless of whether the results are positive or negative.

2
New cards

Reliability Coefficient

A statistic used to quantify reliability, typically ranging from 00, indicating no reliability, to 11, indicating perfect reliability.

3
New cards

Measurement Error

The inherent uncertainty associated with any measurement, consisting of both avoidable mistakes and inevitable imprecision.

4
New cards

True Score (TT)

A theoretical measurement of a quantity that would be obtained if there were no measurement error at all.

5
New cards

Construct Score

A person's standing on a theoretical variable, such as reading ability or depression, independent of any specific measurement tool.

6
New cards

Observed Score (XX)

The actual result obtained from a measurement, relating to the true score and error through the equation X=T+EX = T + E.

7
New cards

True Variance

Variance in test scores that is attributed to genuine differences between individuals.

8
New cards

Error Variance

Variance in test scores attributed to irrelevant or random sources, often identified as noise.

9
New cards

Random Error

Unpredictable fluctuations and inconsistencies in the measurement process that tend to cancel each other out in the long run.

10
New cards

Systematic Error

Error that influences test scores in a constant direction, either consistently inflating or consistently deflating results.

11
New cards

Bias

A statistical term referring to the degree to which systematic error predictably influences a measurement.

12
New cards

Item Sampling

A source of error variance that refers to differences in the specific items selected for inclusion within or between tests.

13
New cards

Carryover Effects

A situation where the act of measurement itself alters the quantity being estimated, such as practice effects or fatigue effects.

14
New cards

Test-Retest Reliability

An estimate obtained by correlating results from the same group of people on two separate administrations of the identical test.

15
New cards

Coefficient of Stability

An estimate of test-retest reliability specifically when the time interval between the two test administrations is greater than six months.

16
New cards

Parallel Forms

Different versions of a test where the means and variances of the observed scores are equal for each form.

17
New cards

Alternate Forms

Different versions of a test designed to be equivalent in content and difficulty but not meeting the strict requirements of parallel forms.

18
New cards

Internal Consistency

A measure of reliability based on the degree to which different items on a single scale relate to one another.

19
New cards

Split-Half Reliability

A method of estimating reliability by correlating two equivalent halves of a test administered a single time.

20
New cards

Spearman-Brown Formula

A mathematical tool used to estimate the internal consistency of a test that has been shortened or lengthened.

21
New cards

Coefficient Alpha (Cronbach's Alpha)

A measure of internal consistency representing the mean of all possible split-half correlations, corrected by the Spearman-Brown formula.

22
New cards

McDonald's Omega

A measure of reliability that estimates internal consistency without assuming all test item loadings (λ\lambda) are equal.

23
New cards

Inter-Scorer Reliability

The degree of agreement and consistency between two or more independent raters or judges.

24
New cards

Classical Test Theory (CTTCTT)

The most widely used psychometric model, also known as the true score model, which assumes an observed score is the sum of a true score and error.

25
New cards

Generalizability Theory

A theory suggesting that a person's test scores vary due to specific variables in the testing situation, referred to as facets.

26
New cards

Universe Score

The long-term average of a person's performance in a given testing situation, analogous to a true score in classical theory.

27
New cards

Item Response Theory (IRTIRT)

A family of psychometric methods that models the probability of a specific performance level based on a person's latent trait or ability.

28
New cards

Standard Error of Measurement (SEMSEM)

The standard deviation of a theoretically normal distribution of test scores that provides a measure of precision for an observed score.

29
New cards

Confidence Interval

A range or band of scores that is likely to contain the true score, calculated using the standard error of measurement.

30
New cards

Standard Error of the Difference (sigmadiff\\sigma_{diff})

A statistical measure used to determine if the difference between two scores is large enough to be considered statistically significant.