Psychometric Reliability Concepts and Models

0.0(0)
Studied by 0 people
call kaiCall Kai
learnLearn
examPractice Test
spaced repetitionSpaced Repetition
heart puzzleMatch
flashcardsFlashcards
GameKnowt Play
Card Sorting

1/24

flashcard set

Earn XP

Description and Tags

Fill-in-the-blank practice flashcards covering classical test theory, reliability models, measurement error, internal consistency, inter-scorer agreement, and psychometric evaluation principles.

Last updated 3:43 AM on 9/24/26
Name
Mastery
Learn
Test
Matching
Spaced
Call with Kai
Chat

No analytics yet

Send a link to your students to track their progress

25 Terms

1
New cards

In 1904, Charles Spearman published a foundational paper titled "__________".

The Proof and Measurement of Association between Two Things

2
New cards

Reliability broadly refers to the __________ and consistency of test scores over time or across different items.

dependability

3
New cards

For most basic research purposes, reliability estimates in the range of __________ and __________ are considered good enough.

.70.70; .80.80

4
New cards

In Classical Test Theory, the fundamental formula expressing an observed score is __________.

X=T+EX = T + E

5
New cards

In Classical Test Theory, XX represents the observed score, TT represents the true score, and EE represents __________.

Error

6
New cards

Classical Test Theory assumes that measurement errors are random and that their distribution is __________.

bell-shaped

7
New cards

The standard deviation of measurement errors is used as the basic measure of error and is known as the __________.

Standard Error of Measurement

8
New cards

If an examinee achieves an observed score of 7070 on a test with an SEM of 44, their possible true score falls between __________ and __________.

6666; 7474

9
New cards

Variation among items within or between tests due to item sampling or content sampling is a major source of error variance in __________.

TEST CONSTRUCTION

10
New cards

Test-retest reliability is also referred to as __________ reliability.

time-sampling

11
New cards

When the time interval between test-retest administrations exceeds 66 months, the resulting coefficient is known as the coefficient of __________.

stability

12
New cards

In parallel-form reliability, two forms are constructed such that the __________ and variances of observed test scores are equal.

means

13
New cards

To correct for half-length when correlating split halves of a single test, researchers use the __________ formula.

Spearman-Brown

14
New cards

The Kuder-Richardson 20 (KR-20) formula is designed to calculate reliability for tests whose items are __________, meaning scored 00 or 11.

dichotomous

15
New cards
<p>According to the common interpretation table for KR-20, a value between $$0.90$$ and $$1.00$$ indicates __________.</p>

According to the common interpretation table for KR-20, a value between 0.900.90 and 1.001.00 indicates __________.

Very high consistency

16
New cards

Cronbach developed coefficient alpha as a general estimate of internal consistency for tests whose items are not scored as __________.

00 or 11

17
New cards
<p>As illustrated in the diagram, a High Cronbach's Alpha ($$\alpha = .90$$) means items are highly correlated, move together, and __________.</p>

As illustrated in the diagram, a High Cronbach's Alpha (α=.90\alpha = .90) means items are highly correlated, move together, and __________.

Measure the same thing

18
New cards
<p>As shown in the diagram, a Low Cronbach's Alpha ($$\alpha = .40$$) means items are weakly correlated, don't move together, and __________.</p>

As shown in the diagram, a Low Cronbach's Alpha (α=.40\alpha = .40) means items are weakly correlated, don't move together, and __________.

May be measuring different things

19
New cards

When assessing inter-scorer agreement, __________ Kappa is used for 22 raters, whereas Fleiss' Kappa is used for 33 or more raters.

Cohen's

20
New cards
<p>According to the inter-scorer agreement table, values greater than $$0.75$$ indicate __________ Agreement.</p>

According to the inter-scorer agreement table, values greater than 0.750.75 indicate __________ Agreement.

Excellent

21
New cards

Tests that contain items measuring a single trait or factor are described as __________.

homogenous

22
New cards

Traits or states that are presumed to be ever-changing as a function of situational experiences are called __________ characteristics.

dynamic

23
New cards

A test that contains relatively easy items but strict, short time limits is known as a __________ test.

speed

24
New cards

Item Response Theory (IRT) uses computer adaptivity to tailor questions based on item __________.

difficulty

25
New cards

In clinical settings where important decisions are made about an individual's future, evaluators should seek tests with a reliability greater than __________.

.95.95