Lecture 4 Psychometric Theory

0.0(0)
Studied by 0 people
call kaiCall Kai
learnLearn
examPractice Test
spaced repetitionSpaced Repetition
heart puzzleMatch
flashcardsFlashcards
GameKnowt Play
Card Sorting

1/57

encourage image

There's no tags or description

Looks like no tags are added yet.

Last updated 4:42 AM on 9/30/26
Name
Mastery
Learn
Test
Matching
Spaced
Call with Kai
Chat

No analytics yet

Send a link to your students to track their progress

58 Terms

1
New cards

What is reliability?

The consistency of measurement

2
New cards

What does reliability ask?

Can i trust the score?

3
New cards

Higher reliability =

less random error, lower SEM

4
New cards

What is the reliability coefficient range?

0-1. closer to 1 = better reliability

5
New cards

Reliability vs validity

Reliability = consistency

Validity = accuracy

A test can be reliable without being valid

6
New cards

What is random error?

Error that occurs unpredictably

  • fatigue

  • illness

  • distraction

  • phone ringing

Random error lowers reliability

7
New cards

What is systematic error

Error that occurs consistently

  • dyslexia

  • ESL

  • vision problems

  • Systematic error may not reduce reliability


8
New cards

What does test-retest reliabilty ask?

Will i get the same score later?

9
New cards

What is test-retest reliability?

Same test + Same people + Different times

10
New cards

What error source does test-retest assess?

Time sampling error

11
New cards

Why is 2 weeks the gold standard?

Short enough to reduce true change. Long enough to reduce memory effects.

12
New cards

Why not test 1 hour later?

practice effects & memory effects

13
New cards

Why not test 2 years later?

Real change may occur

14
New cards

What question does parallel forms reliability ask?

Do form A and form B produce similar scores

15
New cards

Major assumption of parallel forms

Both forms should produce the same true score

16
New cards

Why use parallel forms?

to reduce practice effects

17
New cards

What coefficient is used for parallel forms?

Coefficient of equivalence

18
New cards

What error source is assessed when forms are adminstered immediately (parallel forms)

Content sampling error

19
New cards

What error source is assessed when forms are adminsitered 2 weeks apart (parallel forms)

Content sampling + time sampling error

20
New cards

High same-day correlation but low 2 week correlation suggests…

the test may be measuring a state instead of a trait

21
New cards

What is a state?

temporary condition

mood, stress, fatigue

22
New cards

What is a trait?

Stable characteristic

ex: trait anxiety, introversion

23
New cards

What is internal consistency

Whether items “hang together”

24
New cards

Main question of internal consistency

Do the items measure the same construct?

25
New cards

What happens if a depression scale includes a pizza item?

Internal consistency decreases. The item does not represent the construct

26
New cards

What psychometric property is being assessed by Cronbach Alpha?

Internal consistency reliability

27
New cards

How is split half reliability calculated ?

Split test into 2 halves and correlate scores

28
New cards

Main weakness of split half reliability?

Reliability depends on how the test is split

29
New cards

Why can speeded test inflate split-hald reliabilty

Both halves contain unanswered items when people run out of time. The clock creates the correlation

30
New cards

Problem with speeded test and corelation

correlation may reflect the time limit rather than the construct.

31
New cards

What does KR-20 measure

internal consistency reliability

32
New cards

When is KR-20 used

Dichotomous scoring

Right/wrong

Correct/incorrect

33
New cards

Examples of KR-20 tests

Spelling tests, math tests, achievement tests

34
New cards

Can KR-20 be used on a Likert scale

No. only dichotomous measures

35
New cards

Cronbach’s alppha measures….

internal consistency

36
New cards

Can an alpha be used on likert scales?

yes

37
New cards

Can alpha be used on dichotomous tests?

yes

38
New cards

High alpha means ….

items are highly related and hang together

39
New cards

Does high alpha = unidimensionality

No

40
New cards

AP history test example demonstrates

High alpha can occur even when multiple constructs are measured

41
New cards

Why can Cronbach’s alpha be misleading in a homogeneous sample?

Everyone responds similarly

Artifically inflating alpha

42
New cards

What type of sample gives a better estimate of reliability?

heterogeneous sample

43
New cards

Homogeneous sample =

people are very similar

44
New cards

Heterogeneous sample =

Wide range of scores

45
New cards

What is the spearman brown prochecy formula used for?

Predict reliability after changing test length

46
New cards

What happens to reliability when quality items are added?

Reliability generally increases

47
New cards

Why does adding items increase reliability

Random error has less influence on the total score

48
New cards

Why cant reliability increase forever?

Diminishing returns

Items become redundant

49
New cards

What does interrater reliability assess?

Agreement among judges

50
New cards

Main question of interrater reliability

Do the judges agree?

51
New cards

Main error source of interrater reliabilty?

Interrater differences

52
New cards

When would you use Kendall’s coefficient of concordance?

When judges rank people (beauty pageant)

53
New cards

When would you use Cohen’s Kappa?

When judges classify people into categories

54
New cards

Teacher and parent disagree on a CBCL. What reliability issue is involved

Interrater reliability

55
New cards

Split half assesses what?

content sampling error

56
New cards

KR-20 / Cronbach’s alpha assess what

Content heterogeneity and internal consistency

57
New cards

What is SEM

Standard error of measurement

Estimate of error around an observed score

58
New cards

What does SEM ask?

How much wiggle room exists around the score