A. 3. Common Psychometric (Measurement) Theories

0.0(0)
Studied by 0 people
call kaiCall Kai
Locked
learnLearn
examPractice Test
spaced repetitionSpaced Repetition
heart puzzleMatch
flashcardsFlashcards
GameKnowt Play
Card Sorting

1/37

encourage image

There's no tags or description

Looks like no tags are added yet.

Last updated 11:44 PM on 8/8/26
Name
Mastery
Learn
Test
Matching
Spaced
Call with Kai
Chat

No analytics yet

Send a link to your students to track their progress

38 Terms

1
New cards

Homogeneous Test Items

Test items designed to measure one factor or construct that indicates high internal consistency

2
New cards

Heterogeneous Test Items

Test designed to measure multiple factors or construct that indicates low internal consistency

3
New cards

Dynamic Characteristics

A trait, state, or ability presumed to be ever-changing because of situational factors (e.g mood)

4
New cards

Static Characteristics

A trait, state, or ability presumed to be relatively unchanging (e.g pesonality)

5
New cards

Restricted range

Occurs when scores are too similar (e.g from a homogenous group). This results in little to no variation which lessens the reliability coefficient

6
New cards

Inflated range

Occurs when scores are too different (e.g from a heterogenous group). This means the range is too wide which inflates the reliability coefficient

7
New cards

Power Test

A test where the time limit is long enough or even non-existent to allow test takers to attempt all items. The items increase in difficult the higher they go

8
New cards

Speed Test

Test that contains items of uniform level of difficulty but with a time limit. The goal is to complete as many items as fast as possible

9
New cards

Norm-Referenced Test

Test that compares the score of an examinee to the scores of other people or a group of people

10
New cards

Criterion-Referenced Test

Test that compares the performance of an examinee to a predetermined standard or competency level, usually through certain cut-off scores

11
New cards

Classical Test Theory

Theory that is referred to as true score (or classical) model of measurement and believes that everyone has a "true score" on a test

12
New cards

True Score

A value that, according to CTT, genuinely reflects an individual’s ability (or trait) level as measured by a particular test

13
New cards

Formula for the CTT

X = T + E

14
New cards

Domain Sampling Theory

Core concept in the CTT that states that any test is a subset of items drawn from an infinite universe of potential questions measuring a single trait. Assumes that all items come from the same single definition and tests under this theory have high internal consistency

15
New cards

Domain of Behavior

The universe of items that could measure a certain behavior

16
New cards

Generalizability Theory

Theory that believes that error is not a single pool (Unlike CTT). Rather, it believes that error has different sources that needs to be taken into account

17
New cards

Terminologies in the Generalizability Theory

Universe

Facets

Universe Score

Generalizability Study

Decision Study

18
New cards

Universe

In Generalizability Theory, it it the testing environment. It describes the context/details of the particular test situation.

19
New cards

Facets

In Generalizability Theory, it is the possible sources of error

20
New cards

Universe Score

In Generalizability Theory, it is equivalent to the true score. The average of an individual’s scores across all possible conditions in a universe of generalization

21
New cards

Generalizability Study (G-Study)

In Generalizability Theory, it examines how generalizable scores from a particular test are if the test is administered in different situations

22
New cards

Decision Study

In Generalizability Theory, it is when developers use the G-Study to make decisions on how to change or improve a test

23
New cards

Item-Response Theory (Latent-Trait Theory)

The Modern Test Theory. It models how individuals interact or respond to a particular test question

24
New cards

Latent Trait

The unobservable Trait

25
New cards

Manifestation Trait

The observable Trait

26
New cards

Dichotomous Test Items

Test items with only two possible responses

27
New cards

Three Main Models of IRT (Parameters) for dichotomous items

Parameter A: Discrimination

Parameter B: Difficulty

Parameter C: Pseudoguessing/Guessing

28
New cards

Polytomous Test Items

Test items with more than two responses

29
New cards

Parameter A: Discrimination

In Item-Response Theory, it measures how an item differentiates among people with higher or lower levels of the trait/ability that is being measured

30
New cards

Parameter B: Difficulty

In Item-Response Theory, it is the most basic parameter that tries to evaluate how high the trait/ability is in a certain test taker

31
New cards

Parameter C: Pseudoguessing/Guessing

In Item-Response Theory, it is the parameter that accounts how likely a test taker will guess the correct item

32
New cards

Formula for probability of guessing

Number of correct options / Total number of choices

33
New cards

Standard Error of Measurement/Score

Provides an estimate of the amount of error inherent in an observed score of measurement

34
New cards

Confidence Interval

A range or band of the test scores that is likely to contain the true score

35
New cards

68% confidence interval (±1 SEM)

This means there's a 68% chance that the individual's true score falls within that range

36
New cards

95% confidence interval (±1.96 SEM often approximated as ±2 SEM)

This means there’s a 95% chance that the individual’s true score falls within that range

37
New cards

Standard Error of the Difference

Statistical measure that aids in determining whether the difference between two test scores is due to true variance or error variance

38
New cards

Standard Error of Estimate

Statistical measure that predicts how much the actual data differs from a prediction line in regression