Finishing Up CTT & Threats to Psychometric Quality 1 (9/23/26)

0.0(0)
Studied by 0 people
call kaiCall Kai
learnLearn
examPractice Test
spaced repetitionSpaced Repetition
heart puzzleMatch
flashcardsFlashcards
GameKnowt Play
Card Sorting

1/35

encourage image

There's no tags or description

Looks like no tags are added yet.

Last updated 3:19 PM on 9/23/26
Name
Mastery
Learn
Test
Matching
Spaced
Call with Kai
Chat

No analytics yet

Send a link to your students to track their progress

36 Terms

1
New cards

Non-cognitive item analysis based on CTT

Reliability, item total correlation, item variance (SD), item mean, item-criterion correlation, non response rate

2
New cards

Cognitive item analysis based on CTT (only for binary data)

Reliability, item difficulty, item discrimination index, distribution of responses, point biserial correlation (or biserial correlation)

3
New cards

Item-scale correlation (item-total correlation) for non-cognitive analysis

Correlation between score in the item and total score of the scale (we want high correlations)

4
New cards

Item variance for non-cognitive analysis

Not small

5
New cards

Item means for non-cognitive analysis

Not too high and too low

6
New cards

Reliability for non-cognitive analysis

Acceptable: more than 0.7

Respectable: more than 0.8

Good: more than 0.9

7
New cards

Non-response rate for non-cognitive analysis

0 or very small

8
New cards

Item-criterion correlations for non-cognitive analysis

No clear standard but at least more than .2

9
New cards

Personality tests

Item variance should not be too small and item means should not be too low or too high

10
New cards

Survey items

In some cases, small variance, too high/low mean are okay

11
New cards

Negative item discrimination

People who don’t have the appropriate knowledge were able to correctly answer the question, but people who did have the appropriate knowledge were unable to correctly answer the question

12
New cards

The reliability that should be used for cognitive items (binary data)

KR-20 because the data is dichotomous

<p>KR-20 because the data is dichotomous</p>
13
New cards
<p>What does m stand for in the KR-20 equation?</p>

What does m stand for in the KR-20 equation?

number of items

14
New cards
<p>What does pj stand for in the KR-20 equation?</p>

What does pj stand for in the KR-20 equation?

Proportion of people who correctly answer the item j

15
New cards
<p>What does qj stand for in the KR-20 equation?</p>

What does qj stand for in the KR-20 equation?

Proportion of people who did not correctly answer the item j

16
New cards
<p>What does <span>σ2x</span> stand for in the KR-20 equation?</p>

What does σ2x stand for in the KR-20 equation?

Variance of the test scores

17
New cards

Item difficulty for cognitive items (binary data)

The number of ppl who get a particular item correct. Ex) 84 out of 100 people correctly answered = 84%. Too low (0 - 0.2) or too high (0.9 - 1) is not good. Around 0.5 would be desiriable

18
New cards

Item discrimination for cognitive items (binary data)

Difference in item difficulty between high performers and low performers. 30 or more than 30 would be better but more than 20 is acceptable

19
New cards

Steps for calculating item discrimination for cognitive items (binary data)

  1. Create high & low performer groups (any % between 25% and 35%) using scores of the test

  2. Calculate the % of people who correctly answered of each group

  3. Calculate “high performer %” and “low performer %”


20
New cards

Point biserial correlation for cognitive items (binary data)

Correlation between performance on the item and performance on the total test (technically performance on the total test excluding the item)

<p>Correlation between performance on the item and performance on the total test (technically performance on the total test excluding the item)</p>
21
New cards

Standard for point biserial correlation

Depends on the type of construct but at least should be more than 0.2. For distractors, correlation between total scores and choice of the distractor (select 1 not select 0)

22
New cards

Distractors

Wrong options in tests

23
New cards

Good item distribution

Like an inverted triangle

<p>Like an inverted triangle</p>
24
New cards

Bad item distribution

Like a triangle

<p>Like a triangle</p>
25
New cards

Good distractor

Point biserial correlations for the two distractors are negative; the two distractors worked well.

<p>Point biserial correlations for the two distractors are negative; the two distractors worked well.</p>
26
New cards

Bad distractor

Point biserial correlation for the distractor c) worked well but the distractor b) did not work well. As the biserial correlation for distractor b) is positive and relatively large, higher ability people tend to choose choice b)

<p>Point biserial correlation for the distractor c) worked well but the distractor b) did not work well. As the biserial correlation for distractor b) is positive and relatively large, higher ability people tend to choose choice b)</p>
27
New cards

Response bias

Tendencies for participants to respond inaccurately or falsely to questions. Includes acquiesence, extremity & modesty, social desirability, malingering, random/careless responding, & guessing

28
New cards

Test bias

Systematically obscures differences between groups (and between people from different groups). Includes construct bias and predictive bias

29
New cards

Acquiesence

Consistently endorsing items w/out much regard for their content (“yea-saying”) and consistently rejecting items w/out much regard for their content (“nay-saying”). Caused by cultures and other factors such as personality traits. Would increase internal correlations & correlations between tests. May harm reliability & validity

30
New cards

Extremity & modestry

Tendency to overuse “extreme” response options and underuse “extreme” response options, respectively. Implications similar to those for acquiescence bias. Creates ambiguity in who truly has high (vs low) levels of the construct being measured

31
New cards

Social desirability

Tendency to respond in a way that is social appealing, regardless of one’s true psychological characteristics. Technically this is not “fake good” but there is overlap

32
New cards

Subdimensions of social desiarbility

Impression management and self-deception

33
New cards

Item transparency

How easily test takers understand what the item is assessing. Too high and low of this is not good. Can artificially inflate (or deflate) test scores. Can’t tell who truly has a high level of the construct. Can harm predication/decisions based on the test result.

34
New cards

Malingering

Tendency to respond (intentionally) in a way that suggests psychological problems, regardless of one’s true psychological characteristics; “faking bad.” Most likely to occur when being perceived as having problems would benefit the respondent (e.g., disability evaluation, worker’s compensation claims)

35
New cards

Random/careless responding

Answering questions in random, semi-random or simply careless fashion. Occurrence estimates 1-10% of respondents, though varies by context and scope of randomness. Most likely when not motivated to be thoughtful/careful (anonymity, feelings of coercion). Score isn’t meaningful/interpretable. In research context, night attenuate or spuriously inflate effects

36
New cards

Guessing

For tests with “correct” answers. When respondent motivated to select correct answers. Eg) cognitive ability test