Psychological Testing & Assessment – Comprehensive Bullet-Point Notes

Psychological Testing and Assessment: Key Definitions

  • Psychological Testing

    • Process of measuring psychology-related variables via devices or procedures designed to obtain a sample of behaviour.

    • Objective: obtain a gauge (usually numerical).

    • May be individual or group; yields scores.

  • Psychological Assessment

    • Gathering & integrating data for a psychological evaluation using multiple tools.

    • Objective: answer referral questions, solve problems, or aid decision-making.

    • Typically individualized; evaluator is central; requires educated selection and interpretive skill; outcome = problem-solving recommendations.

Forms of Assessment

  • Collaborative: assessor + assessee are partners from intake to feedback.

  • Therapeutic: assessment itself used for self-discovery and change.

  • Dynamic: evaluation → intervention → re-evaluation; focuses on learning potential during assessment.

Major Assessment Tools

  • Tests

    • Psychological test: measures intelligence, personality, aptitude, interests, attitudes, values.

    • Format: plan/structure, item layout, time limits, administration mode (paper, computer, CAT).

    • Ways tests differ:

    • Administration (one-on-one vs. group, active examiner vs. none).

    • Scoring (hand, self, machine); scores may be raw, cut score\text{cut\ score}, etc.

    • Psychometric soundness (reliability, validity, utility).

  • Interview: reciprocal information exchange; observe verbal & non-verbal cues (eye contact, voice pitch, dress, pauses). Variants include panel interviews.

  • Portfolio: collection of work samples (paper, audio, video, digital).

  • Case-History Data: archival records giving past/current adjustment, pre-morbid neuropsychological status, academic/behavioural standing.

  • Behavioural Observation

    • Naturalistic or laboratory; quantitative & qualitative recording.

  • Role-Play Tests: simulate situations; evaluate expressed thoughts/behaviours.

  • Computer-Based Tools

    • Local Processing, Central Processing, Teleprocessing.

    • Score Reports: simple, extended, interpretive (descriptive, screening, consultative, integrative).

    • CAPA, CAT; pros (speed, reach, cost, “green”); cons (client integrity, mere testing vs. assessment).

  • Other: video scenarios, biofeedback, thermometers, etc.

Stakeholders

  • Test Developer: conceives, prepares, pilots, publishes.

  • Test User: selects, administers, scores, interprets.

  • Test Taker: subject of assessment; variability in anxiety, cooperation, pain, coaching, social desirability.

    • Psychological Autopsy: reconstruction of deceased’s profile from records/interviews.

Typical Settings & Purposes

  • Education (achievement, diagnostic, informal evaluation).

  • Clinical (screen/diagnose behavioural problems).

  • Counselling (adjustment, productivity).

  • Geriatric (quality of life variables).

  • Business/Military (selection, placement).

  • Credentialing, Government.

Assessment of Individuals with Disabilities

  • Accommodation: adapt test/procedure (e.g., Braille, extended time).

  • Alternate Assessment: non-standard evaluation procedures.

  • Must consider capabilities of assessee & assessor, purpose, meaning of scores.

Reference Sources

  • Test catalogues, manuals, reference volumes, journal reviews, online data bases.

Taxonomy of Tests

  • Administration: Individual vs. Group.

  • Ability Tests: Achievement, Aptitude/Prognostic, Intelligence.

  • Personality Tests: Objective/Structured vs. Projective/Unstructured.

  • Interest Tests.

Historical Milestones

  • Ancient China civil-service exams; Greco-Roman humours.

  • 19th c.: Galton (anthropometry, individual differences); Wundt (experimental lab); Cattell (mental tests).

  • 20th c.: Binet–Simon (1905) intelligence scale; WWI Army Alpha/Beta; Wechsler scales; Woodworth Psychoneurotic Inventory; Projective era (Rorschach).

Culture & Assessment Issues

  • Verbal language & translation nuances.

  • Non-verbal behaviours mean different things (eye contact, test timing).

  • Standards of evaluation vary; need culturally appropriate norms.

  • Group membership & fairness (height requirements, affirmative action).

Legal & Ethical Framework

  • Minimum competency testing; Truth-in-testing laws.

  • Expert testimony standards: Frye → Daubert.

  • Test-user qualification levels A/B/C.

  • Test taker rights: informed consent, feedback, privacy/confidentiality, least-stigmatizing label.

Statistics Refresher

  • Measurement & Scales

    • Nominal, Ordinal, Interval, Ratio; properties: Magnitude, Equal Intervals, Absolute Zero.

  • Descriptive Stats

    • Central Tendency: Mean Xˉ=X/n\bar X=\sum X/n, Median, Mode.

    • Variability: Range, IQR=Q<em>3Q</em>1\text{IQR}=Q<em>3-Q</em>1, Variance s2s^2, SD ss.

  • Shapes: Skew (+/–), Kurtosis (platy, meso, lepto).

  • Normal Curve: symmetrical; area under curve interpretable; standard scores (z, T, SAT, GRE).

    • z=Xμσz=\dfrac{X-\mu}{\sigma}, T=10z+50T=10z+50.

  • Correlation

    • Pearson rr, Spearman ρ\rho, biserial, point-biserial, phi, tetrachoric.

    • r2r^2 = coefficient of determination.

  • Regression

    • Simple: Y=a+bXY'=a+bX; SEE; Multiple Regression.

    • Residuals, shrinkage, cross-validation.

Core Assumptions in Testing

  1. Traits & states exist.

  2. They can be quantified.

  3. Test behaviour predicts non-test behaviour.

  4. Tests have strengths & weaknesses.

  5. Measurement includes error (Classical Test Theory X=T+EX=T+E).

  6. Testing can be fair/unbiased.

  7. Testing benefits society.

Characteristics of a Good Test

  • Reliability (consistency, low error variance).

  • Validity (content, criterion-related, construct).

  • Clear instructions & easy administration.

  • Actionable, useful results.

Norms & Standard Errors

  • Norm-Referenced vs. Criterion-Referenced evaluation.

  • Sampling methods: stratified, purposive, incidental.

  • Types of norms: age, grade, developmental, national, subgroup, local, fixed-reference.

  • Standard Errors:

    • SEM=σ1rSEM = \sigma\sqrt{1-r}

    • SEESEE (prediction), SE<em>mean=σ/nSE<em>{mean}= \sigma/\sqrt{n}, SE</em>diffSE</em>{diff}.

Reliability Methods

  • Test–Retest (coefficient of stability).

  • Alternate/Parallel Forms (coefficient of equivalence).

  • Internal Consistency: Split-Half + Spearman-Brown r<em>SB=2r</em>hh1+rhhr<em>{SB}=\dfrac{2r</em>{hh}}{1+r_{hh}}, KR-20/KR-21, Cronbach α\alpha.

  • Inter-scorer reliability (Cohen/Fleiss κ\kappa).

  • Factors: homogeneity, dynamic traits, speed vs. power, restriction of range.

  • Item Response Theory (IRT), Generalizability Theory.

Validity Evidence

  • Content: expert judgement, Lawshe CVR=ne(N/2)N/2CVR=\dfrac{n_e-(N/2)}{N/2}.

  • Criterion-Related: Concurrent vs. Predictive; validity coefficient, expectancy tables, Taylor-Russell, Naylor-Shine; hit/miss, false +/–.

  • Construct: homogeneity, developmental changes, contrasted groups, convergent & discriminant (multitrait-multimethod), factor analysis.

  • Face validity ≠ technical validity.

  • Bias vs. Fairness; rating errors (leniency, severity, halo), DIF analysis.

Utility & Decision Making

  • Utility = benefits – costs; Brogden–Cronbach–Gleser formula; productivity gain.

  • Cut-scores: relative (norm-referenced) vs. fixed (criterion). Methods: Angoff, Bookmark, known-groups, IRT item-mapping.

Test Development Cycle

  1. Conceptualization: define construct, target population, purpose.

  2. Construction: scaling (nominal/ordinal/interval/ratio; Likert, Guttman), item writing (selected- vs. constructed-response), CAT.

  3. Tryout: pilot sample; examine item difficulty pp; item discrimination dd, item-validity & reliability indices.

  4. Item Analysis: ICCs, distractor analysis, qualitative reviews (think-aloud, sensitivity panels).

  5. Revision: cross-validation, co-norming, anchoring protocols; build item banks, address DIF.

Intelligence: Concepts & Theories

  • Interactionism: genes × environment.

  • Factor-Analytic Theories: Spearman g & s; Thurstone 7 primaries; Cattell Gf/Gc; Horn (vulnerable vs. maintained); Guilford SOI; Carroll 3-stratum; CHC model.

  • Information-Processing: Luria (simultaneous vs. successive); PASS.

  • Multiple Intelligences: Gardner.

  • Triarchic: Sternberg (analytical, creative, practical).

  • Developmental: Piaget (sensorimotor → formal operations).

  • Issues: nature vs. nurture, stability, Flynn effect, culture loading, personality correlates, gender, family environment.

Intelligence Tests

  • Individual

    • Stanford–Binet-5: routing tests, adaptive basal/ceiling, verbal & non-verbal, IQDeviationIQ_{Deviation} mean 100 SD 16.

    • Wechsler Series: WAIS-IV, WISC-IV, WPPSI-III; index scores (VCI, PRI, WMI, PSI); short form (WASI).

    • Kaufman (KABC-II, KAIT, KBIT), Woodcock–Johnson-III.

  • Group

    • Army Alpha/Beta, ASVAB, SAT/ACT, GRE, LSAT, MAT.

    • Non-verbal: Raven Progressive Matrices, Culture-Fair Test.

  • Special Populations

    • Infant: Brazelton BNAS, Bayley-III, Cattell CIIS, Gesell.

    • Disabilities: Columbia Mental Maturity, Peabody Picture Vocabulary, Leiter-R.

    • Learning Disabilities: ITPA-3, Woodcock–Johnson; RTI framework.

    • Visual-Motor: Bender-Gestalt, Benton VRT.

    • Creativity: Torrance Tests (fluency, originality, flexibility, elaboration).

End-to-End Workflow Summary

  1. Identify need & context (ethical/legal, cultural, practical).

  2. Choose/develop appropriate test with sound psychometrics.

  3. Administer with accommodations & rapport; collect protocols.

  4. Score (manual, computer, CAT); compute standard/percentile/z/T.

  5. Interpret using norms, reliability, validity, utility; integrate with other data (interviews, history).

  6. Communicate findings respecting rights (feedback, confidentiality, least-stigmatizing labels).

  7. Reassess & revise instruments; monitor fairness and societal impact.