Psychological Testing & Assessment – Comprehensive Bullet-Point Notes
Psychological Testing and Assessment: Key Definitions
Psychological Testing
Process of measuring psychology-related variables via devices or procedures designed to obtain a sample of behaviour.
Objective: obtain a gauge (usually numerical).
May be individual or group; yields scores.
Psychological Assessment
Gathering & integrating data for a psychological evaluation using multiple tools.
Objective: answer referral questions, solve problems, or aid decision-making.
Typically individualized; evaluator is central; requires educated selection and interpretive skill; outcome = problem-solving recommendations.
Forms of Assessment
Collaborative: assessor + assessee are partners from intake to feedback.
Therapeutic: assessment itself used for self-discovery and change.
Dynamic: evaluation → intervention → re-evaluation; focuses on learning potential during assessment.
Major Assessment Tools
Tests
Psychological test: measures intelligence, personality, aptitude, interests, attitudes, values.
Format: plan/structure, item layout, time limits, administration mode (paper, computer, CAT).
Ways tests differ:
Administration (one-on-one vs. group, active examiner vs. none).
Scoring (hand, self, machine); scores may be raw, , etc.
Psychometric soundness (reliability, validity, utility).
Interview: reciprocal information exchange; observe verbal & non-verbal cues (eye contact, voice pitch, dress, pauses). Variants include panel interviews.
Portfolio: collection of work samples (paper, audio, video, digital).
Case-History Data: archival records giving past/current adjustment, pre-morbid neuropsychological status, academic/behavioural standing.
Behavioural Observation
Naturalistic or laboratory; quantitative & qualitative recording.
Role-Play Tests: simulate situations; evaluate expressed thoughts/behaviours.
Computer-Based Tools
Local Processing, Central Processing, Teleprocessing.
Score Reports: simple, extended, interpretive (descriptive, screening, consultative, integrative).
CAPA, CAT; pros (speed, reach, cost, “green”); cons (client integrity, mere testing vs. assessment).
Other: video scenarios, biofeedback, thermometers, etc.
Stakeholders
Test Developer: conceives, prepares, pilots, publishes.
Test User: selects, administers, scores, interprets.
Test Taker: subject of assessment; variability in anxiety, cooperation, pain, coaching, social desirability.
Psychological Autopsy: reconstruction of deceased’s profile from records/interviews.
Typical Settings & Purposes
Education (achievement, diagnostic, informal evaluation).
Clinical (screen/diagnose behavioural problems).
Counselling (adjustment, productivity).
Geriatric (quality of life variables).
Business/Military (selection, placement).
Credentialing, Government.
Assessment of Individuals with Disabilities
Accommodation: adapt test/procedure (e.g., Braille, extended time).
Alternate Assessment: non-standard evaluation procedures.
Must consider capabilities of assessee & assessor, purpose, meaning of scores.
Reference Sources
Test catalogues, manuals, reference volumes, journal reviews, online data bases.
Taxonomy of Tests
Administration: Individual vs. Group.
Ability Tests: Achievement, Aptitude/Prognostic, Intelligence.
Personality Tests: Objective/Structured vs. Projective/Unstructured.
Interest Tests.
Historical Milestones
Ancient China civil-service exams; Greco-Roman humours.
19th c.: Galton (anthropometry, individual differences); Wundt (experimental lab); Cattell (mental tests).
20th c.: Binet–Simon (1905) intelligence scale; WWI Army Alpha/Beta; Wechsler scales; Woodworth Psychoneurotic Inventory; Projective era (Rorschach).
Culture & Assessment Issues
Verbal language & translation nuances.
Non-verbal behaviours mean different things (eye contact, test timing).
Standards of evaluation vary; need culturally appropriate norms.
Group membership & fairness (height requirements, affirmative action).
Legal & Ethical Framework
Minimum competency testing; Truth-in-testing laws.
Expert testimony standards: Frye → Daubert.
Test-user qualification levels A/B/C.
Test taker rights: informed consent, feedback, privacy/confidentiality, least-stigmatizing label.
Statistics Refresher
Measurement & Scales
Nominal, Ordinal, Interval, Ratio; properties: Magnitude, Equal Intervals, Absolute Zero.
Descriptive Stats
Central Tendency: Mean , Median, Mode.
Variability: Range, , Variance , SD .
Shapes: Skew (+/–), Kurtosis (platy, meso, lepto).
Normal Curve: symmetrical; area under curve interpretable; standard scores (z, T, SAT, GRE).
, .
Correlation
Pearson , Spearman , biserial, point-biserial, phi, tetrachoric.
= coefficient of determination.
Regression
Simple: ; SEE; Multiple Regression.
Residuals, shrinkage, cross-validation.
Core Assumptions in Testing
Traits & states exist.
They can be quantified.
Test behaviour predicts non-test behaviour.
Tests have strengths & weaknesses.
Measurement includes error (Classical Test Theory ).
Testing can be fair/unbiased.
Testing benefits society.
Characteristics of a Good Test
Reliability (consistency, low error variance).
Validity (content, criterion-related, construct).
Clear instructions & easy administration.
Actionable, useful results.
Norms & Standard Errors
Norm-Referenced vs. Criterion-Referenced evaluation.
Sampling methods: stratified, purposive, incidental.
Types of norms: age, grade, developmental, national, subgroup, local, fixed-reference.
Standard Errors:
(prediction), , .
Reliability Methods
Test–Retest (coefficient of stability).
Alternate/Parallel Forms (coefficient of equivalence).
Internal Consistency: Split-Half + Spearman-Brown , KR-20/KR-21, Cronbach .
Inter-scorer reliability (Cohen/Fleiss ).
Factors: homogeneity, dynamic traits, speed vs. power, restriction of range.
Item Response Theory (IRT), Generalizability Theory.
Validity Evidence
Content: expert judgement, Lawshe .
Criterion-Related: Concurrent vs. Predictive; validity coefficient, expectancy tables, Taylor-Russell, Naylor-Shine; hit/miss, false +/–.
Construct: homogeneity, developmental changes, contrasted groups, convergent & discriminant (multitrait-multimethod), factor analysis.
Face validity ≠ technical validity.
Bias vs. Fairness; rating errors (leniency, severity, halo), DIF analysis.
Utility & Decision Making
Utility = benefits – costs; Brogden–Cronbach–Gleser formula; productivity gain.
Cut-scores: relative (norm-referenced) vs. fixed (criterion). Methods: Angoff, Bookmark, known-groups, IRT item-mapping.
Test Development Cycle
Conceptualization: define construct, target population, purpose.
Construction: scaling (nominal/ordinal/interval/ratio; Likert, Guttman), item writing (selected- vs. constructed-response), CAT.
Tryout: pilot sample; examine item difficulty ; item discrimination , item-validity & reliability indices.
Item Analysis: ICCs, distractor analysis, qualitative reviews (think-aloud, sensitivity panels).
Revision: cross-validation, co-norming, anchoring protocols; build item banks, address DIF.
Intelligence: Concepts & Theories
Interactionism: genes × environment.
Factor-Analytic Theories: Spearman g & s; Thurstone 7 primaries; Cattell Gf/Gc; Horn (vulnerable vs. maintained); Guilford SOI; Carroll 3-stratum; CHC model.
Information-Processing: Luria (simultaneous vs. successive); PASS.
Multiple Intelligences: Gardner.
Triarchic: Sternberg (analytical, creative, practical).
Developmental: Piaget (sensorimotor → formal operations).
Issues: nature vs. nurture, stability, Flynn effect, culture loading, personality correlates, gender, family environment.
Intelligence Tests
Individual
Stanford–Binet-5: routing tests, adaptive basal/ceiling, verbal & non-verbal, mean 100 SD 16.
Wechsler Series: WAIS-IV, WISC-IV, WPPSI-III; index scores (VCI, PRI, WMI, PSI); short form (WASI).
Kaufman (KABC-II, KAIT, KBIT), Woodcock–Johnson-III.
Group
Army Alpha/Beta, ASVAB, SAT/ACT, GRE, LSAT, MAT.
Non-verbal: Raven Progressive Matrices, Culture-Fair Test.
Special Populations
Infant: Brazelton BNAS, Bayley-III, Cattell CIIS, Gesell.
Disabilities: Columbia Mental Maturity, Peabody Picture Vocabulary, Leiter-R.
Learning Disabilities: ITPA-3, Woodcock–Johnson; RTI framework.
Visual-Motor: Bender-Gestalt, Benton VRT.
Creativity: Torrance Tests (fluency, originality, flexibility, elaboration).
End-to-End Workflow Summary
Identify need & context (ethical/legal, cultural, practical).
Choose/develop appropriate test with sound psychometrics.
Administer with accommodations & rapport; collect protocols.
Score (manual, computer, CAT); compute standard/percentile/z/T.
Interpret using norms, reliability, validity, utility; integrate with other data (interviews, history).
Communicate findings respecting rights (feedback, confidentiality, least-stigmatizing labels).
Reassess & revise instruments; monitor fairness and societal impact.