Professional Education 8: Assessment and Evaluation of Learning

0.0(0)
Studied by 0 people
call kaiCall Kai
Locked
learnLearn
examPractice Test
spaced repetitionSpaced Repetition
heart puzzleMatch
flashcardsFlashcards
GameKnowt Play
Card Sorting

1/26

flashcard set

Earn XP

Description and Tags

Practice flashcards covering core concepts, assessment principles, taxonomies, test design, validity, reliability, item analysis, and statistical distributions from the Assessment and Evaluation of Learning lecture.

Last updated 8:33 AM on 9/6/26
Name
Mastery
Learn
Test
Matching
Spaced
Call with Kai
Chat

No analytics yet

Send a link to your students to track their progress

27 Terms

1
New cards

What is the difference between measurement and evaluation in education?

Measurement is the process of quantifying the degree to which someone or something possesses a given trait, while evaluation is the systematic collection and analysis of data to make a judgment about the desirability of changes in students.

2
New cards

What is the distinction between immediate outcomes and deferred outcomes in Outcomes-Based Education (OBE)?

Immediate outcomes are competencies and skills acquired upon completion of a lesson, subject, grade, or program (such as communication or problem-solving skills), whereas deferred outcomes refer to the long-term application of skills in professional and workplace practice (such as success in professional career planning).

3
New cards

How does William Spady define outcomes in transformational Outcomes-Based Education (OBE)?

Spady defines outcomes as deferred, long-term, cross-curricular outcomes that relate directly to a student's future life roles, such as being a productive worker, responsible citizen, or parent.

4
New cards

What are the four principles of Outcomes-Based Education (OBE)?

  1. Clarity of focus 2. Designing Down 3. High Expectations 4. Expanded Opportunities
5
New cards

What concept describes aligning assessment tasks and specific evaluation criteria directly to intended learning outcomes?

Constructive Alignment

6
New cards

What are the six levels of the Cognitive Domain in Bloom's Taxonomy?

  1. Remembering 2. Understanding 3. Applying 4. Analyzing 5. Evaluating 6. Creating
7
New cards

What are the five levels of the Affective Domain in Bloom's Taxonomy?

  1. Receiving 2. Responding 3. Valuing 4. Organizing 5. Internalising
8
New cards

What are the five levels of the Psychomotor Domain in Bloom's Taxonomy?

  1. Imitation 2. Manipulation 3. Precision 4. Articulation 5. Naturalisation
9
New cards

What three systems comprise Marzano's New Taxonomy model of thinking skills?

  1. Cognitive System 2. Metacognitive System 3. Self-System
10
New cards

What are the differences among Assessment FOR Learning, Assessment OF Learning, and Assessment AS Learning?

Assessment FOR learning is used before or during instruction to guide teaching (e.g., placement, formative, diagnostic); Assessment OF learning evaluates achievement at the end of instruction (e.g., summative); Assessment AS learning focuses on training teachers and learners on how to assess.

11
New cards

How do Traditional Assessment and Authentic Assessment compare regarding Action, Setting, Focus, and Outcome?

In Traditional Assessment, Action is selecting a response, Setting is contrived, Focus is teacher-structured, and Outcome is indirect evidence. In Authentic Assessment, Action is performing a task, Setting is simulation, Focus is student-structured, and Outcome is direct evidence.

12
New cards

What is the difference between Demonstration-type and Creation-type performance-based tasks?

Demonstration-type tasks require no physical product (such as cooking demonstrations or entertaining tourists), whereas Creation-type tasks require tangible products (such as project plans or research papers).

13
New cards

What do the components of the GRASPS acronym stand for in setting criteria for performance tasks?

G - Goal; R - Role; A - Audience; S - Situation; P - Product; S - Standards and Criteria

14
New cards

What sequence of steps is involved in the portfolio development process depicted in this diagram?

  1. Set Goals 2. Collect Evidences 3. Select 4. Organize 5. Reflect 6. Evaluate 7. Exhibit
15
New cards

According to this Venn diagram, how do Checklists, Rating Scales, and Rubrics differ in their purpose?

Checklists show observed traits of a work or performance, Rating Scales show the degree of quality of a work or performance, and Rubrics combine both traits as modified checklists and rating scales.

16
New cards

What is the key structural difference between a Holistic Rubric and an Analytic Rubric?

A Holistic Rubric gives a single rating describing the overall quality of the entire performance or product, whereas an Analytic Rubric rates identified dimensions or criteria independently to provide a detailed assessment.

17
New cards

How do Norm-Referenced Tests and Criterion-Referenced Tests differ in score interpretation?

Norm-Referenced Tests interpret results by comparing one student with other students under competition for limited high scores, whereas Criterion-Referenced Tests interpret results by comparing a student against a fixed set of criteria with no competition.

18
New cards

What distinguishes a Power Test from a Speed Test?

A Power Test consists of items with increasing levels of difficulty taken with ample time to measure difficulty capacity, whereas a Speed Test consists of items with uniform difficulty taken within a time limit to measure speed and accuracy.

19
New cards

What are the four phases in the Development and Validation of an Assessment Instrument?

Phase I: Planning Stage; Phase II: Item Writing Stage; Phase III: Try Out Stage; Phase IV: Evaluation Stage

20
New cards

How are Difficulty Index ranges interpreted in item analysis?

0.000.200.00 - 0.20: Very difficult item; 0.210.400.21 - 0.40: Difficult item; 0.410.600.41 - 0.60: Moderately difficult item; 0.610.800.61 - 0.80: Easy item; 0.810.81 and above: Very easy item

21
New cards

How are Discrimination Index ranges interpreted in item analysis?

1.000.60-1.00 - -0.60: Questionable item; 0.590.20-0.59 - -0.20: Not discriminating item; 0.210.20-0.21 - 0.20: Moderately discriminating item; 0.210.600.21 - 0.60: Discriminating item; 0.611.000.61 - 1.00: Very discriminating item

22
New cards

What is the difference between Concurrent Validity and Predictive Validity?

Concurrent Validity describes present status by correlating test scores with external measures administered concurrently, while Predictive Validity describes future performance by correlating test scores with measures administered after a longer time interval.

23
New cards

What is the difference between Convergent Validity and Divergent Validity?

Convergent Validity is established when an instrument correlates with tests measuring similar traits, while Divergent Validity is established when an instrument describes only the intended trait and does not correlate with tests measuring distinct traits.

24
New cards

Which method for measuring internal consistency reliability involves scoring odd- and even-numbered items separately?

Split Half method

25
New cards

What two types of skewed distributions are shown in this diagram, and how do their scores distribute?

  1. Positively Skewed Distribution: Most scores are low, extremely high scores exist, and the mean is greater than the mode. 2. Negatively Skewed Distribution: Most scores are high, extremely low scores exist, and the mean is lower than the mode.
26
New cards

What three types of kurtosis are shown in this diagram, and what are their respective KK values?

A. Mesokurtic (Normal) where K=0K = 0; B. Leptokurtic (peaked/steeper) where K>0K > 0; C. Platykurtic (flatter) where K<0K < 0

27
New cards

What are the definitions of z-score, stanine, and t-score?

z-score represents standard deviations above or below the mean; Stanine divides a normal distribution into 99 segments numbered 11 through 99; t-score places a score in a normal distribution with a mean of 5050 and a standard deviation of 1010.