Ed.D. Doctor of Education Assessment and Evaluation in Education 2 — Questions and Answers
Question 1: Program evaluation in education most commonly uses which framework to assess whether a program is achieving its intended goals?
- SWOT analysis
- Logic model (Correct answer)
- Balanced scorecard
- Needs assessment matrix
Correct answer: Logic model
A logic model visually maps a program's inputs, activities, outputs, and short-/long-term outcomes, providing a framework for evaluating program effectiveness.
Question 2: The primary distinction between summative and formative evaluation of educational programs is:
- Summative uses quantitative data; formative uses qualitative data
- Summative judges overall program worth/merit; formative guides ongoing improvement (Correct answer)
- Summative is conducted by external evaluators; formative by internal staff
- Summative measures teacher performance; formative measures student performance
Correct answer: Summative judges overall program worth/merit; formative guides ongoing improvement
Summative evaluation assesses a program's overall effectiveness or merit at the end of a cycle, while formative evaluation provides ongoing feedback during implementation to support improvement.
Question 3: Inter-rater reliability in performance assessment refers to:
- Consistency of student scores across test versions
- Agreement between two or more raters scoring the same student work (Correct answer)
- Stability of scores over time
- Alignment between assessment items and learning standards
Correct answer: Agreement between two or more raters scoring the same student work
Inter-rater reliability measures the degree of agreement between independent raters evaluating the same student performance, indicating the consistency of subjective scoring.
Question 4: Bloom's Taxonomy is most commonly used in assessment design to ensure:
- Cultural fairness of test items
- A range of cognitive demand levels across assessment items (Correct answer)
- Consistent test administration procedures
- Statistically normal score distributions
Correct answer: A range of cognitive demand levels across assessment items
Bloom's Taxonomy helps educators design assessments that measure a range of cognitive skills from lower-order (recall) to higher-order (analysis, synthesis, evaluation).
Question 5: Consequential validity refers to the:
- Alignment of a test with its stated purpose
- Social and educational consequences of test score interpretations and uses (Correct answer)
- Statistical relationship between a test and a criterion measure
- Degree to which test items represent course content
Correct answer: Social and educational consequences of test score interpretations and uses
Consequential validity considers whether the intended and unintended social consequences of test use are appropriate and justified.
Question 6: A rubric used in educational assessment is best described as:
- A multiple-choice scoring key
- A set of criteria and performance level descriptors for evaluating student work (Correct answer)
- A norm table for converting raw scores to percentiles
- A test blueprint specifying item distribution
Correct answer: A set of criteria and performance level descriptors for evaluating student work
A rubric is a scoring tool that defines evaluative criteria and describes levels of performance quality, making assessment transparent and consistent.
Program evaluation in education most commonly uses which framework to assess whether a program is achieving its intended goals?