aReading Technical Foundations 3 — Questions and Answers
Question 1: Which measurement framework underpins the item calibration process for aReading?
- Item Response Theory (IRT) (Correct answer)
- Classical Test Theory (CTT) only
- Generalizability Theory
- Confirmatory Factor Analysis
Correct answer: Item Response Theory (IRT)
aReading items are calibrated using Item Response Theory (IRT), which models the relationship between a student's ability and the probability of a correct response.
Question 2: What does a high item discrimination parameter (a-parameter) indicate in IRT?
- The item clearly separates high-ability from low-ability readers (Correct answer)
- The item is too easy for most students
- The item has many correct answer choices
- The item does not fit the IRT model
Correct answer: The item clearly separates high-ability from low-ability readers
A high a-parameter means the item strongly differentiates between students of differing reading ability, making it more informative for the CAT algorithm.
Question 3: Which IRT parameter reflects an item's difficulty level in the aReading item bank?
- b-parameter (Correct answer)
- a-parameter
- c-parameter
- d-parameter
Correct answer: b-parameter
The b-parameter (difficulty) represents the ability level at which a student has a 50% probability of answering the item correctly.
Question 4: In a three-parameter IRT model, what does the c-parameter represent?
- The pseudo-guessing probability for a low-ability student (Correct answer)
- The item discrimination value
- The item difficulty level
- The number of distractors in the item
Correct answer: The pseudo-guessing probability for a low-ability student
The c-parameter (pseudo-guessing) estimates the probability that a very low-ability student answers correctly by chance, common in multiple-choice formats.
Question 5: What is the primary purpose of the aReading item bank's calibration studies?
- To establish stable item parameters before the items are used operationally (Correct answer)
- To rank students nationally before fall screening
- To determine how many items each grade needs
- To verify that teachers are administering the test correctly
Correct answer: To establish stable item parameters before the items are used operationally
Calibration studies collect response data from large samples to estimate reliable IRT parameters for each item prior to operational use.
Question 6: Why is aReading considered 'scale-free' with respect to grade level?
- A student's score reflects ability directly, regardless of which grade-level items were administered (Correct answer)
- Students at all grades receive identical items
- Scores reset to zero at the start of each new grade
- The scale only applies to grades 3–8
Correct answer: A student's score reflects ability directly, regardless of which grade-level items were administered
Because all items sit on a single continuous RIT scale, a student's score describes their reading ability independently of which grade's items were presented.
Question 7: Which of the following best defines 'measurement invariance' in the context of aReading score interpretation?
- Item parameters remain stable across different groups of students (Correct answer)
- All students receive the same score range
- The test never changes from year to year
- Scores cannot be compared across different testing seasons
Correct answer: Item parameters remain stable across different groups of students
Measurement invariance means item parameters function the same way across demographic subgroups, supporting fair and comparable score interpretations.
Which measurement framework underpins the item calibration process for aReading?