Wide Range Achievement Psychometric Properties Questions and Answers — Questions and Answers
Question 1: A school psychologist notes that the WRAT-5 manual reports an average internal consistency reliability coefficient (Cronbach's alpha) of .92 for the Word Reading subtest. What is the most accurate interpretation of this psychometric property?
- The Word Reading subtest will produce very similar scores if administered to the same student a week later.
- The items within the Word Reading subtest are highly inter-correlated and measure a single, unified construct. (Correct answer)
- The Word Reading subtest accurately predicts a student's overall reading ability, including comprehension.
- The scoring of the Word Reading subtest is highly consistent across different examiners.
Correct answer: The items within the Word Reading subtest are highly inter-correlated and measure a single, unified construct.
Internal consistency reliability, often measured by Cronbach's alpha, indicates the extent to which all items on a test measure the same underlying construct. A high coefficient, like .92, suggests that the items are homogenous and work together to measure word recognition skills consistently. Test-retest reliability relates to score stability over time, predictive validity relates to predicting other outcomes, and inter-rater reliability relates to consistency between scorers.
Question 2: To establish the convergent validity of the WRAT-5, the test developers conducted studies correlating its subtest scores with scores from other well-established achievement tests. Which of the following findings would provide the strongest evidence of convergent validity for the WRAT-5 Math Computation subtest?
- A low correlation with a measure of reading fluency.
- A high correlation with the WIAT-III Numerical Operations subtest. (Correct answer)
- A moderate correlation with a measure of nonverbal intelligence.
- A high correlation with the WRAT-4 Math Computation subtest.
Correct answer: A high correlation with the WIAT-III Numerical Operations subtest.
Convergent validity is a type of construct validity that demonstrates a test correlates strongly with other tests that measure the same or similar constructs. The WIAT-III Numerical Operations subtest is a well-established measure of math calculation skills, similar to the WRAT-5 Math Computation subtest. Therefore, a high correlation between them provides strong evidence that both tests are measuring the same underlying ability. Correlation with a previous version of the same test demonstrates reliability over time, not necessarily convergence with external measures.
Question 3: The WRAT-5 was standardized on a large, national sample. Which of the following best describes the primary purpose of using such a normative sample in the test's development?
- To ensure the test items are of appropriate difficulty for all age groups.
- To provide a reference group for comparing an individual's performance and deriving standard scores. (Correct answer)
- To prove that the test can be administered reliably by any qualified examiner.
- To eliminate cultural and linguistic bias from all of the test items.
Correct answer: To provide a reference group for comparing an individual's performance and deriving standard scores.
Standardization involves administering a test to a large, representative sample to establish norms. The primary purpose of these norms is to create a baseline or reference group, which allows an individual's raw score to be converted into a standard score (e.g., mean of 100, SD of 15), percentile rank, or grade equivalent. This comparison indicates how the individual performed relative to their same-age peers in the national sample.
Question 4: An examiner determines that a student's standard score on the Spelling subtest is 88, with a 95% confidence interval of 81-95. The Standard Error of Measurement (SEM) for this subtest at the student's age is 3.5. What does this confidence interval indicate?
- There is a 95% chance that the student's true spelling ability falls somewhere between a standard score of 81 and 95. (Correct answer)
- The student's score is significantly different from 95% of their peers.
- If the student retook the test 100 times, 95 of their scores would be exactly 88.
- The student answered 95% of the spelling items correctly.
Correct answer: There is a 95% chance that the student's true spelling ability falls somewhere between a standard score of 81 and 95.
A confidence interval provides a range of scores within which an individual's 'true score' is likely to fall, accounting for measurement error (SEM). A 95% confidence interval means that if the student were tested repeatedly, we would be 95% confident that their true score lies within that specific range (e.g., 81-95). It acknowledges that any single test score is an estimate, not a perfect measure of true ability.
Question 5: The existence of two equivalent forms (Blue and Green) for the WRAT-5 is a key psychometric feature. This feature is most directly useful for assessing which type of reliability?
- Split-half reliability
- Internal consistency
- Inter-rater reliability
- Alternate-forms reliability (Correct answer)
Correct answer: Alternate-forms reliability
Alternate-forms reliability is determined by administering two different but equivalent versions of the same test (like the WRAT-5 Blue and Green forms) to the same group of individuals and correlating the scores. This feature is valuable for retesting individuals in a short period without practice effects from seeing the exact same items.
Question 6: When evaluating the WRAT-5 for use with a specific clinical population, such as individuals with learning disabilities, which type of validity evidence is most critical for an examiner to consider?
- Content validity
- Face validity
- Discriminant validity (Correct answer)
- Predictive validity
Correct answer: Discriminant validity
Discriminant validity (also known as clinical validity or validity with special groups) is crucial in a clinical context. It refers to the evidence that a test can differentiate between a clinical group (e.g., individuals with a reading disorder) and a non-clinical control group. The WRAT-5 manual provides data on how specific clinical groups perform, which helps examiners determine if the test is sensitive to the specific academic difficulties associated with a diagnosis.
A school psychologist notes that the WRAT-5 manual reports an average internal consistency reliability coefficient (Cronbach's alpha) of .92 for the Word Reading subtest.
What is the most accurate interpretation of this psychometric property?