14. A test was administered to a group of students and the score of each student on each test task were
recorded. The assessor wanted to know whether students responded in a similar way from one
task to another.
15. Three multiple-choice tests were constructed to cover the same curriculum area. Test A had 20
items, Test B had 40 items, and Test C had 60 items. What is the most accurate statement that can
be made about the reliability of these tests?
(a) All the tests will have the same reliability.
(b) Test C will be three times as reliable as Test B and Test B will be twice as reliable as Test A.
(c) Test C will have the highest reliability and Test A will have the lowest reliability.
16. Suppose that a teacher wanted to determine how objective she had been in scoring her students’
essays. Which of the following procedures would be appropriate for her purpose?
(a) Retest her students with a similar test and correlate the two set of scores.
(b) Request a colleague to score the essays and correlate her scores with her colleagues’ scores.
(c) Separate by score the even and odd-numbered essays, creating two scores for each student.
Then correlate the two sets of scores.
17. A principal wanted to award prizes to students who demonstrated the best sense of citizenship.
The principal asked the students’ teachers and their immediate past teachers to rate the students.
With which type of reliability should the principal be most concerned?
(a) Internal consistency reliability
(b) Alternate forms reliability
(c) Scorer reliability
18. The standard error of measurement of a test is 3.5. A student obtained a score of 70 on the test.
How should the student’s score be interpreted?
(a) The student’s test scores will always be between 66.5 and 73.5.
(b) The student’s true score probably lies between 66.5 and 73.5.
(c) The student’s obtained scores should be raised to 73.5 or lowered to 66.5 depending on
information from other assessments.
(d) None of these is an appropriate interpretation.
19. Test A and Test B each have the same value for their standard deviations, 8.0. The reliability
coefficients of the two tests are .75 and .90, respectively. Which, if any, of the two tests has the
smaller standard error of measurement (SEM)?
(a) Test A
(b) Test B
(c) Both would have the same SEM.