Item analysis - Index of difficulty and discrimination, Reliability - One Line Questions

1. Which of the following is generally considered the ideal range for the Index of Difficulty (ID) for most achievement tests? 0.20 to 0.80
2. An item that is answered correctly by 90% of the students and incorrectly by 10% of the students has an Index of Difficulty (ID) of: 0.90
3. If the top 25% of test-takers answered an item correctly at a rate of 80%, and the bottom 25% answered it correctly at a rate of 20%, what is the Index of Discrimination (ID) using this common method? 0.60
4. Which coefficient indicates the highest level of reliability? 0.95
5. Which statement best describes the relationship between reliability and validity? A test can be reliable without being valid.
6. Parallel-forms reliability (also known as equivalent-forms reliability) involves: Administering two different but equivalent forms of a test.
7. Which factor can decrease the reliability of a test? Ambiguous item wording
8. Which method is commonly used to calculate the Index of Discrimination? Comparing the proportion of correct answers between the top and bottom score groups.
9. A test designer wants to create a test that accurately reflects a specific curriculum. Which type of validity is most crucial? Content validity
10. If a student gets a similar score when taking the same test on two different occasions, the test is considered to have high: Test-retest reliability
11. Content validity is assessed by: Determining the extent to which test items represent the domain of content they are supposed to cover.
12. Internal consistency reliability assesses the extent to which items within a test measure the same construct. Which statistic is commonly used for this? Cronbach's Alpha
13. Which of the following is a measure of validity, not reliability? Content validity ratio
14. Test-retest reliability measures the consistency of scores over: Time.
15. An Index of Difficulty (ID) of 0.20 suggests that the item is: Difficult
16. When considering the Index of Difficulty (ID), an item with an ID of 0.30 is considered: Difficult
17. The Kuder-Richardson Formula 20 (KR-20) is used to calculate internal consistency reliability for tests with: True-false or multiple-choice items (dichotomous scoring)
18. If an item has an Index of Discrimination (ID) of 0.50 and an Index of Difficulty (ID) of 0.50, what is the assessment of this item? Good item, discriminating well and at an ideal difficulty level.
19. The Index of Discrimination (ID) measures: How well an item differentiates between high-scoring and low-scoring students.
20. An item analysis is most useful during which stage of test development? Pilot testing and revision
21. An Index of Discrimination (ID) of -0.10 suggests that the item: Functions negatively, with lower scorers more likely to answer correctly.
22. An item with an Index of Discrimination of 0.00 means: It does not differentiate between high and low scorers.
23. Which of the following is a major drawback of the split-half reliability method if not corrected by the Spearman-Brown formula? It underestimates the test's true reliability.
24. An item analysis reports that an item has an ID of 0.95 and an Index of Discrimination of 0.10. What action should likely be taken? Revise the item for clarity and difficulty.
25. The purpose of ensuring high reliability in a test is to: Minimize measurement error.
26. If an item has an Index of Difficulty (ID) of 1.00, it means: All students answered it correctly.
27. An Index of Discrimination (ID) of 0.45 is generally considered: Good
28. When an item analysis reveals an item with a very low Index of Difficulty (close to 0) and a negative Index of Discrimination, it should typically be: Revised or discarded.
29. A positive Index of Discrimination (ID) indicates that: Students who answered the item correctly tend to score higher on the total test.
30. Which of the following is a method for estimating reliability by assessing the consistency of responses to items within a single test administration? Split-half method
31. Which of the following is NOT a type of reliability estimation? Content validity
32. Which type of reliability is most appropriate for a multiple-choice test scored by a computer? Internal consistency reliability
33. The standard error of measurement (SEM) is related to reliability and indicates: The amount of error expected in an individual's score.
34. The Index of Difficulty (ID) for a test item represents: The proportion of students who answered the item correctly.
35. Reliability, in the context of educational measurement, refers to: The consistency or stability of test scores.
36. When an item analysis shows that an item has an Index of Difficulty (ID) of 0.15 and an Index of Discrimination (ID) of 0.40, what might be the interpretation? The item is difficult but discriminates moderately well.
37. An item with an Index of Difficulty (ID) of 0.75 and an Index of Discrimination (ID) of 0.30 suggests: The item is easy and discriminates moderately.
38. What is the implication of a very low Index of Difficulty (ID) for a test item? The item is too easy and may not differentiate well among students.
39. What does a low Index of Discrimination (ID) generally suggest about a test item? The item is not effectively differentiating between students with high and low overall knowledge.
40. If the Index of Difficulty (ID) for an item is 0.50, it means: Approximately half of the students answered the item correctly.
41. The Spearman-Brown Prophecy Formula is used to estimate: The reliability of a test if its length is changed.
42. Inter-rater reliability is important when: The scoring of the test involves subjective judgment.
43. Split-half reliability is a method of internal consistency where: The test is divided into two halves, and scores on the halves are correlated.
44. A high reliability coefficient suggests that: The test scores are free from random error.
45. Which of the following is a primary assumption for using the test-retest method of reliability? The trait being measured is relatively stable over the time interval.
46. What is the primary purpose of item analysis in test construction? To evaluate the effectiveness and quality of individual test items.
47. A test that yields consistent scores for the same individual under similar conditions is considered: Reliable
48. A high Index of Difficulty (ID) value, close to 1.00, indicates that the item is: Very easy for most students.
49. A reliability coefficient of 0.90 indicates: High reliability.