Cognitive Test Development
2026-09-16
Note
The stem and options serve different functions: the stem conveys one clear core idea, while the options provide alternative answers that are equivalent in form. A good item doesn’t mix the two — for example, by putting part of the essential information in only one of the options.
Practical implication
When writing an achievement test item, you start from the TIK. When writing an aptitude test item, you start from the narrow ability you want to measure — for example, induction or quantitative reasoning from the broad ability Gf discussed in Meeting 6.
Three narrow abilities commonly measured on aptitude tests (Meeting 6):
“SHOE : FOOT, as … : …”
Key
A — the relationship between a body part and the item worn/used on it, just like a shoe on a foot.
“All flowers have a pistil. All jasmines are flowers.”
Key
A — a direct logical conclusion drawn from the two premises (categorical syllogism).
“3, 6, 12, 24, …”
Key
C (48) — pattern: each term is multiplied by 2.
Using the TIK from Meeting 4: “Students can match each principle of achievement test construction with an example of it.”
Example item
“The principle of ‘measuring a representative sample of content’ in an achievement test is best illustrated by:”
A final exam that contains questions only from the last meeting
A final exam whose questions are spread evenly across all meetings
A final exam with a uniform level of difficulty
A final exam with a large number of questions
The item below has several problems. Discuss with the classmate next to you what needs to be fixed.
Item to review
“Which of the following statements is NOT a characteristic of an aptitude test?”
Always measures an innate ability present from birth that can never change at all over a person’s lifetime
Is predictive of future performance
Is not tied to any one specific subject area
All of the above are correct
Next up: the Midterm Exam
Meeting 8 is the Midterm Exam, covering the material from Meetings 1–7: types of cognitive tests, stages of test development, taxonomy and instructional objectives, achievement tests, aptitude tests, intelligence tests, and item-writing principles.
Any questions?
These slides were prepared using and Quarto with a template from UNAIR Theme.
Azwar, S. (1987). Tes prestasi: Fungsi dan pengembangan pengukuran prestasi belajar. Liberty.
Haladyna, T. M., Downing, S. M., & Rodriguez, M. C. (2002). A review of multiple-choice item-writing guidelines for classroom assessment. Applied Measurement in Education, 15(3), 309–334.
Lane, S., Raymond, M. R., & Haladyna, T. M. (Eds.). (2016). Handbook of test development (2nd ed.). Routledge.
Osterlind, S. J. (2002). Constructing test items: Multiple-choice, constructed-response, performance, and other formats (2nd ed.). Kluwer Academic Publishers.
Schneider, W. J., & McGrew, K. S. (2018). The Cattell–Horn–Carroll theory of cognitive abilities. In D. P. Flanagan & E. M. McDonough (Eds.), Contemporary intellectual assessment: Theories, tests, and issues (4th ed., pp. 73–163). Guilford Press.