Skip to main content icon/video/no-internet

The interaction of technology and norm-based, or norm-referenced, assessment provides opportunities for the collection of interesting information related to student learning and development as well as introducing several issues and concerns. Reviewing the purpose of assessment is critical to the use and interpretation of assessments. Categorizing assessments can clarify purposes and help educators and policy makers address issues. This entry first discusses the benefits and disadvantages of norm-based assessment and its use in computer-based and computer adaptive testing. It then discusses some of the uses and potential uses of norm-based assessments.

A common type of assessment involves the comparison of scores to a standard— criterion-referenced —or to those from a group of similar individuals— norm-referenced. Royal Van Horn argued for the use of improvement-referenced assessment (criterion- referenced measures taken multiple times), but Peter Behuniak proposed that assessment should be consumer-referenced. When analyzing the efficacy of norm-referenced assessment, two critical questions become (1) what are the disadvantages and benefits of norm-referenced assessment that cannot be addressed with other approaches, and (2) if norm-referenced assessment is warranted, how can technology be used to make the process efficient and effective?

Dylan Wiliam stated that scores derived from norm-referenced testing are relatively insensitive to instruction; this is the primary challenge for their use in assessment of academic knowledge and skills. As a result, a shift of classroom- and state-based standardized assessments in the United States has occurred toward the emphasis of criterion-referenced. Nevertheless, there are several areas where an effort to use norm-referenced assessment might be justified: (1) increasing the efficiency of norm-referenced assessments when comparisons are made on academic content or skills where there is a lack of coherence of expected standards or taught content, (2) providing learners and other stakeholders information about how students compare with large groups of similar individuals, (3) providing important information for program evaluation if learner input characteristics among the comparison groups are matched, and (4) collecting data on important domains of human development and behavior that do not as yet have established standards and benchmarks.

Computer-Based Testing and Computer Adaptive Testing

Two technological contributions to norm-referenced assessment are computer-based testing and computer adaptive testing. The major advantage of computer-based testing is that results can be provided quickly. An additional advantage is that the testing may be more secure. However, many schools do not have the computer resources for all students to take the exam at the same time. Therefore, paper and pencil tests are available where an adequate number of computers or computer access is not available. A second advantage of computer-based testing is the opportunity for more interactivity. For example, the test developers can provide simulations, have students view video clips, or have students engage in activities.

Computer adaptive testing (CAT) provides a different procedure for presenting the tested content to the test taker. In this process, not all items are given to any specific test taker; rather, the computer adapts to the correct and incorrect answers provided by the test taker and presents items accordingly. Advocates suggest that CAT delivers more valid results if most of the items answered by the test taker match his or her level of knowledge. Wim van der Linden and Peter Pashley showed that a major advantage to test developers is that the processes of item selection and the estimation of item difficulty can be made efficient using the Rasch measurement model. A second advantage is that the testing procedure can be more efficient because each examinee does not answer all possible questions. Rather, the items are organized into what are called testlets, which van der Linden described as sets of content-related items. Examinees are systematically provided with testlets until it is determined that they can or cannot answer most of the questions reliably. The test thereby converges on the knowledge and skill of the examinee.

...

  • Loading...
locked icon

Sign in to access this content

Get a 30 day FREE TRIAL

  • Watch videos from a variety of sources bringing classroom topics to life
  • Read modern, diverse business cases
  • Explore hundreds of books and reference titles

Sage Recommends

We found other relevant content for you on other Sage platforms.

Loading