Entry
Reader's guide
Entries A-Z
Subject index
Criterion-Referenced Testing: Methods and Procedures
Introduction
Criterion-referenced tests are constructed to allow users to interpret examinee test performance in relation to well-defined domains of content and/or behaviours. Normally, performance standards are set on the test score reporting scale to permit examinee test performance to be classified into performance categories such as below basic, basic, proficient, and advanced. Criterion-referenced tests are well suited for many of the assessment needs that exist in education, the professions, the military, and industry. Today, criterion-referenced tests are called by many names – domain-referenced tests, competency tests, basic skills tests, mastery tests, performance tests, authentic assessments, objectives-referenced tests, and more. In different contexts, test developers and users have adopted these different names. For example, in school contexts, the term ‘mastery testing’ is common. When criterion-referenced tests are developed to model classroom activities or exercises, the term ‘authentic test’ is sometimes used. When criterion-referenced tests consist of many performance tasks, the terms ‘performance test’ or ‘performance assessment’ are used. Regardless, all of these terms refer to a type of assessment where what examinees know and can do is estimated, and often performance standards are used for interpreting examinee performance.
This entry has been divided into three sections. First, the most important criterion-referenced testing concepts will be presented. Second, criterion-referenced tests will be compared to norm-referenced tests. Finally, some conclusions and predictions about the future for criterion-referenced tests will be offered.
Key Criterion-Referenced Testing Concepts
Defining Content Domains
When this approach to assessment was introduced by Glaser (1963) and Popham and Husek (1969), criterion-referenced tests were constructed to assess a set of behavioural objectives. Over the years, it became clear that behavioural objectives did not have the specificity needed to guide instruction or to serve as targets for test development and test score interpretation (Popham, 1978). Numerous attempts were made to increase the clarity of behavioural objectives including the development of detailed domain specifications that included a clearly written objective, a sample test item or two, detailed specifications for appropriate content, and details on the construction of relevant assessment materials (see Hambleton, 1998). Domain specifications seemed to meet the demand for clearer statements of the intended targets for assessment but they were very time-consuming to write and often the level of detail needed for good assessment was impossible to achieve for higher order cognitive skills, and so test developers found domain specifications to be limiting.
Recently the trend in criterion-referenced testing practices has been to write objectives focused on the more important educational outcomes (fewer instructional and assessment targets seem to be preferable) and then offer a couple of sample assessments, preferably samples that show the diversity of approaches that might be used for assessment (Popham, 2000). Coupled with these looser specifications of the objectives is an intensive effort to demonstrate the validity of any assessments that are constructed.
Writing Valid Test Items
The production of valid test items, that is test items that provide a psychometrically sound basis for assessing examinee level of proficiency or performance, require (1) well-trained item writers, (2) item review, (3) field testing, and (4) the use of multiple item formats. Well-trained item writers are persons who have had experience with the intended population of examinees, know the intended curricula, and have experience writing test items using a variety of item formats. Item review often involves checking test items for their validity in measuring the intended objectives, their technical adequacy (that is, being consistent with the best item writing practices), and ensuring items are free of bias and stereotyping. Field-testing must be carried out on samples large enough to provide stable statistical information and representative of the intended population of examinees. Unstable and/or biased item statistical information only complicates and threatens the validity of the test development process. And, finally, one of the most important changes today in testing is the introduction of new item formats, formats that permit the assessment of higher level cognitive skills (see Zenisky & Sireci, in press).
...
- 1. Theory and Methodology
- Ambulatory Assessment
- Assessment Process
- Assessor's Bias
- Automated Test Assembly Systems
- Classical and Modern Item Analysis
- Classical Test Theory
- Classification (General, including Diagnosis)
- Criterion-Referenced Testing: Methods and Procedures
- Cross-Cultural Assessment
- Decision (including Decision Theory)
- Diagnosis of Mental and Behavioural Disorders
- Diagnostic Testing in Educational Settings
- Dynamic Assessment (Learning Potential Testing, Testing the Limits)
- Ethics
- Evaluability Assessment
- Evaluation: Programme Evaluation (General)
- Explanation
- Factor Analysis: Confirmatory
- Factor Analysis: Exploratory
- Formats for Assessment
- Generalizability Theory
- History of Psychological Assessment
- Intelligence Assessment through Cohort and Time
- Item Banking
- Item Bias
- Item Response Theory: Models and Features
- Latent Class Analysis
- Multidimensional Item Response Theory
- Multidimensional Scaling Methods
- Multimodal Assessment (including Triangulation)
- Multitrait-Multimethod Matrices
- Needs Assessment
- Norm-Referenced Testing: Methods and Procedures
- Objectivity
- Outcome Assessment/Treatment Assessment
- Person/Situation (Environment) Assessment
- Personality Assessment through Longitudinal Designs
- Prediction (General)
- Prediction: Clinical vs. Statistical
- Qualitative Methods
- Reliability
- Report (General)
- Reporting Test Results in Education
- Self-Presentation Measurement
- Self-Report Distortions (including Faking, Lying, Malingering, Social Desirability)
- Test Adaptation/Translation Methods
- Test User Competence/Responsible Test Use
- Theoretical Perspective: Cognitive
- Theoretical Perspective: Cognitive-Behavioural
- Theoretical Perspective: Constructivism
- Theoretical Perspective: Psychoanalytic
- Theoretical Perspective: Psychological Behaviourism
- Theoretical Perspective: Psychometrics
- Theoretical Perspective: Systemic
- Trait-State Models
- Utility
- Validity (General)
- Validity: Construct
- Validity: Content
- Validity: Criterion-Related
- 2. Methods, Tests and Equipment
- Adaptive and Tailored Testing
- Analogue Methods
- Autobiography
- Behavioural Assessment Techniques
- Brain Activity Measurement
- Case Formulation
- Coaching Candidates to Score Higher on Tests
- Computer-Based Testing
- Equipment for Assessing Basic Processes
- Field Survey: Protocols Development
- Goal Attainment Scaling (GAS)
- Idiographic Methods
- Interview (General)
- Interview in Behavioural and Health Settings
- Interview in Child and Family Settings
- Interview in Work and Organizational Settings
- Neuropsychological Test Batteries
- Observational Methods (General)
- Observational Techniques in Clinical Settings
- Observational Techniques in Work and Organizational Settings
- Projective Techniques
- Psychoeducational Test Batteries
- Psychophysiological Equipment and Measurements
- Self-Observation (Self-Monitoring)
- Self-Report Questionnaires
- Self-Reports (General)
- Self-Reports in Behavioural Clinical Settings
- Self-Reports in Work and Organizational Settings
- Socio-Demographic Conditions
- Sociometric Methods
- Standard for Educational and Psychological Testing
- Subjective Methods
- Test Accommodations for Disabilities
- Test Anxiety
- Test Designs: Developments
- Test Directions and Scoring
- Testing through the Internet
- Unobtrusive Measures
- 3. Personality
- Anxiety Assessment
- Attachment
- Attitudes
- Attribution Styles
- Big Five Model Assessment
- Burnout Assessment
- Cognitive Styles
- Coping Styles
- Emotions
- Empowerment
- Interest
- Leadership Personality
- Locus of Control
- Motivation
- Optimism
- Person/Situation (Environment) Assessment
- Personal Constructs
- Personality Assessment (General)
- Personality Assessment through Longitudinal Designs
- Prosocial Behaviour
- Self-Control
- Self-Efficacy
- Self-Presentation Measurement
- Self, The (General)
- Sensation Seeking
- Social Competence (including Social Skills, Assertion)
- Temperament
- Time Orientation
- Trait-State Models
- Values
- Weil-Being (including Life Satisfaction)
- 4. Intelligence
- Attention
- Cognitive Ability: g Factor
- Cognitive Ability: Multiple Cognitive Abilities
- Cognitive Decline/Impairment
- Cognitive Plasticity
- Cognitive Processes: Current Status
- Cognitive Processes: Historical Perspective
- Cognitive/Mental Abilities in Work and Organizational Settings
- Creativity
- Dynamic Assessment (Learning Potential Testing, Testing the Limits)
- Emotional Intelligence
- Equipment for Assessing Basic Processes
- Fluid and Crystallized Intelligence
- Intelligence Assessment (General)
- Intelligence Assessment through Cohort and Time
- Language (General)
- Learning Disabilities
- Memory (General)
- Mental Retardation
- Practical Intelligence: Conceptual Aspects
- Practical Intelligence: Its Measurement
- Problem Solving
- Triarchic Intelligence Components
- Wisdom
- 5. Clinical and Health
- Anger, Hostility and Aggression Assessment
- Antisocial Disorders Assessment
- Anxiety Assessment
- Anxiety Disorders Assessment
- Applied Behavioural Analysis
- Applied Fields: Clinical
- Applied Fields: Gerontology
- Applied Fields: Health
- Caregiver Burden
- Child and Adolescent Assessment in Clinical Settings
- Clinical Judgement
- Coping Styles
- Counselling, Assessment in
- Couple Assessment in Clinical Settings
- Dangerous/Violence Potential Behaviour
- Dementia
- Diagnosis of Mental and Behavioural Disorders
- Dynamic Assessment (Learning Potential Testing, Testing the Limits)
- Eating Disorders
- Health
- Identity Disorders
- Interview in Behavioural and Health Settings
- Irrational Beliefs
- Learning Disabilities
- Mental Retardation
- Mood Disorders
- Observational Techniques in Clinical Settings
- Outcome Assessment/Treatment Assessment
- Palliative Care
- Prediction: Clinical vs. Statistical
- Psychoneuroimmunology
- Quality of Life
- Self-Observation (Self-Monitoring)
- Self-Reports in Behavioural Clinical Settings
- Social Competence (including Social Skills, Assertion)
- Stress
- Substance Abuse
- Test Anxiety
- Thinking Disorders Assessment
- Type A: A Proposed Psychosocial Risk Factor for Cardiovascular Diseases
- Type C: A Proposed Psychosocial Risk Factor for Cancer
- 6. Educational and Child Assessment
- Achievement Testing
- Applied Fields: Education
- Child Custody
- Children with Disabilities
- Coaching Candidates to Score Higher on Tests
- Cognitive Psychology and Assessment Practices
- Communicative Language Abilities
- Development (General)
- Development: Intelligence/Cognitive
- Development: Language
- Development: Psychomotor
- Development: Socio-Emotional
- Diagnostic Testing in Educational Settings
- Dynamic Assessment (Learning Potential Testing, Testing the Limits)
- Evaluation in Higher Education
- Giftedness
- Instructional Strategies
- Interview in Child and Family Settings
- Item Banking
- Learning Strategies
- Performance
- Performance Standards: Constructed Response Item Formats
- Performance Standards: Selected Response Item Formats
- Planning
- Planning Classroom Tests
- Pre-School Children
- Psychoeducational Test Batteries
- Reporting Test Results in Education
- Standard for Educational and Psychological Testing
- Test Accommodations for Disabilities
- Test Directions and Scoring
- Testing in the Second Language in Minorities
- 7. Work and Organizations
- Achievement Motivation
- Applied Fields: Forensic
- Applied Fields: Organizations
- Applied Fields: Work and Industry
- Career and Personnel Development
- Centres (Assessment Centres)
- Cognitive/Mental Abilities in Work and Organizational Settings
- Empowerment
- Interview in Work and Organizational Settings
- Job Characteristics
- Job Stress
- Leadership in Organizational Settings
- Leadership Personality
- Motor Skills in Work Settings
- Observational Techniques in Work and Organizational Settings
- Organizational Culture
- Performance
- Personnel Selection, Assessment in
- Physical Abilities in Work Settings
- Risk and Prevention in Work and Organizational Settings
- Self-Reports in Work and Organizational Settings
- Total Quality Management
- 8. Neurophysiopsychological Assessment
- Applied Fields: Neuropsychology
- Applied Fields: Psychophysiology
- Brain Activity Measurement
- Dementia
- Equipment for Assessing Basic Processes
- Executive Functions Disorders
- Memory Disorders
- Neuropsychological Test Batteries
- Outcome Evaluation in Neuropsychological Rehabilitation
- Psychoneuroimmunology
- Psychophysiological Equipment and Measurements
- Visuo-Perceptual Impairments
- Voluntary Movement
- 9. Environmental Assessment
- Behavioural Settings and Behaviour Mapping
- Cognitive Maps
- Couple Assessment in Clinical Settings
- Environmental Attitudes and Values
- Family
- Landscapes and Natural Environments
- Life Events
- Organizational Structure, Assessment of
- Perceived Environmental Quality
- Person/Situation (Environment) Assessment
- Post-Occupancy Evaluation for the Built Environment
- Residential and Treatment Facilities
- Social Climate
- Social Networks
- Social Resources
- Stressors: Physical
- Stressors: Social
- Loading...
Get a 30 day FREE TRIAL
-
Watch videos from a variety of sources bringing classroom topics to life
-
Read modern, diverse business cases
-
Explore hundreds of books and reference titles
Sage Recommends
We found other relevant content for you on other Sage platforms.
Have you created a personal profile? Login or create a profile so that you can save clips, playlists and searches