Entry
Reader's guide
Entries A-Z
Subject index
Classical and Modern Item Analysis
Introduction
Up until 25 years or so ago, item analysis was straightforward: multiple-choice test items were field-tested on reasonably sized samples of examinees to determine their level of difficulty and discrimination, and distractors were evaluated to determine their effectiveness in attracting examinees who were without the appropriate knowledge required for successfully answering the test items (see Crocker & Algina, 1986; Gulliksen, 1950; Lord & Novick, 1968). Items that were too easy or too hard, or less discriminating than other test items available to the test developer, were less likely to be selected for the final version of a test. In the 1970s, criterion-referenced tests were introduced into the testing field, and item analysis for these tests became less focused on determining levels of item difficulty and discrimination because these item statistics were relatively unimportant in the criterion-referenced test development process. Item congruence with the objectives they were designed to measure became one of the determining factors for item selection. Item difficulties of items measuring the same objective were used to identify potentially flawed items rather than to assess item difficulty per se. Outliers among the item difficulties were helpful in flagging potentially flawed test items. Identifying items with negative or very low item discrimination indices became important but that was about all that was important about item discrimination indices for constructing criterion-referenced tests. Clearly the use of item statistics with criterion-referenced test development was different from norm-referenced test development.
In the 1970s, modern test theory, perhaps better known as ‘item response theory (IRT)’, was introduced into the testing field and the item statistics of interest were different from the classical item statistics and also depended upon the choice of test model (Hambleton, Swaminathan & Rogers, 1991; Lord, 1980; Wright & Stone, 1979). Even the number of item statistics available to the test developer was dependent on the choice of IRT model. Modern test theory was very much focused at the item level as a strategy for gaining more flexibility in the test development process. At the same time, modern test theory is associated with stronger modelling of the item response data. Advantages, in principle, accrue from such an approach, but these advantages only come when the models being applied fit the data (e.g. the one-, two-, and three-parameter logistic test models). Model-data fit then is a critical element of modern test theory. IRT item statistics have the attractive feature that they are invariant across samples of examinees from the population of examinees for whom the test under construction is intended and this item invariance property is a major advantage to test developers. After statistically adjusting item statistics for differences in examinee samples, item statistics can be compared and contrasted, though the examinee samples on which they were based can be quite different.
One other major change in assessment has taken place that impacts strongly on item analysis practices today. Today, it is common to use performance test items that are scored polytomously. There are no multiple-choice item distractors needing to be evaluated. But, item statistics for assessing difficulty and discrimination that can be applied to polytomous response data have become important.
...
- 1. Theory and Methodology
- Ambulatory Assessment
- Assessment Process
- Assessor's Bias
- Automated Test Assembly Systems
- Classical and Modern Item Analysis
- Classical Test Theory
- Classification (General, including Diagnosis)
- Criterion-Referenced Testing: Methods and Procedures
- Cross-Cultural Assessment
- Decision (including Decision Theory)
- Diagnosis of Mental and Behavioural Disorders
- Diagnostic Testing in Educational Settings
- Dynamic Assessment (Learning Potential Testing, Testing the Limits)
- Ethics
- Evaluability Assessment
- Evaluation: Programme Evaluation (General)
- Explanation
- Factor Analysis: Confirmatory
- Factor Analysis: Exploratory
- Formats for Assessment
- Generalizability Theory
- History of Psychological Assessment
- Intelligence Assessment through Cohort and Time
- Item Banking
- Item Bias
- Item Response Theory: Models and Features
- Latent Class Analysis
- Multidimensional Item Response Theory
- Multidimensional Scaling Methods
- Multimodal Assessment (including Triangulation)
- Multitrait-Multimethod Matrices
- Needs Assessment
- Norm-Referenced Testing: Methods and Procedures
- Objectivity
- Outcome Assessment/Treatment Assessment
- Person/Situation (Environment) Assessment
- Personality Assessment through Longitudinal Designs
- Prediction (General)
- Prediction: Clinical vs. Statistical
- Qualitative Methods
- Reliability
- Report (General)
- Reporting Test Results in Education
- Self-Presentation Measurement
- Self-Report Distortions (including Faking, Lying, Malingering, Social Desirability)
- Test Adaptation/Translation Methods
- Test User Competence/Responsible Test Use
- Theoretical Perspective: Cognitive
- Theoretical Perspective: Cognitive-Behavioural
- Theoretical Perspective: Constructivism
- Theoretical Perspective: Psychoanalytic
- Theoretical Perspective: Psychological Behaviourism
- Theoretical Perspective: Psychometrics
- Theoretical Perspective: Systemic
- Trait-State Models
- Utility
- Validity (General)
- Validity: Construct
- Validity: Content
- Validity: Criterion-Related
- 2. Methods, Tests and Equipment
- Adaptive and Tailored Testing
- Analogue Methods
- Autobiography
- Behavioural Assessment Techniques
- Brain Activity Measurement
- Case Formulation
- Coaching Candidates to Score Higher on Tests
- Computer-Based Testing
- Equipment for Assessing Basic Processes
- Field Survey: Protocols Development
- Goal Attainment Scaling (GAS)
- Idiographic Methods
- Interview (General)
- Interview in Behavioural and Health Settings
- Interview in Child and Family Settings
- Interview in Work and Organizational Settings
- Neuropsychological Test Batteries
- Observational Methods (General)
- Observational Techniques in Clinical Settings
- Observational Techniques in Work and Organizational Settings
- Projective Techniques
- Psychoeducational Test Batteries
- Psychophysiological Equipment and Measurements
- Self-Observation (Self-Monitoring)
- Self-Report Questionnaires
- Self-Reports (General)
- Self-Reports in Behavioural Clinical Settings
- Self-Reports in Work and Organizational Settings
- Socio-Demographic Conditions
- Sociometric Methods
- Standard for Educational and Psychological Testing
- Subjective Methods
- Test Accommodations for Disabilities
- Test Anxiety
- Test Designs: Developments
- Test Directions and Scoring
- Testing through the Internet
- Unobtrusive Measures
- 3. Personality
- Anxiety Assessment
- Attachment
- Attitudes
- Attribution Styles
- Big Five Model Assessment
- Burnout Assessment
- Cognitive Styles
- Coping Styles
- Emotions
- Empowerment
- Interest
- Leadership Personality
- Locus of Control
- Motivation
- Optimism
- Person/Situation (Environment) Assessment
- Personal Constructs
- Personality Assessment (General)
- Personality Assessment through Longitudinal Designs
- Prosocial Behaviour
- Self-Control
- Self-Efficacy
- Self-Presentation Measurement
- Self, The (General)
- Sensation Seeking
- Social Competence (including Social Skills, Assertion)
- Temperament
- Time Orientation
- Trait-State Models
- Values
- Weil-Being (including Life Satisfaction)
- 4. Intelligence
- Attention
- Cognitive Ability: g Factor
- Cognitive Ability: Multiple Cognitive Abilities
- Cognitive Decline/Impairment
- Cognitive Plasticity
- Cognitive Processes: Current Status
- Cognitive Processes: Historical Perspective
- Cognitive/Mental Abilities in Work and Organizational Settings
- Creativity
- Dynamic Assessment (Learning Potential Testing, Testing the Limits)
- Emotional Intelligence
- Equipment for Assessing Basic Processes
- Fluid and Crystallized Intelligence
- Intelligence Assessment (General)
- Intelligence Assessment through Cohort and Time
- Language (General)
- Learning Disabilities
- Memory (General)
- Mental Retardation
- Practical Intelligence: Conceptual Aspects
- Practical Intelligence: Its Measurement
- Problem Solving
- Triarchic Intelligence Components
- Wisdom
- 5. Clinical and Health
- Anger, Hostility and Aggression Assessment
- Antisocial Disorders Assessment
- Anxiety Assessment
- Anxiety Disorders Assessment
- Applied Behavioural Analysis
- Applied Fields: Clinical
- Applied Fields: Gerontology
- Applied Fields: Health
- Caregiver Burden
- Child and Adolescent Assessment in Clinical Settings
- Clinical Judgement
- Coping Styles
- Counselling, Assessment in
- Couple Assessment in Clinical Settings
- Dangerous/Violence Potential Behaviour
- Dementia
- Diagnosis of Mental and Behavioural Disorders
- Dynamic Assessment (Learning Potential Testing, Testing the Limits)
- Eating Disorders
- Health
- Identity Disorders
- Interview in Behavioural and Health Settings
- Irrational Beliefs
- Learning Disabilities
- Mental Retardation
- Mood Disorders
- Observational Techniques in Clinical Settings
- Outcome Assessment/Treatment Assessment
- Palliative Care
- Prediction: Clinical vs. Statistical
- Psychoneuroimmunology
- Quality of Life
- Self-Observation (Self-Monitoring)
- Self-Reports in Behavioural Clinical Settings
- Social Competence (including Social Skills, Assertion)
- Stress
- Substance Abuse
- Test Anxiety
- Thinking Disorders Assessment
- Type A: A Proposed Psychosocial Risk Factor for Cardiovascular Diseases
- Type C: A Proposed Psychosocial Risk Factor for Cancer
- 6. Educational and Child Assessment
- Achievement Testing
- Applied Fields: Education
- Child Custody
- Children with Disabilities
- Coaching Candidates to Score Higher on Tests
- Cognitive Psychology and Assessment Practices
- Communicative Language Abilities
- Development (General)
- Development: Intelligence/Cognitive
- Development: Language
- Development: Psychomotor
- Development: Socio-Emotional
- Diagnostic Testing in Educational Settings
- Dynamic Assessment (Learning Potential Testing, Testing the Limits)
- Evaluation in Higher Education
- Giftedness
- Instructional Strategies
- Interview in Child and Family Settings
- Item Banking
- Learning Strategies
- Performance
- Performance Standards: Constructed Response Item Formats
- Performance Standards: Selected Response Item Formats
- Planning
- Planning Classroom Tests
- Pre-School Children
- Psychoeducational Test Batteries
- Reporting Test Results in Education
- Standard for Educational and Psychological Testing
- Test Accommodations for Disabilities
- Test Directions and Scoring
- Testing in the Second Language in Minorities
- 7. Work and Organizations
- Achievement Motivation
- Applied Fields: Forensic
- Applied Fields: Organizations
- Applied Fields: Work and Industry
- Career and Personnel Development
- Centres (Assessment Centres)
- Cognitive/Mental Abilities in Work and Organizational Settings
- Empowerment
- Interview in Work and Organizational Settings
- Job Characteristics
- Job Stress
- Leadership in Organizational Settings
- Leadership Personality
- Motor Skills in Work Settings
- Observational Techniques in Work and Organizational Settings
- Organizational Culture
- Performance
- Personnel Selection, Assessment in
- Physical Abilities in Work Settings
- Risk and Prevention in Work and Organizational Settings
- Self-Reports in Work and Organizational Settings
- Total Quality Management
- 8. Neurophysiopsychological Assessment
- Applied Fields: Neuropsychology
- Applied Fields: Psychophysiology
- Brain Activity Measurement
- Dementia
- Equipment for Assessing Basic Processes
- Executive Functions Disorders
- Memory Disorders
- Neuropsychological Test Batteries
- Outcome Evaluation in Neuropsychological Rehabilitation
- Psychoneuroimmunology
- Psychophysiological Equipment and Measurements
- Visuo-Perceptual Impairments
- Voluntary Movement
- 9. Environmental Assessment
- Behavioural Settings and Behaviour Mapping
- Cognitive Maps
- Couple Assessment in Clinical Settings
- Environmental Attitudes and Values
- Family
- Landscapes and Natural Environments
- Life Events
- Organizational Structure, Assessment of
- Perceived Environmental Quality
- Person/Situation (Environment) Assessment
- Post-Occupancy Evaluation for the Built Environment
- Residential and Treatment Facilities
- Social Climate
- Social Networks
- Social Resources
- Stressors: Physical
- Stressors: Social
- Loading...
Get a 30 day FREE TRIAL
-
Watch videos from a variety of sources bringing classroom topics to life
-
Read modern, diverse business cases
-
Explore hundreds of books and reference titles
Sage Recommends
We found other relevant content for you on other Sage platforms.
Have you created a personal profile? Login or create a profile so that you can save clips, playlists and searches