Entry
Reader's guide
Entries A-Z
Subject index
Validity (General)
Introduction
Tests and other forms of assessment are designed to provide information that will be useful for some purpose. The degree to which the information provided by a test score is useful, appropriate, and accurate is described by the psychometric concept validity. Validity is the extent to which the inferences (interpretations) derived from test scores are justifiable from both scientific and equity perspectives. For decisions based on test scores to be valid, the use of a test for a particular purpose must be supported by theory and empirical evidence, and biases in the measurement process must be ruled out.
Validity is not an intrinsic property of a test. As many psychometricians have pointed out (e.g. Cronbach, 1971; Messick, 1989; Shepard, 1993), in judging the worth of a test, it is the inferences derived from the test scores that must be validated, not the test itself. Therefore, the specific purpose(s) for which test scores are being used must be considered when evaluating validity. For example, a test may be useful for one purpose, such as patient diagnosis, but not for another, such as evaluating the treatment of patients.
Contemporary definitions of validity in testing borrow largely from Messick (1989) who stated ‘validity is an integrated evaluative judgement of the degree to which empirical evidence and theoretical rationales support the adequacy and appropriateness of inferences and actions based on test scores or other modes of assessment’ (p. 13). From this definition, it is clear that validity is not something that can be established by a single study and that tests cannot be labelled ‘valid’ or ‘invalid’. Given that (a) validity is the most important consideration in evaluating the use of a test for a particular purpose, and (b) such utility can never be unequivocally established, establishing that a test is appropriate for a particular purpose is an arduous task. In the remainder of this entry, specific forms of evidence for validity as well as some validation frameworks will be discussed. Before describing these concepts and practices, the following facts about validity in testing should be clear: (a) tests must be evaluated with respect to a particular purpose, (b) what needs to be validated are the inferences derived from test scores, not the test itself, (c) evaluating inferences made from test scores involves several different types of qualitative and quantitative evidence, and (d) evaluating the validity of inferences derived from test scores is not a one-time event; it is a continuous process. In addition, it should be noted that although test developers must provide evidence to support the validity of the interpretations that are likely to be made from test scores, ultimately it is the responsibility of the users of a test to evaluate this evidence to ensure the test is appropriate for the purpose(s) for which it is being used.
Test Validation
To make the task of validating inferences derived from test scores both scientifically sound and manageable, Kane (1992) proposed an ‘argument-based approach to validity’. In this approach, the validator builds an argument based on empirical evidence to support the use of a test for a particular purpose. Although this validation framework acknowledges that validity can never be established absolutely, it requires evidence that (a) the test measures what it claims to measure, (b) the test scores display adequate reliability, and (c) test scores display relationships with other variables in a manner congruent with its predicted properties. Kane's practical perspective is congruent with the Standards for Educational and Psychological Testing (American Educational Research Association [AERA], American Psychological Association [APA], & National Council on Measurement in Education [NCME], 1999), which provide detailed guidance regarding the types of evidence that should be brought forward to support the use of a test for a particular purpose. For example, the Standards
...
- 1. Theory and Methodology
- Ambulatory Assessment
- Assessment Process
- Assessor's Bias
- Automated Test Assembly Systems
- Classical and Modern Item Analysis
- Classical Test Theory
- Classification (General, including Diagnosis)
- Criterion-Referenced Testing: Methods and Procedures
- Cross-Cultural Assessment
- Decision (including Decision Theory)
- Diagnosis of Mental and Behavioural Disorders
- Diagnostic Testing in Educational Settings
- Dynamic Assessment (Learning Potential Testing, Testing the Limits)
- Ethics
- Evaluability Assessment
- Evaluation: Programme Evaluation (General)
- Explanation
- Factor Analysis: Confirmatory
- Factor Analysis: Exploratory
- Formats for Assessment
- Generalizability Theory
- History of Psychological Assessment
- Intelligence Assessment through Cohort and Time
- Item Banking
- Item Bias
- Item Response Theory: Models and Features
- Latent Class Analysis
- Multidimensional Item Response Theory
- Multidimensional Scaling Methods
- Multimodal Assessment (including Triangulation)
- Multitrait-Multimethod Matrices
- Needs Assessment
- Norm-Referenced Testing: Methods and Procedures
- Objectivity
- Outcome Assessment/Treatment Assessment
- Person/Situation (Environment) Assessment
- Personality Assessment through Longitudinal Designs
- Prediction (General)
- Prediction: Clinical vs. Statistical
- Qualitative Methods
- Reliability
- Report (General)
- Reporting Test Results in Education
- Self-Presentation Measurement
- Self-Report Distortions (including Faking, Lying, Malingering, Social Desirability)
- Test Adaptation/Translation Methods
- Test User Competence/Responsible Test Use
- Theoretical Perspective: Cognitive
- Theoretical Perspective: Cognitive-Behavioural
- Theoretical Perspective: Constructivism
- Theoretical Perspective: Psychoanalytic
- Theoretical Perspective: Psychological Behaviourism
- Theoretical Perspective: Psychometrics
- Theoretical Perspective: Systemic
- Trait-State Models
- Utility
- Validity (General)
- Validity: Construct
- Validity: Content
- Validity: Criterion-Related
- 2. Methods, Tests and Equipment
- Adaptive and Tailored Testing
- Analogue Methods
- Autobiography
- Behavioural Assessment Techniques
- Brain Activity Measurement
- Case Formulation
- Coaching Candidates to Score Higher on Tests
- Computer-Based Testing
- Equipment for Assessing Basic Processes
- Field Survey: Protocols Development
- Goal Attainment Scaling (GAS)
- Idiographic Methods
- Interview (General)
- Interview in Behavioural and Health Settings
- Interview in Child and Family Settings
- Interview in Work and Organizational Settings
- Neuropsychological Test Batteries
- Observational Methods (General)
- Observational Techniques in Clinical Settings
- Observational Techniques in Work and Organizational Settings
- Projective Techniques
- Psychoeducational Test Batteries
- Psychophysiological Equipment and Measurements
- Self-Observation (Self-Monitoring)
- Self-Report Questionnaires
- Self-Reports (General)
- Self-Reports in Behavioural Clinical Settings
- Self-Reports in Work and Organizational Settings
- Socio-Demographic Conditions
- Sociometric Methods
- Standard for Educational and Psychological Testing
- Subjective Methods
- Test Accommodations for Disabilities
- Test Anxiety
- Test Designs: Developments
- Test Directions and Scoring
- Testing through the Internet
- Unobtrusive Measures
- 3. Personality
- Anxiety Assessment
- Attachment
- Attitudes
- Attribution Styles
- Big Five Model Assessment
- Burnout Assessment
- Cognitive Styles
- Coping Styles
- Emotions
- Empowerment
- Interest
- Leadership Personality
- Locus of Control
- Motivation
- Optimism
- Person/Situation (Environment) Assessment
- Personal Constructs
- Personality Assessment (General)
- Personality Assessment through Longitudinal Designs
- Prosocial Behaviour
- Self-Control
- Self-Efficacy
- Self-Presentation Measurement
- Self, The (General)
- Sensation Seeking
- Social Competence (including Social Skills, Assertion)
- Temperament
- Time Orientation
- Trait-State Models
- Values
- Weil-Being (including Life Satisfaction)
- 4. Intelligence
- Attention
- Cognitive Ability: g Factor
- Cognitive Ability: Multiple Cognitive Abilities
- Cognitive Decline/Impairment
- Cognitive Plasticity
- Cognitive Processes: Current Status
- Cognitive Processes: Historical Perspective
- Cognitive/Mental Abilities in Work and Organizational Settings
- Creativity
- Dynamic Assessment (Learning Potential Testing, Testing the Limits)
- Emotional Intelligence
- Equipment for Assessing Basic Processes
- Fluid and Crystallized Intelligence
- Intelligence Assessment (General)
- Intelligence Assessment through Cohort and Time
- Language (General)
- Learning Disabilities
- Memory (General)
- Mental Retardation
- Practical Intelligence: Conceptual Aspects
- Practical Intelligence: Its Measurement
- Problem Solving
- Triarchic Intelligence Components
- Wisdom
- 5. Clinical and Health
- Anger, Hostility and Aggression Assessment
- Antisocial Disorders Assessment
- Anxiety Assessment
- Anxiety Disorders Assessment
- Applied Behavioural Analysis
- Applied Fields: Clinical
- Applied Fields: Gerontology
- Applied Fields: Health
- Caregiver Burden
- Child and Adolescent Assessment in Clinical Settings
- Clinical Judgement
- Coping Styles
- Counselling, Assessment in
- Couple Assessment in Clinical Settings
- Dangerous/Violence Potential Behaviour
- Dementia
- Diagnosis of Mental and Behavioural Disorders
- Dynamic Assessment (Learning Potential Testing, Testing the Limits)
- Eating Disorders
- Health
- Identity Disorders
- Interview in Behavioural and Health Settings
- Irrational Beliefs
- Learning Disabilities
- Mental Retardation
- Mood Disorders
- Observational Techniques in Clinical Settings
- Outcome Assessment/Treatment Assessment
- Palliative Care
- Prediction: Clinical vs. Statistical
- Psychoneuroimmunology
- Quality of Life
- Self-Observation (Self-Monitoring)
- Self-Reports in Behavioural Clinical Settings
- Social Competence (including Social Skills, Assertion)
- Stress
- Substance Abuse
- Test Anxiety
- Thinking Disorders Assessment
- Type A: A Proposed Psychosocial Risk Factor for Cardiovascular Diseases
- Type C: A Proposed Psychosocial Risk Factor for Cancer
- 6. Educational and Child Assessment
- Achievement Testing
- Applied Fields: Education
- Child Custody
- Children with Disabilities
- Coaching Candidates to Score Higher on Tests
- Cognitive Psychology and Assessment Practices
- Communicative Language Abilities
- Development (General)
- Development: Intelligence/Cognitive
- Development: Language
- Development: Psychomotor
- Development: Socio-Emotional
- Diagnostic Testing in Educational Settings
- Dynamic Assessment (Learning Potential Testing, Testing the Limits)
- Evaluation in Higher Education
- Giftedness
- Instructional Strategies
- Interview in Child and Family Settings
- Item Banking
- Learning Strategies
- Performance
- Performance Standards: Constructed Response Item Formats
- Performance Standards: Selected Response Item Formats
- Planning
- Planning Classroom Tests
- Pre-School Children
- Psychoeducational Test Batteries
- Reporting Test Results in Education
- Standard for Educational and Psychological Testing
- Test Accommodations for Disabilities
- Test Directions and Scoring
- Testing in the Second Language in Minorities
- 7. Work and Organizations
- Achievement Motivation
- Applied Fields: Forensic
- Applied Fields: Organizations
- Applied Fields: Work and Industry
- Career and Personnel Development
- Centres (Assessment Centres)
- Cognitive/Mental Abilities in Work and Organizational Settings
- Empowerment
- Interview in Work and Organizational Settings
- Job Characteristics
- Job Stress
- Leadership in Organizational Settings
- Leadership Personality
- Motor Skills in Work Settings
- Observational Techniques in Work and Organizational Settings
- Organizational Culture
- Performance
- Personnel Selection, Assessment in
- Physical Abilities in Work Settings
- Risk and Prevention in Work and Organizational Settings
- Self-Reports in Work and Organizational Settings
- Total Quality Management
- 8. Neurophysiopsychological Assessment
- Applied Fields: Neuropsychology
- Applied Fields: Psychophysiology
- Brain Activity Measurement
- Dementia
- Equipment for Assessing Basic Processes
- Executive Functions Disorders
- Memory Disorders
- Neuropsychological Test Batteries
- Outcome Evaluation in Neuropsychological Rehabilitation
- Psychoneuroimmunology
- Psychophysiological Equipment and Measurements
- Visuo-Perceptual Impairments
- Voluntary Movement
- 9. Environmental Assessment
- Behavioural Settings and Behaviour Mapping
- Cognitive Maps
- Couple Assessment in Clinical Settings
- Environmental Attitudes and Values
- Family
- Landscapes and Natural Environments
- Life Events
- Organizational Structure, Assessment of
- Perceived Environmental Quality
- Person/Situation (Environment) Assessment
- Post-Occupancy Evaluation for the Built Environment
- Residential and Treatment Facilities
- Social Climate
- Social Networks
- Social Resources
- Stressors: Physical
- Stressors: Social
- Loading...
Get a 30 day FREE TRIAL
-
Watch videos from a variety of sources bringing classroom topics to life
-
Read modern, diverse business cases
-
Explore hundreds of books and reference titles
Sage Recommends
We found other relevant content for you on other Sage platforms.
Have you created a personal profile? Login or create a profile so that you can save clips, playlists and searches