Psychology Week 10 Assignment: Applying Theory to Practice

profileGMcNeel0
PersonnelSelectionAssessment.pptx

Personnel Selection & Assessment: Psychometric Foundations to Modern Evidence-Based Practice.

Gabrielle McNeely

Capella University

Michael Gontarz

PSY 6710

8/9/2026

Subfield Overview & Objectives

This workshop reviews the history of personnel selection and assessment, beginning with WWI Army Alpha/Beta tests, the Civil Rights Act of 1964 and current times.

Theories which are used as a foundational framework are Classical Test Theory, Item Response Theory, Campbell's validity framework (content, criterion, construct) and the Five-Factor Model of personality.

Situational judgment tests are in use, and emerging tests, such as those using AI and games, are also being utilized in addition to the other traditional tests, such as structured interviews, cognitive ability tests, and personality inventories.

This evaluation is essential regarding the ability of contemporary techniques to meet the basic psychometric standards of reliability, predictive validity/fairness, disparate impact and algorithmic bias.

The workshop, however, finds that with the improvement of delivery and efficiency, technology comes with the same principles from the science that underlies the research.

These principles must be followed for a defensible and ethical assessment.

Validity evidence is clearly more strongly adhered to, and fairness outcomes are less consistent, and therefore must be continually monitored to ensure that the psychological and legal standards are followed.

Speaker notes

To illustrate, personnel selection is an applied science of decision-making that is concerned with matching the individual's KSAOs with the essential requirements of the job. This subfield is applicable to the productivity, retention and diversity of the organization. The main aim of our research is to assess the connection between existing staffing patterns and past influences and psychological theories. We will discuss the relevance of early psychometric models to testing in the corporate world today and critically review the possibility that new phenomena, such as AI video interviewing, uphold classical tenets of construct validity and adverse impact mitigation. Understanding these connections enables practitioners to be able to design evidence-based, well-designed assessment systems.

‹#›

The influences of history on selection (for students).

The way people are recruited today has been greatly influenced by the past.

During this time, the Army Alpha and Beta tests led by Yerkes, Otis and Scott gained credibility in the field of WWII Army testing, which ultimately resulted in the first cognitive standardized tests for military classification and aptitude testing for civilians.

The 1964 Civil Rights Act and the 1978 Uniform Guidelines stipulated that jobs must be relevant, thereby creating a legal framework for employers to establish the validity of their selection process and to test against protected groups, in accordance with adverse impact requirements and testing to prove the process is legally defensible.

In the 1980s-1990s, Hunter and Schmidt provided evidence of the general usefulness of cognitive ability testing by demonstrating that selection validities were applicable to a variety of organizational settings and across different cognitive ability tests.

The digital and AI era (2000s to present) has ushered in selection in the online, gamified, machine-learning predictive assessment field, and added to the ethical and validity concerns that continue to be explored.

Speaker Notes

Multi-method batteries of assessments are used for personnel selection to assess maximum and typical performance. The tests of general mental ability have remained popular predictors of success in complex occupations because they have been shown to have valid prediction (Schmidt & Hunter, 1998). The shift from general to specific interviews has led to a more structured process of interviewing, featuring behaviourally anchored rating scales and questions. In the prediction of contextual performance, citizenship and reduced counterproductive behaviours, personality inventories are used with two sets of measures: one set is integrity measures, and the other is personality inventories based on the Big Five (especially Conscientiousness). These different predictors provide a greater incremental validity than single-method predictors (Sackett et al., 2022).

‹#›

Current Approaches to Personnel Selection

Current tools and practices for predicting job performance and organizational fit for selection of individuals rely on evidence-based approaches.

Cognitive testing refers to the standardized testing of General Mental Ability (GMA) and specific abilities, and always shows good predictive ability for general areas of task performance in complex technical occupations (Schmidt & Hunter, 1998).

Structured interviews are done with an anchored rating scale that asks the interviewer to elicit past behaviours and present future situations on which they rate the response, both to minimise interviewer bias and to increase reliability.

The personality and integrity tests are standardized tests corresponding to the Five-Factor Model, with the most consistent predictor being Conscientiousness for tenure, teamwork and reduced counterproductive work behaviours (Barrick & Mount, 1991).

They are all perceived as the scientific underpinning of the present-day selection process, focusing specifically on standardization, empirical testing and the necessity to maximise the chances of success in hiring while minimising legal and organisational liability.

Speaker notes

The selection practices as they are now are direct results of the evolution in the past 100 years. During World War I, Yerkes and Scott developed the Army Alpha and Army Beta tests to process over 1.7 million soldiers in the army. They advanced the idea of psychological measurement to be applied to organizations. The Civil Rights Act of 1964 and the 1978 Uniform Guidelines set the legal groundwork for non-discriminatory selection instruments that are job-related. In the 1980s, Hunter and Schmidt presented evidence that the validity of cognitive tests, via a meta-analysis, transfers to settings. Today's digital screening systems are based on these milestones, with state-of-the-art computational power added to the mix.

‹#›

Foundational Theoretical Frameworks

The theoretical basis is used to support the personnel selection practices as indicated in the theoretical base.

Classical Test Theory (CTT) and psychometrics make the assumption that observed test scores are the sum of a True Score plus Random Error (X = T + E).

In this context, the internal consistency, test-retest reliability and criterion-related validity of the test are extremely important so that the test can be measured properly.

It serves as a statistical source to evaluate and enhance selection instruments.

Behavioural interviewing and work samples are based on Behaviour Consistency and Trait Theory, which assumes behaviours are stable and consistent, and can be predicted for future behaviours in the workplace.

Trait theory uses the stable constructs of personality as the explanation of the outcome at work: the Big 5 personality types.

Overall, Conscientiousness is the most consistent predictor of job performance, length of tenure, and organizational citizenship in a variety of jobs.

All these are part of a larger strategy to be able to conduct structured hiring assessments and to have legitimate personality testing as legitimate assessment tools for making evidence-based hiring decisions.

The theoretical basis is used to support the personnel selection practices as indicated in the theoretical base. Classical Test Theory (CTT) and psychometrics make the assumption that observed test scores are the sum of a True Score plus Random Error (X = T + E). In this context, the internal consistency, test-retest reliability and criterion-related validity of the test are extremely important so that the test can be measured properly. It serves as a statistical source to evaluate and enhance selection instruments. Behavioural interviewing and work samples are based on Behaviour Consistency and Trait Theory, which assumes behaviours are stable and consistent, and can be predicted for future behaviours in the workplace. Trait theory uses the stable constructs of personality as the explanation of the outcome at work: the Big 5 personality types. Overall, Conscientiousness is the most consistent predictor of job performance, length of tenure, and organizational citizenship in a variety of jobs. All these are part of a larger strategy to be able to conduct structured hiring assessments and to have legitimate personality testing as legitimate assessment tools for making evidence-based hiring decisions.

‹#›

The predictive validity of the selection predictors.

Meta-analytic studies have revealed operational validity coefficients for each of the different selection methods applied in predicting overall job performance, thus giving empirical guidance for evidence-based hiring.

Work sample tests have the highest predictive validity at .54, which involves candidates completing real job tasks.

Structured interviews come in second place at .51, with both questions and ratings anchored in a standardized set of questions and a rating scale.

Cognitive ability also has a .51 validity, which is always reliable in predicting performance in complex technical positions.

Conscientiousness measures had a moderate predictive power of .22, indicating they are useful for predicting contextual performance, organizational citizenship, and decreased counterproductive behaviour.

By comparison, graphology and unstructured chats have very little validity (of .02), and little predictive value for job performance.

The results highlight the need for a focus on structured and validated methods rather than subjective methods.

The allocation of resources to methods that have proven to be valid for the purpose of optimizing selection decisions should be done by organizations.

Speaker notes

This data visualization is a summary of the key results from several meta-analyses regarding the predictive validity of popular selection methods: the most highly operationally valid work samples, structured interviews, and cognitive ability tests. Conscientiousness provides additional incremental validity with cognitive ability because it taps into motivational effort versus maximal ability. On the other hand, non-scientific methods such as graphology have practically no predictive value. These empirical numbers directly confirm Classical Test Theory and Behavioural Consistency, and show that structured, objective tools based on construct validity are consistently superior in corporate environments to informal, unstructured selection tools, and that they are based on construct validity.

‹#›

Comparing current practices with theory.

An examination of existing selection practices and theories of selection shows that there is a lack of uniformity in the degree of adherence to the theories and psychometric risk.

Structured interviews exhibit a high degree of behavioural consistency theory with a sampling of past behaviour in a job-relevant context.

Reliability is high when a behaviourally anchored rating scale is followed by the teacher.

Classical test theory and Spearman's g have a high degree of adherence, and construct validity is good, but adverse impact calls for careful balancing of sub-groups.

In the case of AI and algorithmic video screening, the level of adherence ranges from low to partial.

AI systems can be black-boxed, or algorithmic video screening can fail to meet the Classical Test Theory principles.

Construct validity may be violated due to latent bias in the training data.

The moderately high level of adherence to the assessments is gamified and is based on psychometric measurement and trait theory.

They enhance candidate engagement, but there is a need for a thorough process to validate the games and ensure they accurately assess the required Kassy and other attributes.

An examination of existing selection practices and theories of selection shows that there is a lack of uniformity in the degree of adherence to the theories and psychometric risk. Structured interviews exhibit a high degree of behavioural consistency theory with a sampling of past behaviour in a job-relevant context. Reliability is high when a behaviourally anchored rating scale is followed by the teacher. Classical test theory and Spearman's g have a high degree of adherence, and construct validity is good, but adverse impact calls for careful balancing of sub-groups. In the case of AI and algorithmic video screening, the level of adherence ranges from low to partial. AI systems can be black-boxed, or algorithmic video screening can fail to meet the Classical Test Theory principles. Construct validity may be violated due to latent bias in the training data. The moderately high level of adherence to the assessments is gamified and is based on psychometric measurement and trait theory. They enhance candidate engagement, but there is a need for a thorough process to validate the games and ensure they accurately assess the required Kassy and other attributes.

‹#›

The intellectual and occupational requirements for compliance.

There are legal and professional guidelines governing the proper and defensible use of personnel selection devices.

For purposes of determining potential adverse impact, the Equal Employment Opportunity Commission uses the Four-Fifths Rule, which sets an 80% threshold, meaning that if a selection rate for any protected group is below 80% of the selection rate for the majority group, there is a potential adverse impact.

If adverse impact has occurred, both the Uniform Guidelines of 1978 and the EEO regulations require that employers show that the test is job related, which means evidence of empirical criterion-related, content and/or construct validation.

Assessment procedures should be well documented and free of technical unfairness and bias, per the 2018 Principles of the SIOP on Assessment.

Candidate informed consent, data privacy, test security, score interpretation, and overall, defensible, ethical, and scientifically sound selection practices are upheld by the American Psychological Association Ethical Code.

There are legal and professional guidelines governing the proper and defensible use of personnel selection devices. For purposes of determining potential adverse impact, the Equal Employment Opportunity Commission uses the Four-Fifths Rule, which sets an 80% threshold, meaning that if a selection rate for any protected group is below 80% of the selection rate for the majority group, there is a potential adverse impact. If adverse impact has occurred, both the Uniform Guidelines of 1978 and the EEO regulations require that employers show that the test is job related, which means evidence of empirical criterion-related, content and/or construct validation. Assessment procedures should be well documented and free of technical unfairness and bias, per the 2018 Principles of the SIOP on Assessment. Candidate informed consent, data privacy, test security, score interpretation, and overall, defensible, ethical, and scientifically sound selection practices are upheld by the American Psychological Association Ethical Code.

‹#›

Summary & Practitioner Recommendations

All selection instruments need to be based on thorough job analysis and theories to be job-relevant and scientifically sound.

A combination of complementary predictors (GMA, structured interviews and Conscientiousness measures) results in better incremental validity than using any single method alone and includes both maximal cognitive capacity and motivational effort.

For adverse impact purposes, and because of the possible latent bias or the fact that the algorithm's decisions would contradict the principles of measurement, systematic auditing of automated and AI-driven algorithms is required.

Validation evidence needs to be well documented and kept as a legal defensibility, as well as adverse impact analyses and candidate communication procedures, should be well documented and kept as a legal defensibility, following the guidelines of the Equal Employment Opportunity Commission and the Principles of the Society for Industrial and Organizational Psychologists.

These recommendations will serve as a collective measure to ensure scientific, legal and ethical acceptability of selection systems.

All selection instruments need to be based on thorough job analysis and theories to be job-relevant and scientifically sound. A combination of complementary predictors (GMA, structured interviews and Conscientiousness measures) results in better incremental validity than using any single method alone and includes both maximal cognitive capacity and motivational effort. For adverse impact purposes, and because of the possible latent bias or the fact that the algorithm's decisions would contradict the principles of measurement, systematic auditing of automated and AI-driven algorithms is required. Validation evidence needs to be well documented and kept as a legal defensibility, as well as adverse impact analyses and candidate communication procedures, should be well documented and kept as a legal defensibility, following the guidelines of the Equal Employment Opportunity Commission and the Principles of the Society for Industrial and Organizational Psychologists. These recommendations will serve as a collective measure to ensure scientific, legal and ethical acceptability of selection systems.

‹#›

References

Campion, M. A. (2026). Can legal and professional personnel selection principles be met with machine learning (artificial intelligence)? Human Resource Management, 65(1), 235–255.

Hoover, A. N., Rupp, D. E., & McCauley, R. (2025). Cybervoting best practices: An integrative framework for developing, validating, and implementing social media assessment for personnel selection. Employee Responsibilities and Rights Journal, 1–31.

Kenku, A. A., Yahaya, I., Uzoigwe, T. L., & Buba, J. H. (2026). Theoretical Perspectives on Psychometric Assessment as a Strategic Mechanism in Personnel Selection and Recruitment. NIU Journal of Social Sciences, 12(1), 343-352.

Landers, R. N., & Sanchez, D. R. (2022). Game‐based, gamified, and gamefully designed assessments for employee selection: Definitions, distinctions, design, and validation. International Journal of Selection and Assessment, 30(1), 1–13.

Shinde, S. (2025). The Role of Psychometric Testing in Strengthening Employee Retention and Organizational Success. TECHNO REVIEW Journal of Technology and Management, 5(1), 37–50.

Spain, R. D., Hedge, J. W., Ohse, D., & White, A. (2022). The need for research-based tools for personnel selection and assessment in the forensic sciences. Forensic Science International: Synergy, 4, 100213.