The 0.51 Validity Coefficient: How General Mental Ability Tests Predict Job Performance
General Mental Ability tests predict job performance better than experience or education, but their 0.51 validity coefficient comes with significant trade-offs in demographic equity.
By Madison Lane
- Predictive Validity Advocates
- Industrial-organizational psychologists who prioritize maximizing hiring ROI through mathematically validated assessments.
- Adverse Impact Critics
- Legal and HR professionals focused on mitigating demographic score disparities and ensuring equitable hiring practices.
- Holistic Assessment Proponents
- Practitioners who advocate combining cognitive tests with structured interviews to balance validity and diversity.
Perspectives this story doesn't cover
- Candidates who experience test anxiety
- Neurodivergent applicants disadvantaged by timed formats
The short answer
- General Mental Ability (GMA) tests carry a 0.51 validity coefficient, making them the strongest single predictor of job performance.
- A 2022 academic attempt to lower the coefficient to 0.31 was mathematically rebutted in 2023, restoring the 0.51 standard.
- GMA validity scales with job complexity, rising to 0.58 for professional roles and dropping to 0.23 for unskilled labor.
- Despite high predictive power, cognitive tests produce demographic score disparities that trigger legal scrutiny under the EEOC's four-fifths rule.
- Combining a cognitive test with a structured interview pushes overall predictive validity to 0.65 while mitigating adverse impact.
Structured interviews and General Mental Ability (GMA) tests share the exact same 0.51 validity coefficient for predicting job performance, making them the joint most effective single-method hiring tools in personnel psychology. Yet they differ in one critical respect: structured interviews measure what a candidate already knows how to do, while GMA tests measure how quickly they will learn what they do not. For corporate hiring managers allocating assessment budgets, that distinction dictates whether a new hire will scale with a rapidly changing role or plateau the moment their existing knowledge becomes obsolete.[1]
The 0.51 coefficient is the foundational metric of modern hiring science. It originates from Frank Schmidt and John Hunter’s 1998 meta-analysis in the Psychological Bulletin, which aggregated 85 years of personnel selection research. In industrial-organizational psychology, a validity coefficient operates on a scale from 0 to 1, where 1 represents perfect prediction of future job performance and 0 represents random chance. At 0.51, cognitive ability tests predict job performance more accurately than reference checks (0.26), years of job experience (0.18), and years of education (0.10) combined.[1]
The practical stakes of these figures materialize directly on a company's balance sheet. A selection method with a 0.51 validity coefficient generates massive economic utility by systematically filtering out candidates who will require excessive training or fail to execute complex tasks. When organizations rely instead on unstructured interviews—which carry a validity of just 0.38—they leave significant predictive power on the table, increasing the rate of costly mis-hires.[1][8]
The supremacy of the 0.51 figure faced a severe academic challenge in 2022. A paper published in the Journal of Applied Psychology by Paul Sackett and colleagues proposed slashing the GMA validity coefficient down to 0.31. The researchers achieved this lower number by omitting a statistical adjustment known as the "correction for range restriction." Because companies generally only hire candidates who score well on assessments, the pool of actual employees represents a restricted range of the total applicant population. Evaluating job performance only among high scorers artificially compresses the observed variance, making the test look less predictive than it actually is across a random sample.[2]
The academic rebuttal was swift and mathematically decisive. In 2023, researchers In-Sue Oh, Huy Le, and Philip Roth published a direct response in the same journal, demonstrating that omitting the range restriction correction fundamentally distorts the predictive reality of the tests. They proved the correction is mathematically warranted to estimate true applicant-pool validity. Following the exchange, Sackett and his co-authors published a reply endorsing the correction wherever applicant-pool data exist, restoring 0.51 as the consensus reference standard for GMA validity.[3][4]
The academic rebuttal was swift and mathematically decisive.
That predictive power scales linearly with the complexity of the work. The 0.51 figure represents a cross-job average, but the coefficient fluctuates based on the cognitive demands of the role. For highly complex professional and managerial positions, the validity of GMA rises to 0.58. For medium-complexity roles, it holds at 0.51, and for completely unskilled manual labor, it drops to 0.23. The underlying mechanism is straightforward: cognitive ability dictates the speed and efficiency of job knowledge acquisition. In roles where the environment is static, learning speed matters less; in roles requiring continuous adaptation, it is the primary bottleneck to performance.[1][5]
Despite the mathematical evidence, corporate adoption of GMA testing remains constrained by the risk of adverse impact. Cognitive ability tests consistently produce measurable score disparities across demographic groups. Research indicates that Black and Hispanic candidates score, on average, 0.7 to 1.0 standard deviations below white candidates. These disparities reflect systemic educational inequities and format disadvantages rather than innate ability, but under U.S. employment law, they trigger the Equal Employment Opportunity Commission's four-fifths rule. If a test screens out a protected class at a disproportionate rate, the employer must legally prove the assessment is a business necessity directly related to the job.[8]
This creates a legal paradox for human resources departments. The U.S. Equal Employment Opportunity Commission explicitly permits cognitive testing, provided the tests are validated for the specific position and administered consistently. "The use of any selection procedure which has an adverse impact on the hiring, promotion, or other employment or membership opportunities of members of any race, sex, or ethnic group will be considered to be discriminatory and inconsistent with these guidelines, unless the procedure has been validated in accordance with these guidelines," the EEOC states. Yet companies frequently abandon validated cognitive tests out of legal anxiety, replacing them with unstructured interviews. Unstructured interviews introduce significantly more human bias, predict performance poorly, and simply hide adverse impact behind closed doors rather than measuring it transparently.[8][9]
The most effective hiring architectures do not rely on GMA tests in isolation. Combining a cognitive ability test (0.51) with an integrity test (0.41) or a highly structured behavioral interview (0.51) pushes the overall predictive validity of the selection process to between 0.63 and 0.65. This combinatorial approach captures both the candidate's capacity to learn (GMA) and their behavioral disposition to apply that learning effectively (integrity and structure), while simultaneously diluting the demographic adverse impact produced by the cognitive test alone.[1][5]
Choosing the right selection method requires matching the tool to the specific constraints of the role. General mental ability testing fits well when hiring for complex, rapidly evolving positions where learning agility is paramount, provided the organization has the resources to validate the test and defend its business necessity. It does not fit well when used as a rigid, standalone cutoff score without behavioral context, or when hiring for highly standardized, low-complexity tasks where a simple work sample test would predict immediate proficiency with far less legal friction.[9]
Competing readings
General Mental Ability (GMA) Tests
The highest-predicting, lowest-cost screening method, measuring learning agility and problem-solving.
For: Unmatched predictive power for complex roles, yielding a 0.51 validity coefficient that scales up to 0.58 for professional and managerial positions. Against: High risk of adverse impact, with demographic score disparities often triggering legal scrutiny under the four-fifths rule. Evidence: 85 years of meta-analytic data confirm GMA is the single strongest predictor of both training success and long-term job performance.
Structured Interviews
Standardized behavioral and situational questions scored against a rigid rubric.
For: Equals GMA in predictive power (0.51 validity) while producing significantly less adverse impact across demographic groups. Against: Highly resource-intensive to design, train for, and administer at scale, requiring constant interviewer calibration. Evidence: When combined with a GMA test, structured interviews push the overall predictive validity of a hiring process to 0.65, the highest achievable threshold in personnel selection.
Unstructured Interviews
Conversational, free-flowing interviews based on interviewer intuition and rapport.
For: High candidate acceptance, zero upfront design cost, and universal familiarity among hiring managers. Against: Highly susceptible to affinity bias, poor at predicting actual job performance, and legally difficult to defend if challenged. Evidence: Despite being the most widely used selection method globally, unstructured interviews yield a validity coefficient of just 0.38, leaving massive predictive power on the table.
Years of Education & Experience
Traditional resume screens based on degree attainment and tenure in previous roles.
For: Easy to verify, objective, and deeply ingrained in corporate applicant tracking systems. Against: Extremely weak predictors of future performance in a new environment. Evidence: Years of job experience carry a validity coefficient of just 0.18, and years of education sit at 0.10, making them statistically inferior to almost any active assessment method.
- 0.51
- GMA predictive validity
- 0.38
- Unstructured interview validity
- 0.10
- Years of education validity
- 0.65
- GMA + structured interview validity
Sources
[1]ResearchGatePredictive Validity AdvocatesThe Validity and Utility of Selection Methods in Personnel Psychology: Practical and Theoretical Implications of 85 Years of Research Findings
Read on ResearchGate →
[2]J Appl PsycholPredictive Validity AdvocatesRevisiting meta-analytic estimates of validity in personnel selection: Addressing systematic overcorrection for restriction of range
Read on J Appl Psychol →
[3]J Appl PsycholPredictive Validity AdvocatesRevisiting Sackett et al.'s (2022) rationale behind their recommendation against correcting for range restriction in concurrent validation studies
Read on J Appl Psychol →
[4]J Appl PsycholPredictive Validity AdvocatesCorrecting for range restriction in meta-analysis: A reply to Oh et al. (2023)
Read on J Appl Psychol →
[5]ResearchGatePredictive Validity AdvocatesThe Validity and Utility of Selection Methods in Personnel Psychology: Practical and Theoretical Implications of 100 Years of Research Findings
Read on ResearchGate →
[6]J Appl PsycholPredictive Validity AdvocatesA meta-analytic study of general mental ability validity for different occupations in the European community
Read on J Appl Psychol →
[7]WileyPredictive Validity AdvocatesHistory and development of the Schmidt-Hunter meta-analysis methods
Read on Wiley →
[8]U.S. Equal Employment Opportunity CommissionAdverse Impact CriticsEmployment Tests and Selection Procedures
Read on U.S. Equal Employment Opportunity Commission →
[9]Factlen Editorial TeamHolistic Assessment ProponentsSynthesis by Factlen editorial team
Read on Factlen Editorial Team →
Comments
More in Business
See all →African Markets
Dangote Refinery IPO Aims to Raise $1.5 Billion in Landmark African Market Listing
4 sources
Resource-Based View
How Valuable, Rare, Inimitable, and Organized Resources Determine Sustained Competitive Advantage
7 sources
Corporate Accounting
Cash Basis vs. Accrual Basis: How Timing Revenue Recognition Shifts Tax Liability and Financial Reporting
7 sources
Inventory Strategy
Why a 5% Shift in Consumer Demand Triggers a 40% Swing in Manufacturing Output
6 sources
Every angle. Every day.
Get Business stories with full source coverage and perspective breakdowns delivered to your inbox.




