Anti-Faking Personality Tests: How Executive Assessments Detect and Prevent Response Distortion
When organizations hire executives or leaders, the stakes are extraordinarily high. A misaligned hire at the C-suite level can cost millions in lost productivity, damaged culture, regulatory risk, or reputational harm. Yet one of the most common threats to accurate executive assessment is deceptively simple: candidates faking good on personality tests.
Faking good—intentionally presenting yourself more favorably than reality—is not a rare exception. Research shows that 30 to 50 percent of job applicants admit to or demonstrate evidence of distorting their responses on personality assessments. In executive hiring, where the motivation to succeed is intense and the competitive pressure is relentless, faking becomes even more prevalent. The question isn't whether faking happens; it's whether your assessment tools can detect it and prevent it from invalidating your hiring decisions.
This glossary article explores the science, mechanisms, and practical implications of anti-faking personality tests. We'll examine how modern assessments detect response distortion, compare different formats designed to resist faking, and explain why anti-faking measures are essential for safe, evidence-informed executive assessment.
What Is Faking Good on a Personality Assessment?
Faking good—also called "faking favorable," "impression management," or "response distortion"—is the intentional manipulation of personality test responses to present oneself in a more favorable light than one's actual personality or behavior would warrant. It's a conscious, deliberate strategy to influence how others perceive you.
The distinction between faking good and other forms of response bias is important. Faking good is intentional: the respondent knows they're distorting their answers. In contrast, self-deception is unintentional. A person engaging in self-deception genuinely believes their inflated self-description; they've unconsciously reconstructed their self-image to be more positive than reality. Both distort assessment results, but they have different causes and different detection implications.
In the context of personality assessment, faking good typically manifests as endorsing items that present the respondent as more conscientious, agreeable, emotionally stable, or honest than they actually are. A candidate might claim "I am always punctual" when they're frequently late, or assert "I never lose my temper" when colleagues know otherwise. The motivation is straightforward: to improve the chances of being hired, promoted, or viewed favorably.
| Dimension | Faking Good (Impression Management) | Self-Deception |
|---|---|---|
| Intentionality | Deliberate, conscious | Unconscious, unintentional |
| Awareness | Respondent knows they're distorting | Respondent believes their distorted response |
| Motivation | Strategic (get job, improve image) | Protective (defend self-esteem) |
| Detectability | Can be flagged by validity scales | Often harder to detect |
| Prevalence in Hiring | 30–50% of applicants | Variable, context-dependent |
| Example | Answering "I'm always punctual" when not | Genuinely believing "I'm punctual" despite tardiness |
Historical Context and the Evolution of Anti-Faking Research
The problem of faking on personality assessments didn't emerge overnight. During the 1980s and 1990s, industrial-organizational psychologists began systematically studying whether and how often candidates distorted their responses in hiring contexts. The findings were sobering. Dwight and Donovan's 2003 meta-analysis revealed that up to 63 percent of applicants admitted to faking on personality tests, while Baer and Miller (2002) estimated the actual rate of faking-good behavior at approximately 30 percent of applicants—a substantial minority.
What made these findings particularly significant was that they contradicted earlier assumptions. Prior to this research, some assessment professionals believed that faking was either impossible (because people couldn't accurately manipulate their self-perceptions) or inconsequential (because faking didn't meaningfully alter the validity of personality predictions). The empirical evidence forced a reckoning: faking was both possible and common, and it could distort assessment results in ways that mattered for hiring decisions.
In executive and leadership contexts, the motivation to fake is even stronger. A candidate competing for a VP-level position, a C-suite role, or a board seat faces enormous pressure to present an idealized version of themselves. The financial stakes are higher, the competition is fiercer, and the perceived consequences of honest self-disclosure (admitting to a weakness, acknowledging a limitation) feel more severe. This is why anti-faking mechanisms are particularly critical in executive assessment.
Why Do People Fake on Personality Tests?
Understanding why people fake is the first step toward designing assessments that resist faking or detecting it when it occurs. The answer isn't that people are inherently dishonest. Rather, faking is a normal social behavior that intensifies in high-stakes situations.
The Psychology of Impression Management
Sociologist Erving Goffman famously described social life as a kind of theatrical performance. In everyday interactions, people manage the impressions others form of them—they dress for the occasion, adjust their language to the context, and emphasize certain aspects of their personality while downplaying others. This "impression management" is not pathological; it's a fundamental part of human social functioning.
In most contexts, impression management is mild and context-appropriate. You might be more formal in a job interview than at a casual dinner with friends, but you're still recognizably yourself. However, in high-stakes situations—particularly hiring for competitive, high-status roles—the motivation to manage impressions intensifies dramatically. The candidate's goal is no longer simply to make a good impression; it's to make the best possible impression, even if that means exaggerating strengths or minimizing weaknesses.
Psychologists have identified two universal human motives that drive impression management: getting along with others (cooperation) and getting ahead of others (competition). In hiring, both are active. Candidates want to be liked (appear agreeable, conscientious, emotionally stable) and to be seen as capable and impressive (appear confident, driven, successful). These motivations are so powerful that they can override the instruction to "be honest" on a personality test.
Situational Factors That Amplify Faking
Not all situations create equal faking pressure. Several factors determine whether a candidate is likely to distort their responses:
The stakes of the selection decision. The higher the stakes—in terms of salary, status, or career impact—the stronger the motivation to fake. A candidate applying for an entry-level position may answer honestly; a candidate competing for a C-suite role is far more likely to strategically distort responses.
Perceived job relevance. Candidates fake most on traits they believe are relevant to job success. If a candidate believes that conscientiousness is critical for the role they're applying for, they'll exaggerate their conscientiousness. If they think emotional stability matters, they'll minimize evidence of stress or anxiety. This is why faking patterns often vary by job type: leadership candidates fake heavily on confidence and decisiveness, while candidates for detail-oriented roles fake on conscientiousness and meticulousness.
Transparency of the assessment. If candidates understand what the test is measuring and what the "good" answers are, they're more likely to fake. Conversely, if the assessment is opaque—if candidates can't figure out what each item is measuring—faking becomes harder. This is one reason why forced-choice personality tests, which obscure the "correct" answer, are more resistant to faking than Likert-scale tests, where the desirable response is often obvious.
Perceived likelihood of detection. If candidates believe their faking will be detected, they may be less likely to fake. Some research suggests that simply informing candidates that the assessment includes validity scales designed to detect faking can reduce faking motivation. Others, however, may interpret this as a challenge and attempt to fake more strategically.
Individual Differences in Faking Tendency
Not all candidates fake equally. Some are more prone to faking than others, and research has identified several individual differences that predict faking tendency.
Cognitive ability. Smarter candidates may be better at faking strategically. They can more easily identify what the test is measuring, predict what the "good" answer is, and construct responses that are consistent across items. This creates a troubling paradox: the most cognitively capable candidates may be the most effective fakers, potentially biasing hiring decisions in their favor.
Personality traits. Candidates low in honesty and integrity are more likely to fake. So are those high in Machiavellianism (a tendency toward strategic manipulation) or narcissism (an inflated sense of self-importance). Interestingly, candidates who are naturally high on the traits being faked (e.g., naturally conscientious people faking conscientiousness) may find it easier to fake convincingly, because their faking is consistent with genuine aspects of their personality.
Self-awareness. People with lower self-awareness may fake more unconsciously. They may not realize they're distorting their responses; instead, they may genuinely overestimate their own conscientiousness or emotional stability. This blurs the line between faking and self-deception.
Motivation and stakes perception. Candidates who perceive the hiring decision as more consequential are more likely to fake. This is why executive candidates, who have more to gain (and potentially lose), tend to show higher faking rates than entry-level candidates.
How Do Personality Tests Detect Faking?
The science of detecting faking has evolved substantially over the past three decades. Modern personality assessments employ multiple mechanisms to identify response distortion, ranging from traditional validity scales to sophisticated forced-choice formats and real-time response analysis. Understanding these mechanisms is essential for interpreting assessment results and making informed hiring decisions.
Validity Scales: The Traditional Anti-Faking Approach
Validity scales are the oldest and most widely used anti-faking mechanism. The principle is elegant: identify items that are endorsed infrequently in the normal population but frequently by people who are faking good, then use the frequency of these responses as a signal of possible distortion.
Here's how they work in practice. Psychometricians develop a pool of items that describe positive qualities almost universally endorsed by normal respondents—things like "I have never told a lie," "I always keep my promises," or "I never get angry." In a normal, honest sample, these items are endorsed by only a small percentage of people (say, 5 percent). But when someone is faking good, they endorse many of these items, because they're trying to present themselves as virtuous and flawless.
A validity scale counts how many of these "infrequently endorsed desirable" items a respondent selects. If the count is unusually high, it suggests the respondent may be faking. For example, if 95 percent of a normal sample endorses "I sometimes make mistakes" but only 20 percent of a test-taker's responses suggest they make mistakes, that pattern is suspicious and might indicate faking.
Validity scales are included in most major personality assessments used in hiring. The MMPI-2 includes the Lie (L) Scale and the K Scale; the Personality Assessment Inventory (PAI) includes Positive Impression Management (PIM) and Negative Impression Management (NIM) scales; the NEO-PI-R includes Positive Presentation and Negative Presentation validity scales. These scales have been extensively researched and are well-established in the field.
However, validity scales have significant limitations. The most important is construct contamination: validity scales correlate with actual personality traits. Someone who is genuinely conscientious, agreeable, and morally principled will naturally endorse many of the items on a "Lie Scale," without faking at all. This creates false positives—authentic candidates flagged as fakers—and can lead to unfair rejection of high-integrity candidates who simply have high integrity.
Additionally, some research suggests that mechanically "correcting" personality scores based on validity scale elevations can actually reduce the validity of the personality measure itself. When you adjust scores downward based on a validity scale, you may be removing true personality information along with the noise from faking.
| Assessment | Validity Scale Name | What It Measures | Threshold for Concern |
|---|---|---|---|
| MMPI-2 | Lie (L) Scale | Overly virtuous responding | T > 65 |
| MMPI-2 | K Scale | Defensive, guarded responding | T > 65 |
| PAI | Positive Impression Management (PIM) | Faking good | T > 70 |
| PAI | Negative Impression Management (NIM) | Faking bad / malingering | T > 70 |
| NEO-PI-R | Positive Presentation | Overly favorable self-description | > 1 SD above mean |
| HPI (Hogan) | Prudence subscales (Moralistic, Mastery, Virtuous) | Impression management indicators | Elevated combined score |
| HDS (Hogan) | Infrequency scale | Unusual or random responding | Elevated |
Forced-Choice Personality Assessments: A More Resistant Format
In recent decades, researchers have developed an alternative to traditional Likert-scale personality tests: forced-choice personality assessments, often abbreviated as MFC (multidimensional forced-choice) or FC formats. These assessments work on a fundamentally different principle and show significantly greater resistance to faking.
In a forced-choice format, respondents don't rate their agreement with statements on a scale. Instead, they choose between two, three, or four options that are equally attractive (or equally unattractive). For example, instead of rating "I prefer working with people" on a 5-point scale, a respondent might choose between "I prefer working with people" and "I prefer working with data"—both of which are neutral, positive qualities depending on the job.
The genius of forced-choice is that it removes the "obviously desirable" response option. On a Likert scale, a candidate applying for a leadership role can see that "I am a natural leader" is the good answer and select it, whether or not it's true. On a forced-choice test, both options might be leadership-relevant but in different ways, making it harder to identify the "correct" answer and thus harder to fake strategically.
Research on forced-choice formats has been overwhelmingly positive. A comprehensive meta-analysis by Martínez and Salgado (2021) examined the faking resistance of forced-choice inventories across multiple studies and found that forced-choice formats show significant resistance to faking behavior. The effect sizes tell the story: when candidates are instructed to "fake good" on a Likert-scale personality test, the effect size for Conscientiousness is δ = 1.27 (a very large effect, meaning faking dramatically changes scores). When the same construct is measured with a quasi-ipsative forced-choice format, the faking effect size drops to δ = 0.49—a 60 percent reduction.
Interestingly, the meta-analysis also found that faking effects are larger in laboratory experiments (where researchers explicitly instruct candidates to "fake good") than in real-world hiring contexts. This suggests that in actual job applications, candidates may not fake as intensively as laboratory studies suggest, perhaps because the effort required is substantial or because they fear detection.
However, forced-choice formats come with trade-offs. The most significant is ipsative scoring. In a traditional Likert-scale test, scores are norm-referenced: your score on Conscientiousness is compared to a normative sample, and you can score high on multiple dimensions simultaneously. In an ipsative forced-choice test, scores are relative: if you score high on Conscientiousness, you necessarily score lower on other dimensions, because you've been choosing between them. This makes interpretation more complex and can make it harder to identify candidates who are high on multiple desirable traits.
Additionally, forced-choice formats are less familiar to most candidates, which can create confusion or resentment. Some candidates report finding forced-choice tests frustrating because neither option feels like a good fit. And while the research on forced-choice formats is growing, it's still less extensive than the research on traditional Likert-scale personality tests, so some organizations are hesitant to adopt them without additional validation.
Item-Level Detection and Advanced Methods
Beyond validity scales and format design, modern personality assessments employ additional sophisticated methods to detect faking:
Response time analysis. People who are faking may take longer to respond to items because they're deliberating about which answer will make them look best. Conversely, they may rush through items without fully reading them. By analyzing response time patterns—both absolute times and consistency across items—assessments can flag potentially faked responses.
Inconsistency indices. Fakers sometimes contradict themselves across items that measure the same construct. For example, a respondent might claim "I am very organized" on one item and "I often lose track of details" on another. Inconsistency indices flag these contradictions and signal possible faking. The MMPI-2 includes VRIN (Variable Response Inconsistency) and TRIN (True Response Inconsistency) scales for this purpose; the PAI includes an Inconsistency Index.
Item Response Theory (IRT) approaches. Some assessments use Item Response Theory, a sophisticated statistical framework, to model the probability of a particular response pattern given a respondent's true level on a trait. If a response pattern is extremely unlikely given the respondent's overall trait level, it signals possible faking or careless responding. IRT-based detection can be more nuanced than simple validity scales because it considers the full pattern of responses rather than just counting "infrequent desirable" endorsements.
Machine learning and AI-based detection. Emerging research explores whether machine learning algorithms can identify faking patterns that human-designed validity scales might miss. Monaro et al. (2021) examined whether ChatGPT could outperform humans in faking personality assessments while avoiding detection—a fascinating question about the future of assessment. While AI-based detection is still experimental, it represents a frontier in anti-faking research.
Forced-Choice vs. Likert Scales: Which Resists Faking Better?
For organizations designing or selecting personality assessments, the choice between Likert-scale and forced-choice formats is consequential. Both have advantages and limitations, and the right choice depends on the specific context and priorities.
Likert Scale Format: The Traditional Approach
Likert scales have dominated personality assessment for decades. In a Likert format, respondents rate their agreement with statements on a numerical scale—typically 1 to 5 or 1 to 7, ranging from "Strongly Disagree" to "Strongly Agree."
Advantages: Likert scales are intuitive and familiar to most respondents. People understand what they're being asked to do, and the scoring is straightforward: higher numbers mean higher agreement. The resulting scores are absolute, not relative, which makes interpretation simpler. A person can score high on multiple dimensions simultaneously (e.g., high on both Conscientiousness and Extraversion). Additionally, Likert-scale personality tests have been extensively researched, so there's a robust body of evidence on their reliability and validity.
Faking vulnerability: The downside is that Likert scales are highly vulnerable to faking. The "good" answer is often obvious. If you're applying for a leadership role and the test asks "I am a natural leader," you know that "Strongly Agree" is the desirable response. A candidate faking good will select it, regardless of whether it's true. This is why Likert-scale personality tests show the large faking effect sizes documented in research (δ = 1.27 for Conscientiousness).
Forced-Choice Format: The Modern Anti-Faking Approach
Forced-choice personality assessments represent a newer paradigm, designed specifically to reduce faking vulnerability.
Advantages: As discussed earlier, forced-choice formats significantly reduce faking because neither option is obviously "better." They also reduce social desirability bias more broadly. Additionally, by forcing respondents to choose, they reduce the temptation to respond neutrally or give non-committal answers. The result is more decisive, more forceful personality data.
Limitations: The primary limitation is ipsative scoring, which makes interpretation more complex. Additionally, forced-choice tests are less familiar to most respondents, which can create confusion or frustration. Some candidates report that forced-choice tests feel artificial because they're forced to choose between options that don't feel like good fits. And while research on forced-choice formats is growing, it's still less extensive than research on Likert-scale tests, so some organizations worry about construct validity or applicability to their specific context.
| Dimension | Likert Scale | Forced-Choice (MFC) |
|---|---|---|
| Format | Rate agreement on 5-point scale | Choose from equally attractive options |
| Faking Resistance | Low to moderate | High |
| Effect Size of Faking | δ = 1.27 (Conscientiousness) | δ = 0.49 (quasi-ipsative) |
| Scoring Type | Norm-referenced (absolute) | Ipsative (relative) |
| Interpretability | Straightforward, absolute scores | Complex, requires ipsative interpretation |
| Respondent Familiarity | Very high | Lower |
| Validity Research | Extensive, well-established | Growing, but less extensive |
| Real-World Faking Rate | 30–50% of applicants | 10–20% estimated (lower resistance needed) |
| Best Use Case | General personality assessment, development | High-stakes hiring, executive selection |
| Cost/Complexity | Lower | Moderate to higher |
Research Evidence and Meta-Analysis
The comparative research is clear. Martínez and Salgado's (2021) meta-analysis examined forced-choice inventories across multiple studies and contexts. The findings were consistent: forced-choice formats show significantly greater resistance to faking than Likert scales. The effect size for Conscientiousness faking was reduced by 60 percent (from δ = 1.27 to δ = 0.49) when using quasi-ipsative forced-choice formats.
Importantly, the meta-analysis also found that faking effects are substantially larger in experimental laboratory contexts (where researchers explicitly instruct participants to "fake good") than in real-world applicant samples. This suggests that the faking problem, while real, may be somewhat less severe in actual hiring than laboratory studies suggest. Real candidates may not fake as intensively, perhaps because the cognitive effort required is high or because they fear detection.
The trade-off is clear: forced-choice formats reduce faking but increase complexity. For high-stakes hiring where faking is a major concern—such as executive selection—the trade-off often favors forced-choice. For lower-stakes assessment or development contexts, the simplicity and familiarity of Likert scales may be preferable.
Validity Scales and Anti-Faking Mechanisms Explained
For professionals using personality assessments in hiring, understanding how validity scales and anti-faking mechanisms work is essential for correct interpretation and fair decision-making.
How Validity Scales Work: The Science
The principle underlying validity scales is straightforward: identify items that are endorsed infrequently in the normal population but frequently by people who are faking good, then use frequency of endorsement as a signal of possible distortion.
Consider a specific example. The item "I have never told a lie" is endorsed by perhaps 5 percent of a normal, honest sample. Most people acknowledge that they've told at least small lies or white lies at some point. But when someone is faking good—trying to present themselves as virtuous and flawless—they're more likely to endorse this item, because it contributes to their idealized self-presentation.
A validity scale counts how many of these "infrequently endorsed desirable" items a respondent selects. If the count is unusually high, it suggests the respondent may be faking. For example, if a respondent endorses 30 out of 40 "infrequent-desirable" items, while the average normal respondent endorses only 8, that's a red flag. The validity scale score is elevated, suggesting possible faking.
Importantly, an elevated validity scale score doesn't automatically invalidate the test. It's a signal, not a verdict. A skilled assessment professional will use an elevated validity scale score as a trigger for further investigation: conducting a follow-up interview, checking references, or administering additional assessments to clarify whether the candidate is truly faking or whether they're an authentic high-performer who happens to be genuinely conscientious and morally principled.
Types of Validity Scales and What They Measure
Different assessments employ different validity scales, each designed to detect specific types of response distortion:
Lie/Virtuous Responding Scales. These detect overly positive, moralistic, or virtuous responding. The MMPI-2 Lie (L) Scale and the PAI Positive Impression Management (PIM) Scale are examples. They flag respondents who claim to be almost unrealistically virtuous, honest, and well-adjusted.
Defensive/Guarded Responding Scales. These detect attempts to minimize problems, pathology, or weaknesses. The MMPI-2 K Scale and the PAI Defensiveness Scale are examples. They're particularly useful in clinical contexts where people might be motivated to hide symptoms, but they're also relevant in hiring, where candidates might downplay interpersonal difficulties or emotional volatility.
Infrequency Scales. These detect unusual, random, or careless responding. The MMPI-2 F Scale and the PAI Infrequency Scale count endorsements of items that are endorsed by fewer than 10 percent of the normal population. If someone endorses many of these unusual items, it suggests they're either not reading carefully, not understanding the items, or responding randomly.
Inconsistency Scales. These flag contradictory responses across items that measure the same construct. If a respondent claims to be highly organized on one item but admits to being disorganized on another, that's inconsistent. The MMPI-2 includes VRIN (Variable Response Inconsistency) and TRIN (True Response Inconsistency) scales; the PAI includes an Inconsistency Index. These scales help identify respondents who are responding carelessly or inconsistently, which can signal faking, confusion, or lack of effort.
The Problem with Validity Scales: Construct Contamination
Despite their widespread use, validity scales have a significant limitation: they correlate with actual personality traits. This creates a fundamental problem called construct contamination.
Consider the MMPI-2 Lie Scale. It's designed to detect faking good by flagging overly virtuous responding. However, the items on the Lie Scale also correlate strongly with Conscientiousness and Agreeableness—actual personality traits. A person who is genuinely conscientious, agreeable, and morally principled will naturally endorse many of the items on the Lie Scale, without faking at all.
This creates a troubling false positive rate. A genuinely high-integrity executive—someone whose personality genuinely reflects conscientiousness and moral principles—may score high on a validity scale designed to detect faking, and thus be flagged as a faker. This is unfair and can lead to the rejection of excellent candidates.
Research by Ones and colleagues (1998) examined this issue directly. They found that when personality scores are mechanically "corrected" or adjusted downward based on elevated validity scale scores, the construct validity of the personality measure itself can be reduced. In other words, by trying to remove the "noise" from faking, you may remove true personality information as well.
This is why modern assessments don't rely on validity scales alone. Instead, they use multiple indicators: validity scales, plus forced-choice format, plus inconsistency indices, plus response time analysis. By triangulating across multiple methods, assessments can better distinguish genuine high-performers from fakers.
Anti-Faking Personality Tests in Executive Assessment and Leadership Hiring
The stakes of executive assessment are extraordinarily high. A misaligned hire at the C-suite level—a CEO, CFO, Chief Risk Officer, or Board member who misrepresented their capabilities, integrity, or risk profile—can cost an organization millions of dollars and expose it to regulatory, legal, or reputational risk. This is why anti-faking mechanisms are not a luxury in executive assessment; they're a necessity.
Why Anti-Faking Matters in Executive Assessment
The motivation to fake is intense at the executive level. A candidate competing for a C-suite role, a board seat, or a senior leadership position is competing against other highly qualified candidates for a position of enormous status and financial reward. The pressure to present an idealized version of oneself is relentless.
Additionally, executives have the cognitive sophistication to fake strategically. They understand organizational dynamics, they've likely taken personality assessments before, and they can often predict what the "good" answers are. A bright, ambitious executive can craft responses that are consistent, plausible, and aligned with what the organization is looking for—all while being fundamentally misaligned with their actual personality or risk profile.
The consequences of invalid assessment are severe. An executive who faked conscientiousness during the hiring process but is actually disorganized and unreliable can derail organizational initiatives and demoralize teams. An executive who faked emotional stability but is actually volatile and reactive can create a toxic culture. An executive who faked integrity but is actually willing to cut corners can expose the organization to legal or regulatory risk. These aren't minor hiring mistakes; they're organizational crises.
This is why organizations serious about executive assessment invest in tools and processes specifically designed to detect and prevent faking. Anti-faking mechanisms—validity scales, forced-choice formats, multi-method assessment—are essential safeguards.
Leadership Traits Most Vulnerable to Faking
Not all personality traits are equally vulnerable to faking. Some are more obvious targets for strategic distortion, particularly in executive contexts:
Integrity and Conscientiousness. Executives universally claim high integrity and conscientiousness during hiring, because these traits are universally valued. However, actual integrity and conscientiousness vary widely. Some executives are genuinely principled and reliable; others cut corners or make excuses. Detecting faking on these dimensions is critical, because integrity and conscientiousness are strong predictors of leadership effectiveness and organizational outcomes.
Emotional Stability. Leaders claim to be calm, composed, and resilient under stress. Yet some executives are actually volatile, reactive, or prone to anxiety. Faking emotional stability is easy because the "good" answer is obvious. However, detecting genuine emotional stability is important because emotionally unstable leaders can create toxic cultures, make poor decisions under pressure, and derail organizational performance.
Agreeableness and Interpersonal Style. Executives often fake agreeableness—claiming to be collaborative, team-oriented, and sensitive to others' needs. Yet some executives are actually domineering, dismissive, or interpersonally harsh. This faking is particularly consequential because interpersonal style directly affects team dynamics, talent retention, and organizational culture.
Dominance and Assertiveness. Conversely, some executives minimize their dominance or assertiveness during assessment, fearing they'll be seen as aggressive or difficult. Yet confidence and assertiveness are often necessary for executive effectiveness. Detecting genuine (versus faked) levels of dominance and assertiveness is important for matching candidates to roles.
Best Practices for Anti-Faking in Executive Assessment
Organizations serious about reducing faking in executive assessment should follow these evidence-based practices:
Use multi-method assessment. Don't rely on personality assessment alone. Combine personality tests with structured interviews, reference checks, simulation exercises, and work history analysis. Each method has different faking vulnerabilities, so multiple methods provide triangulation and reduce the likelihood that faking on one method will go undetected.
Employ forced-choice or hybrid formats. If your organization is selecting a personality assessment, prioritize forced-choice or hybrid formats over traditional Likert scales. The research is clear: forced-choice formats significantly reduce faking vulnerability. While they're more complex to interpret, the investment in training is worthwhile for high-stakes hiring.
Include validity scales and inconsistency indices. Ensure that your assessment includes multiple indicators of response distortion: validity scales, inconsistency indices, response time analysis. Use these indicators as signals for further investigation, not as automatic disqualifiers.
Provide context and transparency. Interestingly, research suggests that informing candidates that the assessment is designed to detect faking can actually reduce faking motivation. Candidates who understand that the assessment includes validity scales may be less likely to attempt faking, because they perceive the likelihood of detection as high. Transparency about the assessment's purpose can be a tool for reducing faking.
Use assessments validated on executive/leadership samples. General population norms may not apply to executive contexts. Ensure that your assessment has been validated on leadership or executive samples, so that interpretation is calibrated to the executive population.
Train assessment users. Even the best assessment is only as good as the person interpreting it. Ensure that your HR professionals and hiring managers understand how to interpret validity scales, recognize false positives, and use assessment results in conjunction with other information.
Common Misconceptions About Personality Test Faking
Despite decades of research, significant misconceptions about personality test faking persist. Clarifying these misconceptions is essential for making good hiring decisions.
Misconception #1: "Personality Tests Are Easy to Fake, and Nothing Can Stop It"
Reality: Faking is possible, but modern anti-faking mechanisms significantly reduce its effectiveness. Forced-choice formats reduce faking effect sizes by 60 percent or more. Validity scales, inconsistency indices, and response time analysis provide additional detection mechanisms. While no assessment is 100 percent faking-proof, well-designed assessments make faking costly in terms of effort and detectability.
Implication: Organizations should not resign themselves to accepting faking as inevitable. Investing in assessments with strong anti-faking mechanisms and using them as part of a multi-method approach can substantially reduce faking's impact on hiring decisions.
Misconception #2: "If Someone Fakes, Their Entire Test Result Is Invalid"
Reality: An elevated validity scale doesn't automatically invalidate results. It signals caution, but doesn't automatically reject the candidate. In practice, elevated validity scales often trigger further investigation—a follow-up interview, reference check, or additional assessment—rather than automatic disqualification.
Implication: Validity scale elevations should be used as a starting point for inquiry, not as a final verdict. A skilled assessment professional will investigate further before making a hiring decision based on a validity scale elevation.
Misconception #3: "Honest People Never Score High on Validity Scales"
Reality: Genuinely conscientious, agreeable, or morally-minded people will often score high on "Lie Scales" or similar validity scales without faking. This is the construct contamination problem discussed earlier. Validity scales correlate with actual personality traits, creating false positives.
Implication: Don't automatically reject candidates with elevated validity scales. Use multiple indicators and context (interview, references, behavioral data) to distinguish authentic high-performers from fakers. Some of your best candidates may have elevated validity scales because they're genuinely high on conscientiousness and integrity.
Misconception #4: "Forced-Choice Formats Are Completely Faking-Proof"
Reality: Forced-choice formats significantly reduce faking, but don't eliminate it entirely. Sophisticated fakers can still identify patterns in forced-choice items and manipulate responses. Additionally, forced-choice formats introduce new complexities (ipsative scoring, interpretation difficulty) that require careful handling.
Implication: Forced-choice formats are a significant improvement over Likert scales for faking resistance, but they're not a complete solution. Use them as part of a comprehensive assessment strategy, not as a standalone guarantee against faking.
Misconception #5: "Only Dishonest People Fake on Personality Tests"
Reality: Faking is a normal, everyday social behavior (impression management). Even honest, ethical people fake in high-stakes hiring contexts. The motivation is situational (desire to get the job), not a reflection of dishonesty in daily life. A person might fake good on a personality test for an executive role while being completely honest in other contexts.
Implication: Detecting faking is not about identifying "bad" or dishonest people. It's about distinguishing the presented self from the actual self. A candidate who fakes on a personality test isn't necessarily unethical; they may simply be responding to the high stakes of executive hiring. The goal of anti-faking mechanisms is to see past the presentation and understand the actual person.
The Future of Anti-Faking Personality Assessment
The field of personality assessment is evolving rapidly. New technologies, methodologies, and research insights are shaping the future of anti-faking mechanisms.
AI and Machine Learning in Faking Detection
Machine learning approaches to faking detection represent a frontier in assessment science. Rather than relying on human-designed validity scales or format innovations, machine learning algorithms can analyze vast amounts of response data to identify patterns associated with faking.
Monaro et al. (2021) explored an intriguing question: can ChatGPT outperform humans in faking personality assessments while avoiding detection? This research highlights both the promise and the peril of AI in assessment. On one hand, AI-based detection could identify faking patterns that human-designed validity scales miss. On the other hand, AI-based faking could become more sophisticated, making it harder to detect.
The future likely involves an arms race between AI-based faking and AI-based detection, with assessment designers continuously updating algorithms to stay ahead of faking strategies.
Ecological Momentary Assessment and Real-Time Personality
An alternative approach to reducing faking is to move away from self-report personality tests altogether and instead measure behavior in real time. Ecological Momentary Assessment (EMA) involves repeated sampling of behavior in real-world contexts, often via smartphone-based prompts.
Rather than asking a candidate "Are you conscientious?" and relying on their self-report, EMA would ask them to report their actual behavior throughout the day: "Did you complete your work on time? Did you organize your workspace? Did you follow through on commitments?" By sampling actual behavior, EMA reduces faking opportunity because it's harder to fake real-time behavior than to fake self-perception on a personality test.
The limitation is practical: EMA is invasive and raises privacy concerns. It's not yet viable for hiring, but it represents a potential future direction for personality assessment.
Hybrid and Adaptive Formats
The future of personality assessment likely involves hybrid approaches that combine multiple formats and methods. An assessment might include some Likert-scale items, some forced-choice items, some behavioral questions, and some consistency checks. By combining formats, assessments can reduce faking on each individual format while maintaining interpretability and respondent acceptance.
Adaptive testing—where the difficulty or format of items adjusts based on responses—is another emerging approach. If a respondent's pattern suggests faking on Likert items, the assessment might shift to forced-choice format. This adaptability could make faking more difficult while maintaining engagement and reducing assessment length.
Ethical and Regulatory Developments
As personality assessment becomes more prevalent in hiring, regulatory and ethical scrutiny is increasing. The EEOC, professional psychology organizations, and employment law are all evolving to address concerns about fairness, bias, and transparency in personality assessment.
Future developments will likely include:
- Greater transparency requirements: Organizations may be required to disclose what assessments measure, how they're scored, and how results are used in hiring decisions.
- Fairness and bias audits: Assessments will need to demonstrate that they don't discriminate against protected groups or reduce accessibility for candidates with disabilities.
- Candidate rights: Candidates may gain the right to see their results, understand how they were interpreted, and appeal decisions based on assessment results.
- Validity standards: Assessments will need to demonstrate validity specifically in hiring contexts, not just in research settings. Generic personality tests may face challenges if they can't show job-related validity.
These developments will likely make personality assessment more rigorous, more transparent, and more fair—but also more complex and more regulated.
Frequently Asked Questions
Can you fake a personality test?
Yes, to some degree. Respondents can deliberately present themselves more favorably than reality. However, modern anti-faking mechanisms—validity scales, forced-choice formats, inconsistency indices, and response time analysis—significantly reduce faking effectiveness. Research shows that forced-choice formats reduce faking effect sizes by 60 percent or more compared to traditional Likert scales. While no assessment is completely faking-proof, well-designed assessments make faking costly in terms of effort and detectability.
How do personality tests detect faking?
Personality tests detect faking through multiple mechanisms: (1) Validity scales that count endorsements of infrequently endorsed desirable items—high counts suggest faking; (2) Forced-choice formats that remove obviously desirable response options, making it harder to identify the "correct" answer; (3) Inconsistency indices that flag contradictory responses across items measuring the same construct; (4) Response time analysis that identifies unusually long deliberation times or rushed responding; and (5) IRT-based approaches that model the probability of response patterns and flag unlikely patterns.
What is faking good on a personality assessment?
Faking good—also called "impression management" or "response distortion"—is the intentional distortion of personality test responses to present oneself in a more favorable light than one's actual personality would warrant. It's a conscious, deliberate strategy motivated by desire to succeed in hiring or other high-stakes contexts. Research estimates that 30 to 50 percent of job applicants show evidence of faking good on personality tests.
Are forced-choice personality tests better at detecting faking?
Yes, substantially. Forced-choice personality tests are significantly more resistant to faking than traditional Likert-scale tests. A meta-analysis by Martínez and Salgado (2021) found that forced-choice formats reduce faking effect sizes by approximately 60 percent. For example, the faking effect size for Conscientiousness drops from δ = 1.27 (Likert scale) to δ = 0.49 (quasi-ipsative forced-choice). The trade-off is that forced-choice formats are more complex to interpret and less familiar to respondents.
How common is faking on personality assessments?
Research estimates that 30 to 50 percent of job applicants show evidence of faking or admit to intentional distortion on personality tests. The rate varies depending on the stakes of the selection decision, the transparency of the assessment, and individual differences in faking tendency. In executive hiring, where stakes are highest, faking rates are likely at the higher end of this range.
What are validity scales in personality tests?
Validity scales are items designed to detect response bias by measuring infrequently endorsed desirable responses. Examples include the MMPI-2 Lie (L) Scale, the PAI Positive Impression Management (PIM) Scale, and the NEO-PI-R Positive Presentation scale. If a respondent endorses many "infrequently desirable" items, the validity scale score is elevated, signaling possible faking or response distortion. Elevated validity scales don't automatically invalidate results but signal need for further investigation.
What is impression management in personality testing?
Impression management is the intentional process of controlling how others perceive you, often by presenting a more favorable version of yourself. It's a normal social behavior that intensifies in high-stakes situations like job interviews or executive hiring. In personality testing, impression management manifests as deliberate distortion of responses to present oneself as more conscientious, agreeable, emotionally stable, or honest than one actually is. The motivation is situational (desire to get the job or advance one's career), not necessarily a reflection of dishonesty in daily life.
How can I avoid faking on a personality test?
Be honest and authentic in your responses. Remember that personality assessments are designed to measure how you actually are, not how you wish to be. Faking is often detected by validity scales and inconsistency indices, and detected faking can backfire, leading to rejection or loss of credibility. Additionally, if you do get the job based on faked responses, there's a mismatch between your presented self and your actual self, which can lead to poor job fit and dissatisfaction. Authenticity is the best strategy for both getting hired and succeeding in the role.
What is response distortion in personality assessments?
Response distortion is any systematic deviation from truthful responding on a personality assessment. It includes faking good (impression management), faking bad (malingering or minimization), and unintentional distortions like self-deception (where respondents genuinely but inaccurately overestimate themselves). Response distortion can reduce the validity of personality assessment results and lead to poor hiring decisions. Modern assessments include multiple mechanisms to detect and reduce response distortion.
Do anti-faking mechanisms affect test validity?
Well-designed anti-faking mechanisms improve validity by reducing noise from response distortion. Forced-choice formats, for example, reduce faking effect sizes while maintaining or improving predictive validity. However, validity scales themselves can correlate with actual personality traits (construct contamination), which can reduce validity if scores are mechanically adjusted based on validity scale elevations. The key is to use anti-faking mechanisms as signals for further investigation, not as automatic disqualifiers. Modern assessments use multiple indicators and require skilled interpretation to balance faking detection with validity.