The Big Five personality test is a questionnaire that measures personality across five broad, independent traits: Openness to Experience, Conscientiousness, Extraversion, Agreeableness and Neuroticism, remembered by the acronym OCEAN. Each trait is a continuous scale, not a category, so your result is a position on five dimensions rather than a personality 'type'. It is the model most widely used in academic personality psychology, and the framework that sits underneath most work-based personality questionnaires employers send you. There are no right or wrong answers, but there are profiles that fit a given role better than others.
The Big Five (OCEAN) measures five traits (Openness, Conscientiousness, Extraversion, Agreeableness and Neuroticism) on continuous scales, scored as percentiles against a norm group rather than as pass/fail marks. Employers use it for role fit, team composition and interview probing. Answer as your consistent professional self, at reasonable speed, without inflating every answer: response-consistency and social-desirability scales are designed to flag profiles that look too good to be true.
What Is the Big Five Personality Test?
The Big Five emerged from decades of research into the words people use to describe each other. When researchers factor-analysed those descriptions across languages and cultures, the same five clusters kept reappearing. That is the model's core claim: almost any personality description you can write reduces, statistically, to five broad dimensions. Individual questionnaires then split each dimension into narrower facets, for example, Conscientiousness into orderliness, dutifulness and achievement striving.
Two features matter for candidates. First, the traits are dimensional: you are not 'an extravert' or 'an introvert', you sit somewhere on a scale, and most people sit near the middle. Second, the traits are largely independent: being high on Conscientiousness tells an employer almost nothing about your Openness. That independence is why a Big Five profile carries more usable information than a four-letter type label.
In hiring, you will rarely see a questionnaire called 'the Big Five'. What you will see is a commercial instrument (SHL's OPQ32, the Hogan Personality Inventory, Saville Wave, Korn Ferry's KF4D) whose scales map, more or less closely, onto the same five dimensions. If you understand OCEAN, you understand the machinery behind almost all of them. For the employer's-eye view of why they send these questionnaires at all, read our companion guide on <a href="/blogs/personality-tests-workplace-guide">personality tests in the workplace</a>; this article stays at trait level.
Openness to Experience
Openness describes intellectual curiosity, imagination, aesthetic sensitivity and appetite for novelty. High scorers are drawn to ideas, ambiguity and change; low scorers prefer the concrete, the proven and the familiar. Neither pole is 'better'. They are suited to different work.
- High Openness looks like: generating options nobody asked for, comfort with ambiguous briefs, interest in strategy and experimentation, tolerance for reorganisation and new tools.
- Low Openness looks like: preference for established procedure, scepticism about change for its own sake, focus on execution over exploration, reliability in stable processes.
- Roles that reward high scores: strategy, product, R&D, design, marketing, consulting, anything early-stage.
- Roles that reward moderate or lower scores: compliance, audit, quality assurance, safety-critical operations, high-volume process work.
- Common misreading: Openness is not intelligence. It correlates only modestly with cognitive ability, which is why employers still run a separate personality assessment alongside cognitive tests rather than treating one as a proxy for the other.
Conscientiousness
Conscientiousness covers organisation, dependability, self-discipline, planning and achievement striving. Across the research literature it is the single trait most consistently associated with job performance across occupations, which is why it is the trait employers scrutinise most closely and the one candidates most often try to inflate.
- High Conscientiousness looks like: deadlines met without chasing, structured planning, thorough checking, follow-through on commitments.
- Low Conscientiousness looks like: flexible improvisation, comfort switching priorities, less patience for documentation and process.
- Where very high scores can count against you: roles that demand rapid pivots can be poorly served by perfectionism and rigidity. Some instruments report facet scores precisely so recruiters can tell diligence apart from inflexibility.
- Why faking it here is risky: this is the trait every candidate knows to claim, so questionnaires include multiple differently-worded items on it. Inconsistent answers across those items are exactly what response-consistency scales are built to detect.
Extraversion
Extraversion describes sociability, assertiveness, energy and the tendency to seek stimulation from the outside world. It is the trait most often misunderstood as 'confidence'. An introvert can be entirely confident; they simply draw less energy from social contact and prefer depth over breadth of interaction.
- High Extraversion looks like: comfort speaking first in a group, energy from client contact, quick rapport, visible enthusiasm.
- Low Extraversion looks like: preference for written or one-to-one communication, deep focused work, listening before contributing.
- Roles that reward high scores: sales, business development, account management, recruitment, front-of-house leadership. Our sales assessment practice covers the questionnaires used for these roles.
- Roles where it barely matters: many technical, analytical, actuarial and engineering roles score assertiveness only as a leadership sub-signal, not as a hiring hurdle.
Agreeableness
Agreeableness captures cooperation, trust, empathy, modesty and a preference for harmony over conflict. It is the trait with the most non-linear relationship to job success: too low and you are abrasive, too high and you may struggle to challenge a colleague, hold a boundary or negotiate hard.
- High Agreeableness looks like: collaboration, generous credit-sharing, conflict avoidance, strong service orientation.
- Low Agreeableness looks like: direct challenge, comfort with disagreement, tough-minded negotiation, sceptical questioning.
- Roles that reward high scores: customer service, healthcare, HR, teaching, team-based delivery roles.
- Roles that reward mid-to-low scores: procurement, litigation, trading, turnaround management, any role where the job is to say no credibly.
- What recruiters look for: not the maximum, but the balance, enough warmth to work in a team, enough independence to disagree with a senior stakeholder.
Neuroticism (Emotional Stability)
Neuroticism is the tendency to experience negative emotions (anxiety, worry, frustration, self-doubt) and to react strongly to stress. Most work-based instruments invert it and report the positive pole, Emotional Stability or Resilience, because that language is fairer and less pathologising in a hiring context.
- High Emotional Stability (low Neuroticism) looks like: steady performance under deadline pressure, recovery after setbacks, calm in escalations.
- Lower Emotional Stability looks like: vigilance to risk and detail, higher sensitivity to criticism, more visible stress response under sustained load.
- Where employers weight it heavily: high-pressure client work, emergency services, aviation, trading floors, and leadership roles where volatility is contagious.
- Important caveat: this scale measures a personality tendency, not mental health, and it is not a clinical diagnosis. Employers using it in selection are also bound by data-protection and non-discrimination rules on how the result may be used.
Sit a work-style personality questionnaire built on the same trait structure employers use. The free preview gives you 4 questions per test plus 5 AI coach messages, no payment details needed.
Start the free personality testHow Employers Actually Use the Big Five
The most common mistake candidates make is assuming a personality questionnaire is a pass/fail gate like a numerical test. Usually it is not. It is used in four distinct ways, and knowing which one applies changes how much it matters.
- 1Role fit. Your profile is compared to a role profile the employer has defined, sometimes derived from their own high performers, sometimes from the publisher's generic competency model. Distance from that profile is what gets reported, not a raw score.
- 2Interview input. This is the biggest use and the one candidates underestimate. The report generates probing questions: 'Your profile suggests you prefer autonomy. Tell me about a time you had to follow a process you disagreed with.' A profile you cannot back up with real examples is the actual failure mode.
- 3Team composition. For assessment centres and leadership hiring, the panel looks at how your profile complements the existing team, not just at the absolute scores.
- 4Flag review. Response-consistency and social-desirability indicators are checked. A flagged profile is usually re-tested or discussed, not silently rejected, but it is a bad way to start a conversation.
The instruments themselves vary in length, format and how directly they map to OCEAN. These are the ones you are most likely to meet in a European or UK hiring process.
| Publisher | Questions | Time | Difficulty | Notes |
|---|---|---|---|---|
| SHL OPQ32 | 104 questions | 25-40 minutes | N/A | Forced-choice, 32 scales; results are commonly mapped back onto Big Five-style factors |
| Hogan HPI | 206 items | 15-20 minutes | N/A | 'Bright side' personality; the Hogan suite also includes HDS (derailers) and MVPI (values) |
| Saville Wave | 216 questions | About 40 minutes | N/A | Mixes rating and forced-choice; reports motives, talents and preferred culture |
| Thomas DISC | 24 questions | About 8 minutes | N/A | Behavioural style rather than Big Five traits; fast, widely used for shortlisting |
| Korn Ferry KF4D | Multi-part inventory | Typically 25-45 minutes | N/A | Traits, drivers, competencies and experience in one profile |
| IPIP-NEO (public domain) | 120 or 300 items | 15-40 minutes | N/A | Research instrument, not a hiring tool; useful for seeing facet-level Big Five output |
- Publisher deep-dives: Hogan, Saville and Thomas International: what each instrument measures and how its report is structured.
Big Five vs MBTI vs DISC
Candidates constantly conflate these three, and they are not the same kind of thing. The Big Five is a dimensional research model. MBTI is a type indicator built on a different theoretical tradition. DISC describes observable behavioural style. The practical differences are what determine how a result should (and should not) be used in selection.
| Publisher | Questions | Time | Difficulty | Notes |
|---|---|---|---|---|
| Big Five (OCEAN) | 5 continuous trait scales, usually with facets | 10-40 min depending on instrument | N/A | The dominant model in academic personality research; underpins most work questionnaires; results reported as percentiles vs a norm group |
| MBTI | 4 dichotomies, producing 16 types | About 20-30 min | N/A | Forces a binary on each dimension, so near-midpoint scores flip type; the publisher itself advises against using it for selection |
| DISC | 4 behavioural styles (D, I, S, C) | 5-15 min | N/A | Describes observable style at work rather than underlying traits; popular for team development and fast shortlisting |
| HEXACO | 6 traits: Big Five plus Honesty-Humility | 15-30 min | N/A | Adds an integrity-adjacent factor; used where counterproductive work behaviour is a concern |
| 16PF | 16 primary factors | About 35-50 min | N/A | Older narrow-trait model; its factors roll up into five global factors close to OCEAN |
Use MBTI for self-reflection and team language, DISC for communication style, and the Big Five when the question is 'how is this person likely to behave at work over time'. If an employer's questionnaire looks long and asks the same thing several different ways, you are almost certainly looking at a Big Five-derived instrument.
What a 'Good' Profile Looks Like per Role
There is no universally good profile. That is the entire point of the model. There are, however, recognisable role patterns. Treat these as directional, not as a scoring key, and remember that facet-level detail usually matters more than the headline trait.
- Graduate consulting / Big 4: high Conscientiousness, moderate-to-high Openness, mid-to-high Extraversion, mid Agreeableness (enough to challenge a client politely), high Emotional Stability.
- Sales and business development: high Extraversion, high Emotional Stability (rejection resilience), mid Agreeableness, Conscientiousness that shows up as pipeline discipline rather than perfectionism.
- Engineering, actuarial and data roles: high Conscientiousness, moderate Openness, Extraversion largely irrelevant, high Emotional Stability for on-call and deadline pressure.
- Customer service and healthcare: high Agreeableness, high Emotional Stability, high dependability facets of Conscientiousness.
- Leadership and management: high Emotional Stability, mid-to-high Extraversion (assertiveness facet especially), moderate Agreeableness, and Openness high enough to change course. See the leadership assessment category for the instruments used at this level.
- Compliance, audit and safety-critical operations: very high Conscientiousness, lower Openness, high Emotional Stability, mid Agreeableness.
How Personality Questionnaires Are Scored
Understanding the scoring removes most of the anxiety, because it makes clear what you can and cannot control.
- 1Your answers are aggregated into scale scores. Each trait is measured by many items, some of them reverse-worded, so a single 'wrong-looking' answer changes almost nothing.
- 2Raw scores are converted to percentiles against a norm group. The norm group is the reference population: graduates, managers, sales professionals, a national working population. The same raw answers can produce a 55th percentile in one norm group and a 75th in another, which is why a percentile is meaningless without knowing the comparison group.
- 3Scores are often reported as stens or stanines. A sten is a 1-10 band; 5 and 6 are the middle of the distribution. Most people are average on most traits, and that is not a problem.
- 4Format matters. A Likert questionnaire asks you to agree or disagree on a scale, and it is easy to answer, easy to inflate. A forced-choice (ipsative) questionnaire makes you pick the statement that is 'most like me' and the one that is 'least like me' from equally desirable options. Forced-choice is deliberately harder to game, which is why SHL's OPQ32 and similar instruments use it.
- 5Validity scales run in the background. Social-desirability, response-consistency and infrequency indicators are calculated from your answer pattern, not from any single item.
When every option looks like a strength ('I am well organised' versus 'I am creative' versus 'I stay calm') there is no safe answer. That is by design: the format removes the option of agreeing with everything positive, and it forces the questionnaire to measure your relative preferences rather than your willingness to self-promote.
Can You Fake a Personality Test? And Should You?
You can shift a Likert-based profile in a socially desirable direction. That is well documented, and test publishers have known it for decades. The relevant question is not whether faking is possible but whether it survives contact with the rest of the process. Usually it does not, for four reasons.
- 1Social-desirability scales. These are built from items that describe behaviour that is virtuous but statistically rare: never being irritated, never being late, never having told a white lie. Agreeing with all of them produces an implausibly saintly pattern and pushes the impression-management score into the flagged range.
- 2Response-consistency scales. The same trait is asked about in several differently-worded items scattered through the questionnaire. Answering the person you want to be rather than the person you are produces contradictions between those items, and consistency indices exist purely to catch that.
- 3Forced-choice formats. Ipsative questionnaires make 'agree with everything good' structurally impossible: you must rank desirable statements against each other, so inflation gets converted into a shape rather than a level.
- 4The interview. A faked profile creates the questions you will be asked. If the report says you are highly detail-oriented and structured, you will be asked for two examples, and 'the interview probes the report' is the mechanism that catches most successful fakers a stage later.
There is also the self-interested argument. A profile that gets you into a job that does not fit you is a slow, expensive failure. Probation periods exist. Employers are not looking for a mythical perfect candidate; they are looking for someone whose real working style matches the demands of the role, and a genuine mismatch is better discovered before you resign from your current job than six months after.
How to Prepare Honestly
- 1Read the job description as a trait brief. 'Thrives in a fast-changing environment' is Openness and Emotional Stability. 'Meticulous attention to detail' is Conscientiousness. 'Builds relationships across the business' is Extraversion and Agreeableness. You are not changing your answers. You are calibrating which version of yourself is relevant: the professional one, not the weekend one.
- 2Take a full practice questionnaire and read your own report. Most candidates have never seen their profile before an employer sees it. Knowing your own high and low scales in advance is the entire preparation advantage.
- 3Prepare evidence for your extremes. For every scale where you sit near the top or bottom, have one concrete work example ready. Those are precisely the scales the interviewer will probe.
- 4Answer at a steady pace. Overthinking single items is what creates inconsistency. Your first instinct is usually the most reliable, and most questionnaires are not tightly timed.
- 5Answer as your consistent professional self. Not your best day, not a fictional ideal. If you are asked how you behave, answer for how you behave at work, and answer that way for the whole questionnaire.
- 6Do not leave the questionnaire half-finished or rush the last third. A response pattern that changes character halfway through is one of the easiest anomalies to detect.
MakingMoves.ai covers 50+ test categories with 113,000+ practice questions, including personality questionnaires in both Likert and forced-choice formats, with an AI coach that explains what each scale is measuring. The free preview gives you 4 questions per test and 5 AI coach messages. Paid plans are GROW at EUR 19,95 per week, PRO at EUR 49,95 per month, and MAX at EUR 79,95 per month.
Our personality category mirrors the Likert and forced-choice instruments used by SHL, Hogan, Saville and Korn Ferry, with trait-level explanations on every scale.
Explore personality practiceFrequently Asked Questions
Not in the way you fail a numerical test. There is no pass mark and no correct answer. You can, however, be screened out for poor fit against a role profile, or have your results flagged as unreliable by consistency and social-desirability scales. The realistic failure modes are a bad fit and an implausible answer pattern, not a low score on a trait.
Yes. OCEAN is simply the acronym for the five traits: Openness, Conscientiousness, Extraversion, Agreeableness and Neuroticism. Some sources use CANOE for the same five, and many work-based instruments rename Neuroticism as Emotional Stability and reverse its direction.
Conscientiousness is the trait most consistently associated with performance across occupations in the research literature. The others matter more selectively: Extraversion for sales and leadership, Agreeableness for team and service roles, Emotional Stability for high-pressure work, Openness for change-heavy and creative roles.
Most work-based instruments run between 15 and 45 minutes. SHL's OPQ32 is typically 25-40 minutes, the Hogan HPI around 15-20 minutes, and Saville Wave around 40 minutes. Most are untimed or generously timed, because speed is not what is being measured.
No. Employers receive a report of scale scores, usually as percentiles or stens against a norm group, plus any validity flags and often a set of suggested interview probes. They do not review your item-by-item responses. Under the GDPR you can generally ask what data was processed and request feedback on the outcome.
Sometimes, but it depends on the employer, and retakes are far less common than for cognitive tests. Personality is meant to be stable, so a materially different second profile raises questions rather than answering them. The invitation email or the employer's FAQ is the only reliable source.
Yes, but for a different reason than a cognitive test. You are not training a skill. You are removing surprise: learning the forced-choice format, seeing your own profile before an employer does, and preparing the examples an interviewer will ask for based on your extreme scales.
Our assessment tips hub covers preparation plans, test-day checklists and format-specific strategy across all 50+ test categories.
Read the assessment tipsIngmar van Maurik is the founder of MakingMoves.ai and Assessment-Training.com. With 10+ years in psychometric assessment design, he has helped over 1 million professionals prepare for job assessments.