![]()
CliftonStrengths, Reviewed: Where It Came From, What It Measures and What the Evidence Says
- 26 million people. 34 themes. 177 paired statements, 20 seconds each, and your five highest themes ranked.
- The origin deserves respect. Don Clifton asked what was right with people. That question changed the industry for good.
- Read their own numbers. Re-test reliability is about 0.70 – so a “Top 5” can shuffle between sittings.
- The evidence splits in two. Gallup's outcome figures are large. Independent meta-analysis finds modest, real effects.
- Not for hiring – their words. Gallup says it plainly: not designed or validated for selection. Credit for that.
What's the most generous question anyone has asked in the history of management? My nomination goes to a University of Nebraska educational psychologist named Donald O. Clifton, who spent the 1950s watching his discipline catalogue everything wrong with people and asked, in effect: what would happen if we studied what was right with them instead?
That question grew into the CliftonStrengths assessment - formerly StrengthsFinder - which more than 26 million people had taken by 2022, making it one of the most widely used workplace assessments ever built. I've spent much of my career in the field it helped create, so this review is written with respect. It's also written with the psychometric record open on the desk, because an instrument taken by that many people deserves to be examined as an instrument, and the marketing rarely lingers on the technical detail. Here's what it measures, how it measures it, and what the evidence says - cited as I go, so you can check everything for yourself.
Where it came from
Clifton (1924–2003) was a decorated wartime pilot who returned to Nebraska, took degrees in mathematics and educational psychology, and taught and researched there until 1969. He then founded Selection Research, Incorporated - a business built on studying high performers through structured interviews - and in 1988 SRI acquired the Gallup Organisation, the polling institution whose name it took. Over the following decade, Clifton's team distilled the patterns from those interviews into a catalogue of what they called talent themes, and in 1999 the Clifton StrengthsFinder went online.
The idea reached the wider world in 2001, when Marcus Buckingham and Clifton published Now, Discover Your Strengths. Tom Rath's follow-up, StrengthsFinder 2.0 (2007), became one of Amazon's ten best-selling books and stayed there for the best part of a decade. Along the way the American Psychological Association gave Clifton a Presidential Commendation as "the father of strengths-based psychology and the grandfather of positive psychology" - a fair summary of his influence, since the positive psychology movement that followed owes him its founding instinct.
That instinct deserves to be stated in full, because it was a real act of humanity in a harsh genre. Workplace assessment in the mid-twentieth century was largely a deficit hunt: find the gaps, write the development plan, repeat. Clifton's insistence that excellence is better understood by studying excellence changed the emotional register of an entire industry. Whatever else this review finds, that contribution stands.
What it actually measures
What does CliftonStrengths actually measure?
CliftonStrengths measures 34 talent themes - naturally recurring patterns of thought, feeling or behaviour - grouped into four domains and ranked from a 177-item paired-statement survey, with 20 seconds per pair. Gallup distinguishes a talent (the raw pattern) from a strength (consistent near-perfect performance): the assessment measures the former and infers your route to the latter. Your five highest themes arrive as Signature Themes.
CliftonStrengths measures what Gallup calls talent themes - naturally recurring patterns of thought, feeling or behaviour. There are 34 of them, with names that have entered workplace vocabulary: Achiever, Woo, Learner, Deliberative, Ideation. The 34 sit in four domains - Executing, Influencing, Relationship Building and Strategic Thinking - and every person who completes the assessment receives their themes ranked, with the top five presented as "Signature Themes".
Can CliftonStrengths be used for hiring?
No - and Gallup says so plainly in its own technical documentation: the instrument is not designed or validated for employee selection or mental health screening, and comparing profiles between people is discouraged. It is a development tool, aimed at the person who took it. Publishers do not always draw their own boundaries so clearly; this one should be acknowledged.
Two definitional points matter more than they first appear. First, Gallup distinguishes a talent (the raw pattern) from a strength (the ability to deliver consistent, near-perfect performance), with the formula that talent multiplied by investment produces strength. The assessment measures the former and infers your best route to the latter. Second - and to Gallup's credit, stated plainly in its own technical documentation - the instrument is explicitly not designed or validated for employee selection or mental health screening, and comparisons between people's profiles are discouraged. It is a development tool, aimed at the person who took it. Publishers do not always draw their own boundaries so clearly, and this one should be acknowledged.
The uniqueness claim is central to the appeal: Gallup notes that the chance of two people sharing the same top five in the same order is about one in 33 million, and its own analysis of 8.4 million respondents found 233,282 distinct top-five combinations. Your result feels like yours alone, and statistically it very nearly is.
How the survey works
The assessment presents 177 paired self-descriptors - for example, "I get to know people individually" against "I accept many types of people". For each pair you choose which statement describes you better, and how strongly, on a five-point line between the two anchors. A timer gives you 20 seconds per pair before the system moves on - a deliberate design choice intended to capture instinctive rather than considered responses. The whole exercise takes 30 to 45 minutes. A proprietary formula then scores each of the 34 themes, and your report ranks them using percentiles drawn from Gallup's database of more than ten million respondents.
For readers who follow questionnaire design, the format is worth locating precisely. Assessment items come in three broad families: ipsative items force a choice between options; Likert items let you agree or disagree in steps; a visual analogue scale lets you answer anywhere along a continuous line. CliftonStrengths is a hybrid: the pairing is a forced choice, but the five-point response between anchors softens it. Gallup's technical report states that fewer than 30 per cent of items are ipsatively scored and concludes that ipsativity does not compromise interpretation. The 20-second limit is the more unusual choice - defensible as a guard against self-presentation, though it also means a distracted moment becomes part of your permanent profile.
One design decision sits oddly with the philosophy. The instrument's premise is that people are not fixed types, and its output is a nuanced rank order of 34 continuous scores. Yet what most people remember, share and print on their email signature is a list of five labels. The reduction from 34 shaded scores to five named labels is where a spectrum quietly becomes a box - a pattern worth noticing in any assessment, not just this one.
The psychology underneath
CliftonStrengths did not grow out of a formal trait theory. Its 34 themes were derived empirically - distilled from the patterns Clifton's interviewers heard across decades of structured conversations with high performers - and the theoretical framing arrived largely afterwards, through the positive psychology movement Clifton helped inspire. That origin is neither disqualifying nor unusual, but it does mean the deepest question about the instrument is how its themes relate to the personality science that came before and after it.
Gallup's own research answers part of that. A study reported in the technical report correlated theme scores with the Big Five personality factors across 1,462 respondents and found substantial overlap in places: the Communication theme correlated 0.71 with Extraversion, Woo 0.69; Discipline correlated 0.60 with Conscientiousness; Strategic and Ideation correlated 0.65 and 0.64 with Openness. In plain terms, several themes are recognisable Big Five territory wearing warmer names. Gallup presents this as evidence of construct validity, and it is - though it also invites the question of what the additional 34-theme vocabulary adds beyond familiarity and charm. The answer, I think, is real but modest: the language is the innovation. "Woo" starts conversations that "high extraversion" never has.
Does it hold up? The technical record
Is CliftonStrengths scientifically valid?
Partly - and Gallup publishes the numbers, which deserves credit. Internal consistency runs 0.52 to 0.79 across the 34 themes, several below the conventional 0.70 line. Test-retest reliability averages about 0.70, so results move between sittings. Several themes overlap heavily with Big Five factors measured for decades - Communication correlates 0.71 with Extraversion. Most validation is authored by Gallup researchers; independent reviewers find the instrument engaging and useful for development, with limited evidence for stronger claims.
Gallup publishes its psychometrics, which again deserves credit - many commercial instruments publish less. The headline figures from the technical report, for readers who want them:
Internal consistency (how coherently each theme's items hang together) ranges from 0.52 to 0.79 across the 34 themes - Input and Relator at the bottom, Woo and Discipline at the top. Values above 0.70 are conventionally considered adequate; a number of themes sit below that line, which Gallup attributes to the small number of items per theme.
Can your CliftonStrengths top five change?
Yes. Test-retest reliability averages around 0.70 at the full-profile level, and individual themes range from roughly 0.48 to 0.82 - so themes sitting sixth and twelfth today can trade places next quarter, and some of your top five can change without anything being wrong with you or the instrument. Gallup advises treating results as prompts for reflection rather than identity.
Test-retest reliability - whether you get the same result twice - averages around 0.70 at the full-profile level across one-, three- and six-month intervals. Individual themes range from roughly 0.48 (Maximizer) to 0.82 (Woo). A 0.70 is respectable for a personality-style measure. It also means real movement between sittings: with 34 themes packed into a ranked list, themes sitting sixth and twelfth today can trade places next quarter, and some of your "top five" can change without anything being wrong with you or the instrument. Gallup's response - that the assessment captures enduring tendencies and that results should prompt reflection rather than define identity - is reasonable, and sits a little awkwardly with how permanently those five labels tend to follow people around.
Independent scrutiny is the thinner file. Most published validation is authored or co-authored by Gallup researchers, which is standard practice for commercial instruments but means the strongest claims rest on the publisher's own studies. Independent reviews - such as the assessment guides and academic commentary summarised by PositivePsychology.com - tend to reach a consistent verdict: the instrument is engaging, face-valid and useful for development conversations, while its evidence base for anything stronger remains limited.
Does it work? The efficacy evidence
Here the evidence splits into two piles worth keeping separate: Gallup's own outcome research, and independent studies of strengths-based development in general.
Gallup's pile is substantial and impressive on its face. Its 2016 meta-analysis of strengths-based development covered 49,495 business units and 1.2 million employees across 22 organisations, 45 countries and seven industries, and reported that units receiving strengths interventions saw sales up 10 to 19 per cent, profit up 14 to 29 per cent and engagement up 9 to 15 per cent, with sharp falls in turnover and safety incidents. Two cautions belong alongside those numbers: they describe the higher-performing end of the distribution studied, and the research was conducted and published by the company that sells the intervention. That doesn't make it wrong. It makes it a claim awaiting independent replication.
The independent pile is smaller and more sober. A 2021 systematic review and meta-analysis in Frontiers in Psychology (Björk, Bolander and Forsman) examined bottom-up workplace interventions and found that strengths-based approaches produced a statistically significant but modest improvement in work engagement - a standardised effect of 0.34, the strongest of the intervention families tested. The authors also noted that only 12 of the 31 studies reviewed met their highest quality standard. So the fairest one-sentence summary of the independent evidence reads: strengths-based development does something real, the something is modest rather than transformational, and the research could be better.
Two further critiques come from inside the strengths tradition rather than outside it. Kaplan and Kaiser's Harvard Business Review work, "Stop Overdoing Your Strengths" (2009), documented how strengths pushed too hard become liabilities - the decisive leader who steamrolls, the supportive one who avoids hard calls. And Tomas Chamorro-Premuzic argued in HBR that an exclusive focus on strengths can be counterproductive: high performers grow by developing new capabilities, not only by amplifying existing ones, and a weakness ignored does not stop mattering. Neither critique demolishes the strengths idea. Both put a boundary around it that the poster versions leave off.
The commercial model
How much does CliftonStrengths cost?
The assessment costs $24.99 for a Top 5 report or $59.99 for the full ranked 34, with role-specific reports at $49.99. Around it stands a larger paid ecosystem - books, courses, team tools and a global network of Gallup-certified coaches - which is where much of the model's revenue sits. Modest sums individually; considerable at organisational scale.
The assessment costs $24.99 for a Top 5 report or $59.99 for the full ranked 34, with role-specific reports at $49.99 - modest sums individually, considerable at organisational scale. Around it stands a large paid ecosystem: books, courses, team tools, and a global network of Gallup-certified strengths coaches, with more than 700 schools and universities and, at the 2015 count, 467 of the Fortune 500 using the approach. None of this is improper - instruments cost money to build and maintain. It is simply worth seeing clearly that the certification-and-coaching layer is where much of the model's revenue and advocacy lives, and that an accredited interpreter between you and your own results is a design choice, not a necessity of measurement.
What to take from it
Is CliftonStrengths worth it?
For development conversations, often yes: the language is warm and memorable, and the strengths movement moved workplace conversation from deficit to capability. Independent evidence finds strengths-based approaches produce a real but modest improvement in work engagement - a standardised effect of 0.34. The limits are equally real: a top five less stable than its cultural permanence suggests, and a one-off, static result in a working life that refuses to stand still.
A fair verdict, then. CliftonStrengths earns its place in history: it moved workplace conversation from deficit to capability, it publishes its technical evidence, it draws its own boundary against selection use, and its language gives millions of people a warm, memorable way to talk about difference. Those are real achievements, and the first one changed the industry for good.
Its limits are equally real: reliability figures that make the famous top-five list less stable than its cultural permanence suggests; heavy overlap with personality factors measured elsewhere for decades; efficacy evidence that is strong from the publisher and modest from independents; and a one-off, static result in a working life that refuses to stand still. The most searching question about any assessment is not whether it's insightful on the day - most are - but what happens six months later, when you've changed and the report hasn't.
Before you trust any assessment: six questions to keep
Copy, paste and use these on any instrument, ours included:
- Does it measure something changeable, or does it hand me a permanent label?
- Are the reliability figures published - and would the result survive a re-test next quarter?
- Is it validated for how my organisation actually intends to use it?
- Can I understand my results without paying an accredited interpreter?
- What does the publisher say the tool must never be used for - and is that boundary built in, or just policy?
- When was the evidence last checked by someone who doesn't sell the instrument?
Where our thinking sits
Readers of this site will know where Sariio stands: we think the strengths movement asked the right question and settled on the wrong unit - that the working day runs on movable preferences rather than ranked labels, and that any reading of a person should expect them to change. That argument is made across our other essays, and it owes an obvious debt to the man from Nebraska who first asked what was right with people.