Leadership Isn’t Tenure. And It Isn’t an Assessment Score.

Jul 07, 2025

Leadership assessments are everywhere, and they promise a lot: identify who will thrive in complex roles and who won't. These tools can surface genuinely useful data. But the uncomfortable truth is that they rarely predict who will actually succeed in a leadership role. What they capture well are traits, preferences, and patterns that can inform development. What they can't capture is the situational, high-stakes, in-motion reality of leading.

Where assessments run into their limits

Start with context, because context is decisive. The same leader who thrives in a volatile startup may stumble in a stable enterprise, and the reason isn't a flaw in the leader — it's that leadership effectiveness is a product of the match between a person and a situation, not a fixed property of the person. This is one of the oldest and most durable findings in leadership research: there is no single best leadership style, and whether a given trait becomes a strength or a liability depends on the favorableness and demands of the specific context (Fiedler, 1967). An assessment scores the person. It cannot score the fit.

Then there's the gap between a static test and dynamic capability. Sitting in a simulated exercise or completing a self-report survey is not the same as navigating live ambiguity, where pressure, politics, and timing collide in real time. This matters for a technical reason worth being precise about. It is not that all assessment is worthless — general mental ability, structured interviews, and work samples do predict performance reasonably well. It's that the popular leadership instruments are among the weakest: personality relates to leadership only modestly, explaining a limited share of who leads well (Judge, Bono, Ilies, & Gerhardt, 2002), and widely used personality typologies tend to have weak predictive validity and shaky reliability. Even the strongest predictors capture potential, not enacted performance in context.

So the honest summary is not "assessments don't work." It's that their ability to forecast actual leadership performance is inconsistent, heavily moderated by context, and weakest precisely in the popular tools organizations lean on most. Assessments are diagnostic, not determinative. The mistake comes when organizations confuse data with destiny.

What leadership actually requires

Leadership is not granted by tenure, degrees, or a score. It's demonstrated in how someone behaves when conditions are unclear, the stakes are high, and people are watching. A handful of capabilities consistently matter most, and each corresponds to something real in the research.

The first is cognitive agility — the ability to process competing inputs, hold several possibilities at once, and still move forward. Strong judgment isn't a matter of always trusting analysis or always trusting instinct; it's knowing which the moment calls for, and knowing when intuition can be trusted at all, which depends on whether the domain offers valid, learnable patterns (Kahneman & Klein, 2009).

The second is decision-making under pressure — choosing clarity over paralysis when no playbook exists. This is genuinely hard, because pressure works against it: under threat, individuals and groups tend to narrow their thinking and revert to rigid, familiar responses at exactly the moment flexibility is most needed (Staw, Sandelands, & Dutton, 1981). And the decisions that most define leadership are usually the adaptive ones, where no existing expertise supplies the answer and the leader has to act into genuine uncertainty (Heifetz, 1994).

The third is relational intelligence — managing conflict, aligning stakeholders, and creating followership without coercion. This is skill, not charisma: the capacity to take others' perspectives (which, notably, power itself tends to erode unless a leader works at it) and to build the conditions in which people will speak honestly and move together.

The fourth is resilience in ambiguity — bringing structure and stability when external conditions are shifting. That is largely the work of sensemaking, turning a confusing, ambiguous situation into a coherent account others can act on (Weick, 1995), paired with the harder-won capacity to remain in uncertainty without grasping prematurely for a false resolution (Bion, 1970).

None of these are abstractions, and none are untrainable. They're observable, developable, and even testable — but in lived behavior, not on a standardized instrument.

The right role for assessments

This is not an argument against assessment; it's an argument for using it correctly. Used well, an assessment can spark valuable conversations about blind spots and developmental needs, offer a shared language for talking about leadership qualities, and provide data points that genuinely enrich a decision — when they are triangulated with structured interviews and a real track record rather than read in isolation. This is simply good practice: combining multiple methods yields far better insight than any single instrument, because each method adds information the others miss.

What an assessment should never be is a verdict. A scorecard doesn't decide who will hold the line in a crisis, create clarity in ambiguity, or move people to follow when the way forward is uncertain.

The bottom line

Leadership is not what's written on a résumé, and it's not the number an assessment prints out. It's the ability to step into uncertainty, make decisions that hold up under stress, and create the conditions in which others can move with you. That can't be captured in a multiple-choice format. It can only be seen in action — which is exactly where it should be looked for.


References

Bion, W. R. (1970). Attention and interpretation. Tavistock.

Fiedler, F. E. (1967). A theory of leadership effectiveness. McGraw-Hill.

Heifetz, R. A. (1994). Leadership without easy answers. Harvard University Press.

Judge, T. A., Bono, J. E., Ilies, R., & Gerhardt, M. W. (2002). Personality and leadership: A qualitative and quantitative review. Journal of Applied Psychology, 87(4), 765–780.

Kahneman, D., & Klein, G. (2009). Conditions for intuitive expertise: A failure to disagree. American Psychologist, 64(6), 515–526.

Schmidt, F. L., & Hunter, J. E. (1998). The validity and utility of selection methods in personnel psychology: Practical and theoretical implications of 85 years of research findings. Psychological Bulletin, 124(2), 262–274.

Staw, B. M., Sandelands, L. E., & Dutton, J. E. (1981). Threat-rigidity effects in organizational behavior: A multilevel analysis. Administrative Science Quarterly, 26(4), 501–524.

Weick, K. E. (1995). Sensemaking in organizations. Sage.

Stay connected with news and updates

Stay ahead with insight-driven leadership strategies that rewire thinking, enhance decision-making, and decode human dynamics.

Decode Human Dynamics. Rewire Thinking. Act with Clarity.
Close

50% Complete

Master Leadership Psychology. Make Smarter Decisions. Thrive Under Pressure.

The best leaders don’t just react—they think with precision, operate with clarity, and execute with confidence.

Subscribe to our Leadership Insights Newsletter and stay ahead of the curve with high-impact strategies designed for high-agency executives who play at the highest levels.