top of page

Research Lab

Leadership, and Governance by Design.

Cada imagen de este portafolio forma parte de un método de investigación visual donde la arquitectura se convierte en especulación. Estas obras no son renders finales, sino atmósferas conceptuales: estudios interpretativos del espacio, la materia y la coexistencia. A través de la visualización generativa, Poiétika explora cómo las ciudades pueden respirar, recordar y transformarse: la arquitectura como hipótesis viva.

Hardwired to Overlook One Another
 
From the Captain of Guam to the Algorithm — Fifty Thousand Years of the Same Tribal Instinct

Author: Ricardo Serrano Ayvar
Architect · Capital Projects Director

Date: July 28, 2026

Series: Leadership and Governance by Design — Vol. 1, No. 1 (2026)

Category: Transdisciplinary Research — Diagnostic Essay

Section: Essays / Working Papers

License

This work is licensed under the Creative Commons Attribution–NonCommercial–NoDerivatives 4.0 International License (CC BY-NC-ND 4.0).

How to cite

Este documento es un working paper de investigación independiente, de divulgación transdisciplinaria. No ha sido sometido a arbitraje por pares ni forma parte de una publicación indexada. Se comparte como parte de una serie de investigación en desarrollo, con el objetivo de generar discusión pública antes de una eventual formalización académica.

Serrano Ayvar, R. (2026). Hardwired Not to Recognize One Another: From the Captain of Guam to the Algorithm — Fifty Thousand Years of the Same Tribal Instinct [Working paper]. Governance and Leadership by Design Series, Vol. 1, No. 1. ricardoserranoayvar.com

Other citation formats

Serrano Ayvar, R. (2026). Programados para no reconocer a los demás: Del capitán de Guam al algoritmo — cincuenta mil años del mismo instinto de tribu [Working paper]. Serie Gobernanza y Liderazgo por Diseño, Vol. 1, N.º 1. ricardoserranoayvar.com

APA 7: Serrano Ayvar, R. (2026). Hardwired not to recognize one another: From the captain of Guam to the algorithm — Fifty thousand years of the same tribal instinct. [Working paper]. Governance and Leadership by Design Series, 1(1). ricardoserranoayvar.com

Chicago (author-date): Serrano Ayvar, Ricardo. 2026. "Hardwired Not to Recognize One Another: From the Captain of Guam to the Algorithm — Fifty Thousand Years of the Same Tribal Instinct." Working paper, Governance and Leadership by Design Series, vol. 1, no. 1. ricardoserranoayvar.com.

MLA 9: Serrano Ayvar, Ricardo. "Hardwired Not to Recognize One Another: From the Captain of Guam to the Algorithm — Fifty Thousand Years of the Same Tribal Instinct." Governance and Leadership by Design Series, vol. 1, no. 1, 2026, ricardoserranoayvar.com.

Keywords

Organized by thematic clusters, not in alphabetical order — this reflects the essay structure, not a simple list.


Leadership and Followership

Leadership · Followership · Epistemic authority · Transformational leadership · Situational leadership · Servant leadership · Adaptive leadership · Complexity leadership theory · Positional command

​​

Recognition and Cognitive Bias

Implicit leadership theory · Mental prototype · Dominance bias · Babble hypothesis · Dunning-Kruger effect · Sensemaking · Romance of leadership · Role congruity theory · Backlash effect

Ecosystem and Organizational Structure

Organizational culture · Psychological safety · Power and organizational politics · Role ambiguity · Ecological systems theory · Reverse dominance hierarchy · Governance of the commons

Organizational Violence and Dysfunction

Lateral violence · Upward bullying · Abusive supervision · Moral disengagement · Bad apples · Workplace aggression · Workplace incivility

Selection and Talent Management Person-organization fit · Cultural fit / cultural add · KSAO model · Predictive validity · Work sample evidence · Voluntary turnover · Boundaryless career · Culture mold · Evidence-based hiring

​​

Psychology, Personality, and Neurodivergence

Evolutionary psychology · Evolutionary mismatch · Growth mindset · Neurodivergence · Personality traits · Narcissistic personality disorder · Antisocial personality disorder · System justification theory

Systems Design and Governance

Systems architecture · Organizational design · Complex systems theory · Organizational governance · Authentic leadership · Organizational theatricality · Algorithmic amplification

Executive Summary

Hardwired Not to Recognize One Another From the Captain of Guam to the Algorithm — Fifty Thousand Years of the Same Tribal Instinct

Leadership is not a problem of individual character. It never was. In 1997, Korean Air Flight 801 crashed in Guam after two crew members saw the captain's approach error and flagged it — "don't you think it's raining harder?" — without enough clarity for him to process it in time. The accident wasn't the sum of two individual errors. It was an emergent property of an entire sociotechnical system that never learned to recognize, in time, the judgment of whoever held less formal authority. Nine thousand flight hours weren't enough that night. Not for lack of character. For lack of a system capable of recognizing the otherness of whoever was in front of him — not a mistaken reading waiting to be corrected, but another legitimate way of seeing what he, from his own instrument, couldn't see.

The essay maps seven leadership styles with verified authorship — from positional command to epistemic authority, the original synthesis: leading by showing the reasoning instead of imposing presence — Kelley's five followership patterns, and a twenty-five-cell matrix crossing both, where only one turns out genuinely generative. The finding underneath it all: people don't recognize leadership by observing behavior neutrally, they recognize it against a fixed mental prototype — firm voice, volume, speaking time — that systematically confuses performative confidence with real judgment. That's loudest-voice leadership: MacLaren and colleagues' "babble hypothesis" proves that whoever speaks longest, not best, emerges as leader, and the Grant, Gino, and Hofmann study itself shows that bias produces worse business results exactly where the most exemplary followers are available to listen. There's a sister trap, older than any office: theatricality. Visible activity, projected confidence, constant presence on the front stage — guaranteed smoke, with none of those three variables predicting whether there's real substance behind them.

That bias wasn't born in any modern corporation. It's an evolutionary inheritance: we're still evaluating authority with the same instrument once used to pick a tribal chief fifty thousand years ago, and the "primitivization" of organizational systems isn't that we went back to tribal patterns — it's that we never fully left. Today that ancient instinct has an accelerant no one designed with that intent: platforms promise connection with diverse voices, with genuine otherness, and the mechanism that sustains them economically produces, at planetary scale and in milliseconds, exactly the opposite — the same old tribal instinct, with better production values.

And here's something almost never said out loud: the genuinely authentic person almost always gets canceled. Not with a firing — with the silent wear of the backlash effect, with measured discomfort toward whoever simply performs better without asking permission, with lateral violence from peers who'd rather not have such a demanding mirror nearby. It happens the same way to the leader with epistemic authority and to the exemplary follower who questions with independent judgment — both get punished for the same reason: their authenticity sets a standard the system never agreed to sustain. And a widespread confusion is worth dismantling: "strong" and "soft" character aren't a gradient of more or less leadership. A soft character can be epistemic authority with better documented results than its louder counterpart. A strong character can just as often be pure smoke. Confusing volume with substance is, almost always, the same trap under a different name — and evaluation schemes that rely on charisma instead of evidence aren't just blind to genuinely toxic profiles: they frequently prefer them actively, because they project exactly the confidence that scheme confuses with "leadership material" (Babiak & Hare) — when what they bring, documented, isn't order, it's relational violence and waves of negative consequences.

The real counterexample: NUMMI, Fremont, California, 1984. The same workers who, under General Motors, never dared pull the emergency cord, out of fear of punishment, learned under Toyota's system — with time and sustained evidence that punishment no longer followed — to use it. Quality rose, with that, to levels that same workforce had never reached before. Proof, with name and date, that the system — not character — decides whether the right information arrives in time.

The implications land on every actor: for HR, the unstructured interview has the lowest predictive validity of any known method (0.38), far below actual work-sample evidence (0.54). For headhunters, role ambiguity — confusing "Principal" with "Job Captain," sometimes with a third layer of poorly specified language on top — guarantees failed processes before they even start. For the Board and senior leadership, the implicit prototype they carry internalized decides, with no written policy, who they promote. And on the generic suspicion toward non-linear career paths: clinically toxic profiles are a real minority (6.2% and 3.6%, per available epidemiology), so reading ambiguity as pathology is, almost always, the same prototype bias, in darker vocabulary.

The close offers no formula. It hands the question back: how many times we've been the captain and not the co-pilot — and whether we're willing to design, on top of the wiring we all carry, a system capable of recognizing in time whoever holds a different and valid reading, even without the rank to be heard.

Hardwired Not to Recognize One Another

From the Captain of Guam to the Algorithm — Fifty Thousand Years of the Same Tribal Instinct

There's a question most people would rather not ask themselves out loud: how many times have we been, without knowing it, the captain and not the co-pilot. How many times has someone told us something with all the clarity their position allowed, and we didn't listen because it didn't fit the idea we already had of ourselves. It's not a comfortable question. It doesn't have a quick answer either. And there's something even more uncomfortable behind it: it's not just that we sometimes fail to recognize one another. It's that, to a significant degree, we come hardwired not to — by biology, by culture, and now also by platform design. That is, at bottom, the complete argument that sustains everything that follows.

I. Guam, 1997 — The Captain Didn't See What He'd Already Been Told

On August 6, 1997, Korean Air Flight 801 was approaching Guam's airport amid heavy rain and reduced visibility. The airport's glide slope system — the one that tells a crew the correct angle of descent — had been out of service for weeks, a documented fact that in practice required the crew to reconstruct that information using other instruments and their own judgment, in real time, in the rain, at night.

There were three people in the cockpit: the captain, with more than nine thousand flight hours, a spotless record on paper, internal commendations for seniority and performance; the first officer, with fewer hours and a clearly subordinate position in the hierarchy; the flight engineer, third in the chain of command. In the minutes before impact, the black box recordings — later analyzed in near-forensic detail by the U.S. National Transportation Safety Board — show something investigators would take years to interpret with precision. The first officer, seeing that the approach didn't match what was expected, said: "Don't you think it's raining harder?" Moments later, the flight engineer noted, almost in passing, that the glide slope system wasn't available.

Neither observation was a direct challenge. Neither included the word "danger." They were, in the cultural register both men had been trained in, the correct way — the respectful way, the way that doesn't call a captain's authority into question — to flag a concern to a superior.

The captain didn't adjust course in time. The plane crashed into the side of Nimitz Hill, three kilometers from the runway. 228 of the 254 people aboard died.

Why didn't he listen? That's the question that sustains everything that follows.

When Malcolm Gladwell reconstructed this case in Outliers, drawing on the work of aviation researcher Suren Ratwatte and on the "power distance" framework of Dutch social psychologist Geert Hofstede — one of the five cultural dimensions Hofstede identified while studying how different societies distribute and accept inequality of authority — what he documented wasn't a negligent captain or an incompetent crew. It was a communication architecture in which the social cost of contradicting an authority figure was, in the everyday practice of that cockpit, higher than the risk of not doing so with sufficient clarity. The co-pilot saw the problem. He said something. He said it the only way his system had taught him was acceptable to say it. And that system, at three hundred kilometers an hour over a Pacific island, proved lethal.

It's worth naming with precision what kind of explanation this is, because it completely changes the frame for everything that follows. The accident wasn't the sum of two individual errors — a captain who didn't listen, a co-pilot who didn't speak clearly enough. It was an emergent property of the entire sociotechnical system: culture, hierarchy, training, protocol, and technology interacting in a way that none of the three crew members, on their own, could have produced or corrected alone. Hunting for individual culprits inside that cockpit — the classic error of almost any superficial accident investigation — is exactly the same error of analytical level that this article is going to name again and again in the corporate world: asking "what's wrong with this person?" when the right question lives a level higher, in the system that made it possible, or impossible, for the right information to reach whoever could act on it in time.

There's a second reading, less often cited, that matters more for what this article will propose later on. The captain had the formal authority to decide. What he didn't have — what no protocol of that era had given him — was a system capable of recognizing that the most valuable information in that cockpit, at that exact moment, wasn't coming from whoever held the highest rank. It was coming from whoever had looked at the right instrument. Authority and relevant knowledge were not, that night, the same person. And the whole system only knew how to listen to the first one.

Korean Air wasn't an isolated case within its own history. Between 1970 and 1999 it had a fatal accident rate several times higher than the average for comparable major Western airlines, to the point that Delta Air Lines and Air France temporarily suspended their code-share agreements with the Korean carrier, and Lloyd's of London raised its premiums significantly. In pure actuarial terms, the airline was technically on the edge of operational unviability without deep structural intervention.

The change that finally worked wasn't "hire better pilots," or "fire whoever was responsible for the Guam accident," or any of the reflexive responses a pressured board of directors tends to produce when something goes catastrophically wrong. David Greenberg, an American executive from Delta Air Lines hired in 2000 to lead the transformation of Korean Air's flight operations, imposed a complete systems overhaul, with several mutually reinforcing pieces. The cockpit's operating language became English — breaking, without that being its explicit purpose, much of the honorific hierarchy that formal Korean imposes grammatically between superior and subordinate, something English simply doesn't build into its structure. Crew Resource Management protocol — originally developed by a joint NASA and major U.S. airline team following a series of accidents in the seventies that showed exactly the same underlying pattern — was adopted rigorously and audited. And captains were explicitly trained to actively invite disagreement from their crew at the start of every flight — a brief, almost liturgical verbal ritual — while co-pilots were trained to use standardized phrases — "Captain, I consider this situation unsafe" — specifically designed to trigger an immediate cognitive alarm, leaving no room for the cultural ambiguity that had cost 228 lives at Guam.

In a little over a decade, Korean Air went from being among the world's worst-rated airlines for safety to joining the SkyTeam alliance in 2000 and, years later, earning Skytrax's seven-star rating, the highest possible. They didn't mass-replace the people. They changed the system of relationship between whoever leads and whoever follows — and they changed what counted as a valid signal that something was wrong, regardless of how well that signal matched the prior image of what a trustworthy captain should recognize as a threat.

That is the complete thesis of this article, applied to terrain far more everyday than a flight deck, but governed by exactly the same psychological and structural mechanisms: running organizations, project teams, architecture practices, consulting firms, hundred-million-dollar real estate developments. Leadership is not a problem of character. It never was. And almost everything that circulates about leadership in a corporate feed — including the infographics with gold crowns and the ten golden rules of the effective leader — looks, at best, at only half the system.

If this still sounds far removed from the daily work of a boardroom or a construction site, it's worth pausing for a moment before going on. Because the question that follows isn't about aviation. It's the same question, applied to any organization: are boss and leader the same thing?

II. Boss and Leader Are Not the Same Question

Abraham Zaleznik established it with surgical precision in a 1977 Harvard Business Review essay — "Managers and Leaders: Are They Different?" — that remains, nearly fifty years later, one of the most cited pieces in the field: the manager administers the existing system. Plans, organizes, controls, seeks predictability and risk reduction. The leader questions the system itself, articulates a direction toward where the system hasn't yet arrived, and tolerates — in fact needs — the ambiguity that building something that didn't exist before requires.

John Kotter, two decades later, put it in an equally useful way: management deals with complexity — it maintains the order and predictability of a large system; leadership deals with change — it generates direction where there wasn't any, aligns people behind a vision, motivates people to overcome obstacles the system doesn't yet know how to solve.

The distinction is today so recognizable that even the most superficial corporate content repeats it without citing anyone: "A leader is someone others follow. A manager is someone who controls or administers a business." The consensus already exists. What almost no one does after naming it — and what this article sets out to do — is the hard part: rigorously mapping what concrete forms it takes, who receives it, under what conditions it delivers, and under which ones it fails silently and expensively.

III. The Complete Map of Leadership Styles — With Authors, With Examples, Without Padding

Leadership literature has spent nearly ninety years producing taxonomies, and the vast majority of what circulates today in infographic form is a variation — often uncredited — of a handful of original studies. It's worth naming the authors, not just repeating the names of the styles, because knowing where each one comes from changes how it reads, and because each one, brought into the context of any real organization, produces a recognizable scene.

Positional command. Authority that exists solely because the position exists. It isn't documented, doesn't transfer, doesn't survive the absence of whoever holds it. In an architecture practice, it looks like this: decisions about spec changes get made verbally in the hallway or on site, never recorded in any minute, and their validity depends entirely on who said them, not on the technical logic behind them. In hotel or hospital operations, the same pattern shows up when a service protocol changes depending on who's on shift that week, with no one documenting it as a standard. When that person goes on vacation or changes companies, the project — or the operation — has no memory of why what was decided got decided.

Transactional (James MacGregor Burns, 1978; formalized by Bernard Bass in 1985). Leads through explicit exchange: you hit the milestone, you get the agreed pay or recognition; you fail, you're corrected under closer supervision. It works in repeatable, well-defined tasks — quality control against a regulatory checklist. It collapses under high ambiguity because it doesn't build independent judgment: it builds compliance conditional on active supervision.

Situational (Paul Hersey and Ken Blanchard, 1969). Four positions — directing, coaching, participating, delegating — based on the follower's maturity level relative to a specific task, not their general personality. A newly graduated site resident, faced with a complex structural drawing, needs explicit direction. The same resident, two years later, facing vendor coordination they've already mastered, needs light-touch delegation. The same thing happens with a new store manager facing a complicated month-end close, or a corporate-office project coordinator facing their first negotiation with a difficult landlord. I've seen the most common mistake up close, repeated across more than one organization and more than one industry: treating it as a fixed trait of the person ("So-and-so needs a firm hand, always has") instead of a trait of the task and the specific moment.

The results-driven repertoire (Daniel Goleman, HBR, March 2000). Based on a Hay/McBer study of more than three thousand executives: leaders with sustained top performance master several registers — visionary, coaching, affiliative, democratic, pacesetting, coercive — and switch repertoire according to context and urgency, with emotional intelligence (self-awareness, self-management, social awareness, relationship management) as the decision engine. It gets confused with Situational because both are called "adaptive": Hersey-Blanchard adapts within a closed four-position framework based on follower maturity; Goleman adapts across an open catalog based on the full context — including when a genuinely urgent coercive style is appropriate, and when it's become the only register available, which is, functionally, positional command dressed up as firmness.

Transformational (Burns, 1978; Bass and Bruce Avolio). Idealized influence, inspirational motivation, intellectual stimulation, individualized consideration. In real estate development, it shows up when a project director doesn't just hand out tasks by phase, but explains the regulatory and financial reasoning behind a decision, so the team internalizes the logic, not just the instruction.

Servant (Robert Greenleaf, 1970). The leader exists first to serve the growth needs of those they lead. Greenleaf called for constant self-examination of whether those being served become, over time, more autonomous and more capable of serving others.

Epistemic authority. A synthesis proposed here: leads by showing the work — the finding, the data, the pattern, the full reasoning — not just the conclusion. Builds transferable judgment, not dependence — to the point that if that person leaves, the reasoning they left behind remains valid and usable.

The Rest of the Map, Without Padding the List

Autocratic, democratic, and laissez-faire come from the first-ever experimental study of leadership styles — Kurt Lewin, with Lippitt and White, 1939. Autocratic is the pure form of positional command. Laissez-faire is its dysfunctional opposite, documented by Lewin himself as equally damaging. Democratic precedes Goleman's repertoire.

Charismatic comes from Weber (ideal type of authority) and House (formal theory, 1976): a variant of positional command where the source of power is personality, not the position — the same structural fragility. It's, in fact, the style where the "loudest leadership" mechanism described later in Part IV operates in its purest form: charismatic personality is, in substance, often volume and performed confidence read as real substance. Bureaucratic is positional command institutionalized in written rule. Delegative is Hersey-Blanchard's "delegate" position without the rest of the diagnostic. Coaching, affiliative, visionary, pacesetting, and coercive are the six specific names in Goleman's repertoire. Cross-cultural is a cross-cutting competency, not a style of its own. Results-oriented is, functionally, Goleman's pacesetter — the one that scales worst against low-maturity teams, a fact directly relevant later on.

And here's the uncomfortable question Guam had already raised without naming it: if seven distinct styles can each lead legitimately depending on context, why do the labor market, most job interviews, and much of the corporate boardroom recognize, in practice, only one or two of them as "real leadership"? The answer isn't in style theory. It's somewhere else entirely: in the image each person already carries fixed in their head before meeting the person in front of them.

IV. The Fixed Mental Prototype of the Leader We Carry in Our Heads Before Meeting the Person in Front of Us

Implicit Leadership Theory (Robert Lord, Roseanne Foti, and Christy DeVader, 1984; formalized by Lord and Karen Maher in Leadership and Information Processing, 1991) documents something uncomfortable: people don't recognize leadership by observing behavior neutrally. They recognize it by comparing what they see against a fixed mental prototype of the leader — a mental image, a set of impressions already loaded in advance — that they carry with them before meeting anyone: a firm voice, a certain speech rhythm, a certain physical presence, a certain kind of assertiveness. If the behavior matches, the person is perceived as a leader, whether or not they produce a verifiable result. If it doesn't match, the observer has to do extra cognitive work to recognize it — and often, under time pressure, simply doesn't.

Later research has measured this with concrete physical variables and surprisingly consistent results. Timothy Judge and Daniel Cable, in a 2004 longitudinal study in the Journal of Applied Psychology, found a significant correlation between physical height and both the probability of holding a leadership position and the income level attained, even controlling for gender and weight. Casey Klofstad, Rindy Anderson, and Susan Peters documented in 2012 that voters of both sexes, exposed to the same sentence with electronically manipulated voice pitch, consistently preferred deeper-voiced speakers as more competent leaders, with no change in content. Neither variable predicts real managerial competence. Both predict perception.

The Loudest-Voice Leadership Problem — Why Volume Gets Confused with Capability

There's one more variable, perhaps the most decisive of all within the prototype, and it deserves its own treatment because it explains a pattern anyone who's ever sat in a boardroom recognizes immediately: whoever speaks first, speaks most, and speaks with the greatest apparent confidence tends to be read as the most competent person present, even when the actual content of what they're saying doesn't back it up.

Cameron Anderson and Gavin Kilduff, in a study published in the Journal of Personality and Social Psychology in 2009, documented the exact mechanism: dominant personalities — measured by traits of assertiveness, volume of verbal participation, and tendency to speak first in a group — gain influence and are perceived as more competent not because they actually solve the group's problems better, but because their dominant behavior functions as a competence signal that the rest of the group interprets as evidence, without verifying it. It's a correlation between behavior and perceived ability, not between behavior and real ability.

Cameron Anderson, along with Sebastien Brion, Don Moore, and Jessica Kennedy, went even further in a 2012 study: they found that the relationship between overconfidence and status gained within a group was practically independent of actual competence measured objectively — overconfident people gained status at almost the same rate whether or not they were right, simply by projecting enough security that the group stopped questioning them.

There's a technical name for the more specific variant of this same phenomenon, and it's worth attributing it precisely: the "babble hypothesis" holds that the amount of time someone speaks in a group, not the quality of what they say, predicts who emerges as leader. The study that tested it with the most complete methodological rigor to date is by Nathan MacLaren, Francis Yammarino, Shelley Dionne, Hiroki Sayama, Michael Mumford, Shane Connelly, and Robert Martin, published in The Leadership Quarterly (2020): controlling for intelligence, personality, and gender, they found that speaking time retained a direct and significant effect on leadership emergence, even after ruling out alternative explanations that earlier studies hadn't managed to eliminate.

This connects directly to a finding far more cited outside the leadership literature, but from the same family of mechanism: Justin Kruger and David Dunning, in their 1999 paper that gave rise to the popular term "the Dunning-Kruger effect," documented that people with lower actual competence at a task systematically tend to overestimate their own ability by a wider margin than genuinely competent people — who, by contrast, tend to slightly underestimate themselves because they wrongly assume that a task that's relatively easy for them must be equally easy for everyone else. The intersection of both findings is exactly the mechanism that sustains loudest-voice leadership: whoever knows least often speaks with the most apparent certainty — because they lack the information to doubt themselves — and that apparent certainty is, according to Anderson and colleagues, precisely the signal the group confuses with real competence.

Susan Cain popularized, with rigor and a broad synthesis of prior academic literature, this same tension in Quiet: The Power of Introverts in a World That Can't Stop Talking (2012), naming what she called the "extrovert ideal" — the cultural assumption, particularly strong in Anglo-American corporate contexts, that ideal leadership speaks more, takes up more room in the room, and projects constant confidence, regardless of whether that projection corresponds to greater capacity for judgment.

And there's an empirical counterpoint that gives this pattern its sharpest edge, and that connects back to the cross-tabulation matrix in Part VIII: Adam Grant, Francesca Gino, and David Hofmann, in a study published in the Academy of Management Journal in 2011 with real retail sales teams, found that extraverted leaders produced better results than introverted ones only when leading teams of relatively passive followers, who needed to be activated from outside. Facing teams made up of proactive followers — the close equivalent of Kelley's exemplary follower — the pattern reversed completely: introverted leaders, who listened more than they talked and let the team's ideas flow instead of dominating the conversation, produced superior business results, measured in actual sales, than their extraverted counterparts. The reason the authors themselves propose is direct and fits exactly with this article's argument: a leader who dominates the conversation, without meaning to, drowns out the ideas of a team that already had independent judgment available; a leader who listens builds the space for that judgment to be expressed and turned into results.

Loudest-voice leadership, then, isn't simply a matter of aesthetic preference with no consequence, or an office-culture anecdote. It's a measurable selection mechanism, named and backed by solid empirical evidence, that systematically favors performative confidence over real judgment — and that, according to the Grant, Gino, and Hofmann study itself, produces worse business results in exactly the ecosystems where the most exemplary followers are available to listen. It is, put another way, a mechanism almost perfectly designed to silence epistemic authority at the exact moment it's needed most.

Why the Test Preference Persists, and Why It Doesn't Solve the Underlying Problem

Before getting into this, one precision is worth making, because "psychometric" is a broad term and not every psychometric tool is under discussion here. There exist psychometric instruments for numerical aptitude, logical reasoning, or general cognitive ability, with purposes and validity profiles completely different from what follows. What's being questioned in this section is specifically a subset: personality models and their associated tests — from the MBTI to more rigorous Big Five batteries — which much of the modern process for evaluating managerial potential uses as a filter, and which almost all, one way or another, end up putting disproportionate weight on the introversion-extraversion axis as a predictor of leadership. It's worth asking why, and whether that preference resolves anything or simply institutionalizes, with the appearance of scientific rigor, the same bias we've already documented without any instrument involved.

Timothy Judge, Joyce Bono, Remus Ilies, and Megan Gerhardt, in the reference meta-analysis on personality and leadership (Journal of Applied Psychology, 2002), aggregated 222 correlations from 73 samples using the Big Five model. Extraversion turned out to be, indeed, the most consistent correlate of the five with leadership — both emergence and effectiveness — with a correlation of 0.31. But there are two facts in that same study almost never cited alongside the first, and they completely change the reading: a correlation of 0.31 means extraversion explains, on its own, only around ten percent of the variance in leadership — the remaining ninety percent goes unexplained by that trait. And the full Big Five model together — extraversion, openness, conscientiousness, agreeableness, and neuroticism, combined — reaches a multiple correlation of just 0.48 with leadership, leaving more than three-quarters of what actually determines leadership outside anything a personality test can measure.

This doesn't invalidate the Big Five model as a serious research tool — it's, in fact, considerably more psychometrically robust than popular instruments like the MBTI, whose test-retest reliability and predictive validity for job performance have been consistently questioned for decades (Pittenger, 1993). What it does reveal is that relying on personality traits — extraversion above all — as the primary filter for managerial potential is, at best, betting an organization's most expensive decision on a variable that predicts a modest fraction of the outcome, and at worst, is exactly the same mechanism already described — Anderson, Kilduff, MacLaren — dressed up with the authority of an apparently objective instrument: instead of an evaluator confusing volume with competence at a glance, now a standardized test institutionalizes that same confusion and puts a numerical score on it with the appearance of science. The underlying problem — that epistemic authority and extraversion are independent variables, not the same thing measured under a different name — remains exactly where it was. Only now it comes with a personality-test certificate.

Why It Persists: Not a Modern Design Flaw, But an Evolutionary Inheritance

It's worth asking why, with all the available evidence, prototype bias remains so resistant to change. The answer isn't in any specific corporation or any particular generation of executives — it goes much further back. Mark Van Vugt, Robert Hogan, and Robert Kaiser proposed, in a 2008 American Psychologist article, what they called evolutionary leadership theory: human leadership and followership psychology was shaped, over the vast majority of our history as a species, within small hunter-gatherer tribes — groups of between twenty and a hundred fifty people, coordinating hunting, defense, and resource-sharing face to face. That same psychological machinery, without substantial updating, is what we still use today to evaluate authority inside corporations of thousands of employees, digital hierarchies, and teams distributed across time zones. It is, in the most literal sense, an evolutionary mismatch: the environment changed radically in a few thousand years; the instrument we use to read who deserves to be heard did not. Judge and Cable's height, Klofstad's deep voice, the speaking time of the babble hypothesis documented by MacLaren and colleagues — all of these are, viewed this way, the same signals that once identified a good tribal chief fifty thousand years ago: apparent physical capability, vocal projection audible at a distance, a willingness to speak first and longest in the face of a threat. None of them predicts complex judgment on a board of directors. All of them predicted, reasonably effectively, something different and much older.

And that same ancient machinery doesn't stop at deciding who speaks first in a room. It generates, from the same evolutionary root, the full narratives an organization — or an entire society — builds to explain who deserves to rise and who doesn't, who deserves social mobility and who should stay put. John Jost and Mahzarin Banaji, in their formulation of system justification theory (1994), documented something genuinely uncomfortable: people — including, counterintuitively, many of those least favored by a given hierarchy — tend to construct and sustain narratives that legitimize the existing order as fair and deserved, precisely because believing the system is arbitrary is psychologically more costly than believing it's correct. Those narratives of merit — "they got there on their own merit," "they didn't get promoted because they were missing something" — are neither neutral nor recent. They are, in all likelihood, the modern, moralized version of the same tribal instinct that decided, fifty thousand years ago, who deserved the larger share of the kill. The difference is that today that instinct comes with the vocabulary of merit, performance, and organizational culture — which makes it harder to question out loud, not less primitive at its root.

There's one more piece, less often cited, that completes the picture with an almost hopeful twist. Christopher Boehm, an anthropologist, documented in Hierarchy in the Forest (1999) that ancestral human tribes were not passive victims of whoever accumulated the most dominance. They developed, actively and consistently across generations, what he called reverse dominance hierarchy: deliberate coalitions of lower-power members, organized specifically to monitor, publicly shame, or expel anyone attempting to impose themselves by force or accumulate authority beyond what the group considered legitimate. The mechanism worked, according to Boehm's reconstruction from dozens of contemporary small-scale egalitarian societies, precisely because no individual, however dominant, could impose themselves alone against an organized coalition of everyone else.

It is, with fifty thousand years of anticipation and no engineer involved, the exact same mechanism as Toyota's andon cord: a deliberate structure that gives whoever has the least individual power the ability to collectively stop anyone attempting to exercise power illegitimately. And it suggests something that changes the tone of this entire article: the so-called "primitivization" of modern organizational systems. It isn't that we've gone back to tribal patterns after having outgrown them. It's that we never fully left — we're still running corporations of thousands of people on the same psychological hardware that evolved for tribes of a hundred, and the difference between a healthy ecosystem and a neo-feudal one isn't which of the two is "more primitive," but which of the two managed, like Boehm's tribe or Toyota's line, to deliberately design a reverse dominance hierarchy of its own, fitted to its scale — and which one kept operating, without knowing it, on nothing but tribal instinct, rewarding the loudest voice and punishing whoever questions it.

The Trap of Theatricality: Guaranteed Smoke, Optional Substance

There's one more mechanism, timeless and older than any algorithm, worth naming before getting to the digital accelerant, because it's the one the accelerant later multiplies. Erving Goffman, in The Presentation of Self in Everyday Life (1959), described all social interaction as having dramaturgical structure: a "front stage," where a person performs a managed version of themselves for their audience, and a "backstage," where that management relaxes because no one's evaluating. The most visible corporate leadership — the conference talk, the publication, the executive presentation, the interview itself — happens almost exclusively on the front stage. And that's exactly the trap: the front stage is precisely where volume, theatricality, and performed confidence — what we've already described with Anderson and the babble hypothesis — deliver the maximum possible return, and it's also, almost by structural definition, where the least verifiable evidence of real work exists, because no one audits the backstage of a successful presentation.

Barry Jones and Thane Pittman formalized, in 1982, a still-cited taxonomy of strategic self-presentation tactics — ingratiation, self-promotion, exemplification, intimidation, supplication — each designed, consciously or not, to produce a specific impression on the observer, regardless of whether that impression corresponds to the reality behind it. Leadership "positioning," in the modern sense of the term, is almost always a deliberate combination of self-promotion and exemplification: showing achievements, real or conveniently edited, and projecting stated principles that aren't always verified in practice.

The result is, put plainly, guaranteed smoke: visible activity, projected confidence, constant presence on the front stage — with none of those three variables, on their own, predicting whether there's real substance behind them. It's the same pattern Schmidt and Hunter already demonstrated with data: the most visible signal (interview, presence, theatricality) is the least predictive of real performance; the least visible signal (verifiable work evidence) is the most predictive. Theatricality isn't, in itself, proof of emptiness — some leaders with the most real substance also command the front stage well. But theatricality alone is never proof of substance. Confusing one with the other is precisely the mechanism sustaining much of what this article has described, from the cockpit of Flight 801 to the algorithm deciding, right now, which post you see first.

The Accelerant No One Designed to Recognize One Another

There's a more recent dynamic, and it needs to be said with the same honesty as everything else, precisely because this article is going to be read on the very platform that best exemplifies it. If prototype bias and loudest-voice leadership were already hardwired in by fifty thousand years of evolution, in the last two decades an accelerant appeared that no one designed with that intention, but that produces exactly that effect at a scale no ancestral tribe or boardroom could ever have reached on its own.

William Brady and colleagues, in a study published in PNAS (2017) analyzing more than half a million messages, documented that each additional word of moral-emotional charge in a message increased its probability of spreading on social media by roughly 20% — an effect measured by the message's emotional language, not by its accuracy or rigor, which the study did not directly assess. Platform recommendation systems, optimized to maximize attention time, quickly learn to amplify precisely the performed confidence, assertive tone, and apparent certainty that Anderson and colleagues had already identified as a status signal, not a competence one. The algorithm didn't invent loudest-voice leadership. What it did was take a cognitive bias that used to be limited to the reach of a physical room, and hand it a megaphone with a reach of millions.

And here's a paradox worth naming plainly, because it's the heart of this section: these same platforms market themselves as spaces for connecting with diverse voices, of global reach, of genuine encounter with perspectives different from one's own — the explicit, almost foundational promise is more otherness, not less. But the amplification mechanism that sustains them economically does, in measurable practice, almost exactly the opposite: it optimizes toward the most uniform, highest-emotional-charge signal, not toward the genuine diversity of judgment it claims to promote. They promise otherness. They produce, at planetary scale and in milliseconds, the same tribal instinct we already described with Van Vugt and Boehm — only now that instinct decides, for hundreds of millions of people at once, which voice deserves to be heard first.

And there's a second, more specific piece for this particular professional environment: Alison Hearn, in her academic analysis of personal branding culture (2008), documented how professional profile platforms train their users to present a curated, performed version of themselves — the "brand self" — instead of verifiable evidence of real work. It's, under another name, the same substitution Schmidt and Hunter already flagged in the unstructured interview: signal gets evaluated instead of content. Only now that signal doesn't last half an hour in a room — it lives permanently, gets polished with every post, and actively competes for the same attention that should go toward the real-impact evidence described in Part IX.

There's no need for an accusatory tone to say this — the author of this article himself takes part in the same system by publishing it here, and that's precisely why it's worth naming plainly rather than pretending to stand outside it. If prototype bias is an inheritance we didn't choose, and the algorithm is an accelerant we didn't design with that intention either, the question left for every reader isn't how to completely avoid a system we're all part of. It's more modest and more honest: the next time something captures attention through its volume, its confidence, or its reach, it's worth asking, even for a second, whether that's the same as asking if it's right — and by now we know, after this whole journey, that isn't even the right question either. The right question is whether behind that volume there's a different reading of reality genuinely worth hearing, or just the same old tribal instinct, with better production values.

This separates perceived leadership — matches the prototype, gets credited with authority even when the outcome doesn't back it up — from emergent or unperceived leadership — doesn't match, produces real results, but isn't recognized until someone checks the numbers instead of the image. Small-group literature has a name of its own for this: emergent leadership, informal influence that arises without formal appointment, because its source of authority doesn't depend on matching the dominant prototype.

Jim Meindl, with Sanford Ehrlich and Janet Dukerich, coined in 1985 "the romance of leadership" — the collective bias of crediting a single figure, almost heroically, with results that are actually the product of system, context, and shared luck. We don't just manufacture a fixed image of what a leader should look like; we systematically overestimate how much that image explains when something goes well, and how much blame it deserves when something goes wrong.

Lord and Maher describe the prototype as something relatively static — a template that's already there, preloaded, waiting for comparison. Karl Weick, in his reference work on organizational sensemaking (Sensemaking in Organizations, 1995), adds the dynamic piece that model is missing: people don't apply the prototype passively, they actively and retrospectively construct it, making sense of ambiguous signals after they occur, not before. In an airplane cockpit in the rain, or a boardroom under pressure, no one has time to calmly compare against a fixed prototype — they do sensemaking on the fly, and that process, according to Weick, systematically tends to favor the interpretation that best confirms the action already underway, not the one that best describes reality. The captain of Flight 801 didn't just have a "trustworthy captain" prototype — he was, in real time, making sense of an ambiguous situation in a way that reinforced a course of action he'd already decided on, instead of opening it up to reconsideration.

And here Weick complicates something about this article that's worth stating without softening it: if every interpretation of an ambiguous signal is, by definition, an act of sense-construction and not a neutral discovery of the truth, then no system — however well-designed, with every andon cord and every KSAO in the world — guarantees that the "correct reading" gets recognized in time, because there's no correct reading sitting there waiting to be neutrally discovered. There is, at best, a system that widens the number of interpretations taken seriously before acting, not a system that eliminates interpretation itself. This doesn't invalidate the article's central argument — but it does make it more modest than it sounds in Part IX: designing better isn't the same as guaranteeing objective recognition of the other. It is, at most, making it more likely that more than one construction of meaning reaches the table before it's too late to act on it.

The captain of Flight 801 had nine thousand flight hours and fit perfectly into the "trustworthy captain" prototype. That very match with the prototype may have been, paradoxically, part of what kept him from processing in time a signal that didn't fit the image he had of himself.

The Image Clash: When the Candidate Doesn't Match, and the Cost Isn't Just Invisibility

Up to this point, Implicit Leadership Theory describes something relatively passive: not recognizing someone as a leader because they don't match the prototype. But there's a second, more active layer, that explains why, when the real qualities a job needs don't match the fixed image an interviewer carries in their head, the result isn't just "they weren't recognized" — it's often hostile treatment, active disqualification, or a mid-process rule change no one declares as such.

Alice Eagly and Steven Karau, in their Role Congruity Theory (2002), documented that when someone doesn't match the expected prototype for an authority role, the incongruity isn't read neutrally. It generates actively more negative evaluation — the same behavior that would pass without comment in someone "expected" for the role reads, in someone who breaks the prototype, as incompetence, arrogance, or "something's off," even with identical or superior objective performance.

Julie Exline and Marci Lobel, under the name "the perils of outperformance" (1999), documented something even more uncomfortable: being outperformed by someone of lower formal rank measurably generates discomfort in the evaluator — and that discomfort frequently produces active suppression of whoever provokes it, not recognition of merit. It's not that the system simply fails to see someone who doesn't fit the prototype. It's that when that person also demonstrates a capability that threatens the room's implicit order, the system can react actively against them. ("La Ilusión del Camino Recto" ["The Illusion of the Straight Path," a prior essay in this author's series] named this same pattern from another angle, with an image worth bringing in explicitly here: when a complex or transformative profile enters a mediocre ecosystem, the typical institutional response isn't to make use of that capability, it's to neutralize it — an immune defense from sick systems that perceive health as pathology. It's, in different vocabulary, exactly the mechanism Exline and Lobel documented with data: the system doesn't ignore the person who performs better, it attacks them.)

This needs to be named with precision, because "hostility" can sound abstract, and what actually happens takes concrete form with an identifiable range of intensity. Backlash effect (Laurie Rudman and Peter Glick, 2001) is the technical name for the most severe end of that range: when someone violates what an evaluator considers "prescribed" behavior for their role — not just failing to passively fit, but actively contradicting it with performance they "shouldn't" have, given that prototype — the documented social punishment isn't neutral or proportional. It's punitive. Rudman and Glick originally demonstrated this in gender contexts; the mechanism generalizes to any violation of role expectation: the more clearly someone demonstrates competence they "shouldn't" have, the stronger the reaction against them can be, not in their favor.

Robert Baron and Joel Neuman, in their classic typology of workplace aggression (1996), distinguish verbal from physical, direct from indirect or obstructive forms. Applied to a hiring process, the full range of likely reactions to an image clash includes, from lowest to highest intensity:

  • Micro-invalidation — repeated interruptions, minimizing verifiable achievements ("anyone could have done that"), sustained condescending tone.

  • Passive obstruction — changing process rules without notice, withholding information about the real scope of the role, prolonged silence with no feedback.

  • Moving the goalposts (tied to confirmation bias and motivated reasoning): when a candidate's evidence contradicts the evaluator's prototype, instead of updating the prototype, the evaluator changes the criteria mid-process so that evidence no longer counts as sufficient.

  • Direct verbal aggression — explicit disqualification, deliberately humiliating tone, language that far exceeds anything any legitimate technical evaluation would require. This is no longer biased evaluation — it's aggression, in the literal sense of the term, and it should be named as such instead of softened into "a tough process."

The full pattern is consistent: mismatch with the prototype doesn't get resolved through genuine curiosity toward what wasn't expected. It gets resolved, with documented frequency, by defending the prototype — through cold distance, through silent obstruction, or through open aggression when the mismatch comes paired with undeniable competence.

It's worth saying it plainly, because it repeats often enough to risk becoming normalized: leadership exercised through violence is not leadership. It is, at best, positional command defending itself against the mismatch with the crudest tools it has available. Not one of the seven styles described in Part III — not even the most autocratic, the most pacesetting, the most legitimately coercive in Goleman's version — includes aggression as an acceptable component of its definition. When aggression shows up, we're not looking at an extreme variant of leadership. We're looking at its total absence, disguised with the formal authority of whoever wields it.

And here it's worth pausing, not to accuse anyone in particular, but to ask honestly: which side of this mechanism have we been on, at some point, without realizing it? It's easier to recognize oneself as the victim of an image clash than as its author. But whoever has ever interviewed, evaluated, or made a decision about another person probably also, at some point, defended a prototype without knowing they were doing it.

And this doesn't run in only one direction, and it's worth saying so with the same clarity. If violence exercised from authority toward whoever follows isn't leadership, neither is, on the other side, the lateral or upward violence that followership without integrity exercises toward whoever leads, or toward a peer who simply dared to stand out — the same one we describe in Part VIII under the name lateral violence. Failing to recognize the other's judgment, failing to give them a genuine opportunity to contribute, to err, and to correct, is the same basic denial, whether it runs top-down or the other way. A leader, in the sense this article has been building, doesn't run over others to impose their judgment — if they do, they've already stopped being epistemic authority and gone back to positional command under a different name.

Carol Dweck documented, in her research on growth mindset (2006), that environments sustaining real development — for anyone, in any position — share a minimum condition: they treat error and disagreement as useful information for growth, not as a threat to be crushed. But it's worth saying something the more popular version of that idea tends to omit: growth mindset isn't, and can't be, a purely individual trait that a leader possesses and everyone else simply inherits by proximity. Mary Murphy and Dweck herself, in a later study (2010) on what they called a "culture of genius" versus a growth culture within organizations, documented that the real effect occurs at the systems level, not the individual: when an organization operates, implicitly or explicitly, under the theory that talent is a fixed trait a few possess and the rest don't, that collective belief shapes the cognition, emotion, and behavior of all its members — not just whoever holds it at the top — even when no official document declares it. What actually matters for an organization, then, isn't that its leader has a growth mindset while everyone else doesn't. It's that the entire system is genuinely willing to grow together, because a growth mindset that lives only in one person, surrounded by a system that doesn't share it, is exactly the same power-conditioned mindset named above — just with kinder vocabulary.

An ecosystem that only applies that mindset downward, and abandons it the moment someone with less formal authority questions someone who has it — or that abandons it just the same when someone with genuine, epistemic rather than positional, authority is laterally attacked by those who should recognize them — doesn't have, in any honest sense of the term, a growth mindset. It has one conditioned on power. And that, under another name, is the same hierarchy Guam already taught us costs lives.

And there's an additional pattern connecting back to Part V: those who most often provoke this reaction aren't the passive followers or the conformists — precisely because they don't question the prototype, they almost never trigger it. It's the exemplars, and those who already showed emergent leadership capacity without formal appointment — the ones this article just described as invisible to the prototype until someone checks the numbers instead of the image. They are, with too much documented frequency in the outperformance literature cited above, the focus of envy, suspicion, and sabotage — lateral, from their peers, or downward, from a positional command that reads their competence as a threat rather than a contribution — precisely because of the same quality that should make them the system's most valuable asset. The captain of Flight 801 didn't crash for lack of capable pilots aboard. He crashed with two capable people sitting next to him, whose ability to see the problem in time the entire system — not just him — was designed not to be able to hear.

Everything said so far has looked at the captain. What's missing is a look at the co-pilot — at the other side of the cockpit, the side most leadership literature almost never visits.

V. The Half Almost No One Maps — and That Doesn't Even Have Its Own Word

There isn't yet a natural, established word of its own in Spanish for "followership." A clumsy neologism circulates occasionally in mechanical translations, and almost no one uses it in real professional conversation. The usual practice, even among specialists, is to leave the term directly in English — much the same way "leadership" itself circulated as a pure loanword for years before "liderazgo" fully naturalized into the language. That lexical gap isn't an accident: it's additional, almost literal, evidence of this article's own thesis — we built so much vocabulary to name and admire whoever leads that we didn't leave a single serious word to name, with the same dignity, whoever sustains the outcome from the other side of that same relationship.

Robert Kelley challenged that imbalance directly in a 1988 HBR essay — "In Praise of Followers," a deliberately provocative title for its time — that has aged considerably better than much of what that decade's issue of the magazine published. His starting point was statistical and uncomfortable: in most organizations, between seventy and ninety percent of a project's success or failure depends on the actions of followers, not the leader — simply because they're the numerical majority and execute the real operational part of any decision made above them.

And there's something worth saying here with the same clarity as in Part IV, because it isn't only a problem for whoever's hiring: Lord and Maher's implicit prototype doesn't operate only in a recruiter's head evaluating a résumé. It operates, every day, in the head of any follower trying to decide who to follow. A Kelley exemplar with epistemic authority — who leads by showing the reasoning instead of imposing presence — may simply not be recognized as a leader by coworkers carrying the same internalized image of loud command already described: if they don't sound like what "a leader" is supposed to sound like, plenty of people, in good faith and with no ill intent, won't follow them, however good their read of the situation is. Followership, in that sense, also needs to unlearn its own prototype — not just wait for the system above it to learn to recognize better.

Kelley crossed two variables — independent critical thinking and active commitment — to identify five patterns:

Exemplary. High independent judgment, high commitment. Questions when it's needed, engages when it's needed. Doesn't confuse loyalty with automatic obedience or autonomy with detachment. The scarcest pattern and the most valuable — and, paradoxically, the one that most often generates friction with positional command, precisely because it questions.

Alienated. High independent judgment, low commitment. Sees the problem clearly, often more clearly than anyone else in the room, and decides, through a rational risk calculation, not to risk anything by flagging it. It's not gratuitous cynicism. It's, almost always, the learned consequence of having flagged something before and paid a cost for it.

Conformist. High commitment, without exercised independent judgment. Executes with genuine energy, whether the instruction is correct or not. The pattern positional command tends to reward fastest, because it produces the — sometimes deceptive — feeling that the system is working.

Passive. Neither judgment nor active commitment. Waits for explicit instruction and rarely takes initiative even when the situation clearly calls for it.

Pragmatic. Moves according to whatever seems safest at the moment, without a fixed compass of values.

Kelley's most uncomfortable finding — the backbone of this article — is that excellent leadership, on its own, is necessary but radically insufficient if the surrounding ecosystem is made up mostly of passive or alienated followers. Leadership is not, and never was, a unilateral act. It's a relationship — and like any relationship, it needs reciprocity from both sides to produce something neither could produce alone.

Leadership Isn't Possessed, It's Activated

Shared leadership (Craig Pearce and Jay Conger, 2003) and distributed leadership (Peter Gronn, 2002) document that in high-performing teams the leadership role rotates according to who holds, at that moment, the most relevant knowledge for the decision at hand — it isn't a fixed post, it's a function that switches on and off. Kelley's exemplary follower isn't "a follower forever": they become a momentary leader when their judgment is most relevant, and return to followership without friction when it isn't. It's the strongest argument against fixed positional command: in a well-designed system, leadership moves constantly among people.

VI. Leadership Isn't Only Top-Down

Everything described so far, without saying it explicitly, has assumed a single direction: someone with formal authority directing toward those who don't have it. That's just one of three documented directions, and the other two are often what determines whether an ecosystem functions or stalls.

Top-down is the conventional direction — the one Part III's seven styles describe in their most usual form.

Bottom-up, or managing up (John Gabarro and John Kotter, HBR, 1980), is a subordinate's ability to effectively influence a superior's decisions without the formal authority to demand it — through evidence, timing, building incremental trust. An exemplary follower who never develops this capacity is left, no matter how much independent judgment they have, with no way to translate that judgment into real influence over decisions that aren't formally theirs to make. On Flight 801, the first officer and the flight engineer had, in theory, a managing up channel available — what was missing was a protocol that made that channel usable under real pressure.

Horizontal or lateral (lateral leadership, widely documented in matrix project-management literature, where no one has hierarchical authority over anyone) is coordinating, influencing, and aligning peers who neither report to you nor to whom you report. In a complex real estate development, it's exactly what happens between the lead design architect and the construction manager: neither commands the other, and the project only moves forward if both exercise genuine lateral influence. This same lateral dimension, when it fails instead of coordinating, doesn't stay at simple disagreement — it becomes lateral violence, the mechanism developed further in Part VIII: the same absence of hierarchy that here allows genuine cooperation between peers is, there, what allows one peer to attack another with no formal authority stepping in.

This completes what Pearce, Conger, and Gronn already suggested: leadership doesn't just rotate among styles based on context and among people based on who knows most — it rotates in direction too. An ecosystem that only recognizes top-down leadership is, by design, blind to the other two-thirds of where real leadership happens every day.

VII. A Necessary Counterexample: The Cord Anyone Can Pull

Before moving to the layers of the ecosystem, a second case is worth telling, deliberately different in tone from Guam's, because it shows what a system designed from the outset to do exactly the opposite of what failed in that cockpit looks like in real industrial practice, with a name, date, and place that can all be verified — and what happens when that system meets people trained for years under the opposite regime.

On the production line of the Toyota Production System, developed primarily by Taiichi Ohno starting in the 1950s — itself inspired by an automatic loom Sakichi Toyoda had invented decades earlier, capable of stopping on its own the instant a thread broke, so it wouldn't keep weaving defective cloth — there is the andon: literally, a light or signal lamp. Any line worker, regardless of seniority or formal rank, has explicit, formal authority, backed by management, to pull a physical cord the exact moment they spot a defect or a reasonable doubt. That act, carried out by the person with the least formal authority in the entire chain, immediately stops the whole line.

The case that best shows what happens when that design meets people trained under the opposite system happened in Fremont, California, in 1984. The General Motors plant there had closed in 1982, after years of being, according to the auto industry's own records, one of the worst in the entire corporation for both quality and labor relations — high absenteeism, occasional sabotage on the line, near-total distrust between workers and management. Toyota and GM reopened it as a joint venture, NUMMI (New United Motor Manufacturing, Inc.), and, deliberately, rehired a large share of the same workers who had been laid off two years earlier. Same people. A completely different system on top of them.

Susan Helper and Rebecca Henderson, in a landmark article on the decline of General Motors published in the Journal of Economic Perspectives (2014), document something that precisely confirms this article's central psychological mechanism: under the old GM regime at that same plant, the unwritten, shared norm among workers was clear — never pull the emergency cord, out of fear of punishment. Stopping the line, even for a legitimate reason, read as a sign of incompetence or insubordination — exactly the same social-cost calculation the first officer made in the cockpit of Flight 801. That unwritten norm also had a metrics-driven logic behind it: when management rewards superficial speed and uninterrupted line continuity above everything else, the documented result in quality-control literature since the 1960s is what Armand Feigenbaum — the engineer who formalized total quality control — dubbed the "hidden plant" (1961): the enormous proportion of time, material, and effort an organization spends silently correcting, outside any official metric, defects the system itself pushed workers to hide instead of expose in time. A line that never stops isn't necessarily the most efficient one — it often just has better-looking reports. When NUMMI opened with Toyota's andon system — the same formal authority to stop the line, now explicitly encouraged by management instead of punished — many of those same workers, carrying the reflex learned over years under the old regime, simply didn't use it. The cord was there. The authority to pull it, too. Learned fear didn't disappear just because the manual said it no longer applied.

What did change behavior, documented in later work by Paul Adler, Barbara Goldoftas, and David Levine on NUMMI's organizational evolution, wasn't a memo or a single training session. It was sustained repetition, week after week, that pulling the cord — even when the reason turned out to be minor, even when it stopped the line without strict necessity — never generated a negative consequence for whoever did it, not once, verifiably. Over time, cord-pull rates rose steadily, and with them, finished-product quality rose too, to levels that same workforce had never reached under GM's previous regime.

It is proof, with verifiable name, date, and location, of this article's complete thesis: it wasn't the character of those workers that changed between 1982 and NUMMI. They were, largely, the same people. What changed was whether the system above them punished or rewarded flagging a problem in time — and how much sustained repetition it took for accumulated evidence to displace a fear learned over years under a different regime.

What's counterintuitive, extensively documented by operations researchers since, is that this design — which appears to sacrifice speed in exchange for giving veto power to whoever holds the least formal authority in the whole system — produced, consistently across generations and in Toyota plants around the world, some of the highest quality and efficiency levels ever achieved in large-scale industrial manufacturing. The reason, seen through this article's framework, is direct: Toyota deliberately designed a system with the level of psychological safety and epistemic authority distribution exactly inverse to that of Flight 801's cockpit — and NUMMI proved that design can work even with workers trained for years under the opposite regime, as long as the evidence that punishment no longer follows is sustained long enough to be believed.

Elinor Ostrom, the only woman to date to receive the Nobel Prize in Economics, documented over decades — work synthesized in Governing the Commons (1990) — how real communities around the world manage to sustainably govern shared resources without external centralized authority or privatization, through their own rules, peer monitoring, and graduated sanctions designed by the resource's own users. It's the same principle Boehm documented in ancestral tribes and that Toyota institutionalized on a production line, now with the broadest empirical backing that exists in social science on decentralized governance: the systems that work best over the long run are almost never the ones that concentrate monitoring power in a single authority — they're the ones that distribute that capacity among those actually close to the problem.

And here too it's worth resisting the temptation to oversimplify: Ostrom's real finding isn't "decentralize authority and everything works." Her successful cases share a demanding set of simultaneous conditions — clear boundaries on who belongs to the system, rules congruent with specific local conditions, accessible conflict-resolution mechanisms, graduated sanctions instead of immediate expulsion, and external recognition of the group's right to govern itself — and when more than one of those conditions is missing, decentralized governance collapses just as easily as any other institutional arrangement. NUMMI worked because, beyond the physical cord, it had training, protocol, and above all, sustained time with consistent evidence behind it. An organizational ecosystem that only "distributes authority" without that complete scaffolding isn't applying Ostrom — it's using her name to justify ambiguity without design, which is, in fact, the same underlying error we already named in Part VIII under the name structure and process.

It's no coincidence this example shows up precisely here, in the transition toward the ecosystem layers that follow. The andon cord isn't a leadership style. It's a piece of structure and process — the first of the ecosystem layers — designed with enough precision to let Kelley's exemplary follower activate as a momentary leader, in any of the three directions described above, exactly the instant their judgment is most valuable, without needing to negotiate authority first — as long as, like the workers in Fremont, they've had enough time to believe the system genuinely changed this time.

VIII. The Ecosystem Has Layers, and Not All of Them Are People

Before naming them, an honest precision is due: a complete ecosystem, in the broadest sense of systems theory, has more variables than this article is going to develop. Urie Bronfenbrenner, in his ecological theory of human development (The Ecology of Human Development, 1979), proposed a far more exhaustive model of nested layers: the microsystem — the immediate environment with its physical objects and direct face-to-face interactions —, the mesosystem — relationships between different microsystems —, the exosystem, the macrosystem — the surrounding culture —, and the chronosystem — time itself as a variable. A real organizational ecosystem includes, in that full sense, not just people: it includes the physical space where the work happens, the objects and tools that mediate interaction, the interactions themselves among the components, and the entities that make up the complete system, human and non-human.

This article doesn't attempt to model that complete ecosystem. It deliberately focuses on the five layers most directly bearing on the specific problem that matters here — whether the system recognizes in time the judgment of whoever holds less formal authority — knowing that real variables are left out: the physical design of the workspace, the mediation of tools and objects, and the concrete interactions among all the system's components, not just among people. With that honesty about scope stated up front, the five layers relevant to this article's argument are:

Structure and process. Explicit governance, clear breakdown of work — the WBS/PBS/EPS chain — visible rules versus deliberate ambiguity. Without this layer, even the most brilliant transformational leader ends up, without meaning to, exercising positional command, because it's the only tool available.

Culture (Edgar Schein, Organizational Culture and Leadership). Three levels: visible artifacts, espoused values, and underlying basic assumptions — what actually gets rewarded and punished, almost never written down. The gap between what's declared and what's assumed produces alienated followers: they saw the distance between the "open door" rhetoric and the reality of who can question without consequence.

Power and politics. Informal networks, unwritten coalitions, and who has real access to information. When information is managed as a personal control asset instead of a shared resource, the system produces more conformists — who learn that asking has a cost — than exemplars.

Psychological safety (Amy Edmondson, The Fearless Organization, 2019). It isn't "a pleasant atmosphere." It's the measurable certainty that flagging an error or a disagreement doesn't cost reputation or position. Edmondson documented that top-performing hospital teams reported more errors, not fewer — their psychological safety let them name errors instead of hiding them.

The nature of the challenge (Ronald Heifetz, Leadership Without Easy Answers, 1994). Technical challenges — solvable with existing expertise — versus adaptive challenges, which require the system to change its values or practices. An organization built on personalist authority is always facing an adaptive challenge. Positional command only knows how to solve technical challenges: that's why, there, even the best-intentioned leadership crashes without understanding why.

It's worth naming here, with precision, which academic conversation this article's argument actually belongs to, because it isn't the same one that supports most of Part III's taxonomy. Mary Uhl-Bien, Russ Marion, and Bill McKelvey proposed, in The Leadership Quarterly (2007), Complexity Leadership Theory: an explicit paradigm shift away from leadership models inherited from the industrial era — hierarchical, top-down, designed for a predictable physical-production economy — toward a framework that understands leadership as a complex interactive dynamic from which adaptive outcomes emerge — learning, innovation, responsiveness — with no single actor, on their own, producing or controlling them. It is, with formal academic naming, exactly the shift this article has been making since Guam: from studying the isolated leader, to studying the relationship between whoever leads, whoever follows, and the complete system that produces that interaction. The cockpit of Flight 801 wasn't an individual leadership failure — it was, in Uhl-Bien's vocabulary, a failure of the complex dynamic that was supposed to produce adaptation, and didn't produce it in time.

And there's a genuine tension here this article shouldn't hide, because it's precisely the source from which Uhl-Bien and Heifetz launch their critique of much of the language this very text has used too comfortably: "designing systems," "building structure," "installing an andon cord." That's engineering language, and complexity theory is, in its original formulation, a direct critique of the idea that emergent adaptive dynamics can be fully controlled or designed from above, the same way a manufacturing process is designed. Heifetz insists on something similar from another angle: adaptive challenges aren't solved with a technical fix imposed by whoever holds authority — they require the whole system to do the work of rebuilding its own values and practices, with real loss and real dissent along the way, something no org chart or protocol can substitute for. Put plainly: part of what this article proposes — a matrix, a corrected KSAO, a cord anyone can pull — is real and helps, but it doesn't "solve" the problem the way an engineer solves a circuit. At best, it creates more favorable conditions for the slower, more uncomfortable adaptive work to happen. Confusing that favorable condition with the solution itself would, ironically, be the same kind of overconfidence Part IV spent its length documenting.

When the Lateral Direction Fails: Lateral Violence

There's a layer of friction the matrix below doesn't fully capture, because it occurs between people at the same level, not between leader and follower — it's the same lateral dimension described in Part VI, now seen through its failure mode instead of its cooperative form. Nursing literature — where the phenomenon was first documented with the greatest rigor, precisely because clinical teams depend on honest error reporting — has a name for it: lateral violence or horizontal hostility (Martha Griffin, 2004, among others): active hostility between peers of the same rank, frequently directed at whoever threatens to stand out or question the group's social norm.

It's the mechanism by which a Kelley exemplar can not only clash with positional command above them — they can be laterally boycotted by conformists at their own level, who read their independent judgment not as a threat to the formal hierarchy but as a threat to the group's comfortable social stability. It's followership attacking followership, with no formal leader needing to intervene for the ecosystem to end up expelling, through horizontal attrition, exactly the profile it most needed to retain.

It's worth quickly mapping what this looks like in practice, because "hostility" between peers tends to be quieter than the vertical aggression already described, and that's precisely why it goes more unnoticed. Griffin documents a consistent catalog of manifestations, ranging from nearly invisible to open: double-meaning comments made in front of the group, deliberate withholding of information a peer needs to do their job well, sabotage of tasks or deliverables, systematic social exclusion from conversations or informal decisions, gossip or talking behind someone's back, and deliberate breaches of confidentiality. None of these forms requires hierarchical authority — all of them happen between equals, which is precisely what makes them harder to name and to correct: there's no positional command to appeal to, because the attack doesn't come from above.

The Other Two Directions Fail Too — and Almost No One Maps Them

Part VI established three leadership directions: top-down, bottom-up, and lateral. If the lateral one has a named failure mode, the other two have one too, and leaving them out of this map would unintentionally suggest that only the horizontal level gets corrupted. It doesn't.

Top-down violence was already documented in detail in Part IV — the full Baron and Neuman spectrum, from Rudman and Glick's backlash effect to direct verbal aggression — so there's no need to repeat it here. It's enough to name it in its proper place within this complete map: it's the failure mode of the downward direction, and it's the one management literature has studied most, precisely because it's the most visible and the one most often correctly identified as pure positional command.

Bottom-up violence is the least documented of the three, and probably the least named in corporate conversation, though it exists and carries real consequences. Sandra Branch, Sheryl Ramsay, and Michelle Barker, in their reference review on workplace bullying (2013), document what they call upward bullying: sustained hostility, sometimes coordinated among several subordinates, directed at a superior — systematically refusing to share critical information, quietly sabotaging the execution of decisions already made, socially isolating whoever's leading within their own team, or building a shared narrative that erodes their credibility with the rest of the organization without necessarily any real basis for doing so. Unlike top-down violence, which is easier to name because it matches the popular image of a "toxic boss," bottom-up violence rarely gets read for what it is — it frequently gets read, backwards, as evidence that the leader "didn't know how to win over the team," when in reality the team, or a coordinated part of it, actively decided not to be led, even by someone exercising genuine epistemic authority and real transformational leadership.

This matters because it closes a real gap in this essay: not every leadership failure is the leader's failure. Aggressive positional command can produce real harm (top-down). Followership without integrity can sabotage a legitimate leader from below (bottom-up). A group of peers can expel, through horizontal attrition, exactly the profile it most needed to retain (lateral). All three are documented system failures, not necessarily failures of character in whoever stands at the center of each one — and confusing the three with each other, or assuming only the first exists, is exactly the same error of analytical level this article has been naming since Guam.

The Matrix Almost No Leadership Article Dares to Cross

The vast majority of leadership material treats the two taxonomies — the one for whoever leads and the one for whoever follows — as if they lived in separate universes. But the real outcome of any leadership relationship never depends on just one side. It depends on the intersection. Five leadership styles by five followership patterns give twenty-five possible combinations; developing all twenty-five would turn this section into a catalog, not an argument, so here are the nine that best show the full range — from the most fragile to the one genuinely generative:

Positional command + conformist follower. Works, on the surface, while command is present. Collapses the moment it's absent, because no one developed independent judgment: everyone executed, no one decided.

Positional command + exemplary follower. Constant, near-structural friction. The exemplar questions; command reads the question as a threat. Sustained over time, it produces alienation through attrition, or the departure of the profile most worth retaining.

Transactional + pragmatic. Stable, uninspiring, functional in predictable tasks.

Situational + any pattern, well diagnosed. The most adaptable of the twenty-five — as long as the maturity diagnosis is accurate and not a retroactive excuse built after the fact.

Results repertoire + alienated. Reversing alienation requires sustained evidence that the real cost of speaking up has changed — not just a change of tone when asking for input.

Transformational + alienated. Counterintuitive: inspiration doesn't automatically reverse a learned history of the cost of speaking up.

Servant + conformist. Can look virtuous on both sides, but without the intellectual-stimulation dimension, it can end up reinforcing the conformist's lack of independent judgment instead of developing it.

Epistemic authority + passive. The leader wears themselves out showing evidence and reasoning to someone who's never going to use it to decide anything on their own.

Epistemic authority + exemplary. The one genuinely generative crossing among the twenty-five possible ones. Here Porter and Kramer's shared value stops being an abstract concept: it isn't a fixed pie being divided, it's a pie that grows because both sides invest simultaneously.

This matrix suggests something the fast-consumption material almost never dares to say clearly: "What kind of leader am I?" is, most of the time, the wrong question. The one that actually predicts the outcome is "What crossing am I generating right now, with whom, within what structure, what culture, what power, what psychological safety, and in what direction?"

It's worth saying clearly, before moving on, what this means for all the variables this article has been accumulating up to here — seven leadership styles, five followership patterns, three directions, five ecosystem layers, and one full evolutionary layer beneath all the previous ones. It doesn't mean eliminating the friction between them. Positional command and an exemplary follower are almost always going to grate against each other, by the design of both roles, not by either one's defect. An evaluator with an implicit prototype and a candidate who doesn't match it are almost inevitably going to generate tension the first time they meet. That friction isn't the problem to be solved. It is, in fact, the signal that two legitimate, distinct ways of reading the same reality are in the same room. The congruence this article defends isn't forced harmony or absence of disagreement. It's a system's capacity to sustain that friction long enough for it to resolve into an understanding of otherness, instead of resolving, by the shortest and oldest route our biology knows, into dominance. That — not the absence of friction, but how a system sustains it — is what separates Guam from Toyota, and what ultimately separates positional command from epistemic authority.

Everything above is, up to here, diagnosis. What's missing is the question that actually moves something: what does whoever holds the power to hire, promote, or protect someone do with this information.

IX. Why This Matters — The Question HR, Headhunters, and Boards Almost Never Ask Out Loud

Given that a real, documented diversity of effective leadership styles exists, and given that Implicit Leadership Theory shows that most evaluators do so against a fixed mental prototype of the leader — a mental image, a set of already-formed impressions — without knowing they're doing it, the question rarely asked — and that should be the first one in any senior hiring process — is:

Is the leadership actually being hired, or is it the image of leadership already fixed in our heads before the first interview even opens?

The Formula Almost Always Left Out

The KSAO model — Knowledge, Skills, Abilities, Other characteristics —, standard in industrial-organizational psychology for decades, already contains the answer to what should be measured. Knowledge and skills are the two letters almost every process evaluates with reasonable rigor. Abilities — underlying capacity, not just trained skill — already starts to dilute. And the "O" — traits, values, motivation, congruence — is systematically the one left out, precisely because it's the hardest to verify quickly.

Kristof-Brown, Zimmerman, and Johnson, in a 2005 meta-analysis, formalized a distinction almost no process separates: Person-Job fit (can they do the work?) versus Person-Organization fit (are their values congruent with the culture?). Most processes evaluate the first and take the second for granted, unverified.

One more honesty is worth adding before moving on, because it concerns the real limit of everything this article has proposed so far. KSAO, Person-Organization fit, the twenty-five-cell matrix from Part VIII — these are frameworks, not closed equations. They give evaluation a common structure, comparable across people and across different organizations: that's what makes them useful, and it's real. But no framework, however rigorous, resolves on its own the particular case in front of you. It's, allowing for the analogy, like solving an integral: the general formula predicts the shape of the curve — and then there's the constant, the +C, that no formula can fix in advance because it depends entirely on the specific initial condition it's given. The values any healthy ecosystem needs — integrity, congruence, psychological safety — are, in that sense, the universal part of the equation. But how those values show up in this particular person, at this particular moment in their career, within this particular culture, is the constant no framework calculates ahead of time. The formula guides judgment. It doesn't replace it.

When Not Even the Job Itself Is Clearly Named

There's a layer to this problem that happens before evaluation even begins, and it deserves its own name because it isn't evaluator bias — it's a design failure in the job posting itself. Robert Kahn, Donald Wolfe, Robert Quinn, Diedrick Snoek, and Robert Rosenthal, in their foundational work on organizational stress (1964, still one of the most cited in all of organizational psychology), formalized the concept of role ambiguity: the lack of clarity about a job's expectations, scope, and success criteria, documented as one of the most consistent and best-studied sources of work-related stress that exist.

In architecture and engineering, this ambiguity takes a very concrete, very recognizable form: posting a vacancy for "Principal" — a role of technical and business leadership, often on a partnership track, oriented toward client development and strategic project direction — when what the organization actually needs, and what it effectively interviews and evaluates for, is a "Job Captain" — a role focused on coordinating and technically reviewing construction documents, centered on drawing consistency and quality control, without the strategic or business dimension the title "Principal" implies anywhere in the world. In industry practice, these are two roles with almost opposite responsibilities, authority expectations, and career trajectories, and confusing them in the posting isn't a semantic nuance — it's exactly the role ambiguity Kahn and colleagues documented as a measurable source of harm, now transferred from the already-hired employee to the candidate who doesn't even know yet what real position they're being evaluated for.

And that ambiguity rarely runs along a single axis. It often compounds across several at once: not just seniority (Principal vs. Job Captain), but also function (Project Manager, oriented toward scope, budget, and client relationship, confused with a shop or office head, oriented toward internal operations and production). When both axes go misaligned at the same time, the candidate isn't calibrating their preparation against a single poorly defined role — they're calibrating against two distinct, overlapping definitions of success, without anyone in the process having noticed the overlap. It's the same role ambiguity from Kahn and colleagues, just compounded: not one axis of uncertainty, but two crossing inside the same job posting.

And when a third layer gets added on top of that — the language the process is conducted in doesn't match the language the posting was written in, or the required proficiency level was never precisely specified — the candidate ends up absorbing three simultaneous sources of role ambiguity in a single process: they don't know for certain what the job is, they don't know for certain what function is expected of them, and they don't know for certain what language they'll be evaluated in or by what standard. None of these three ambiguities is the candidate's fault. All three are, exactly as Kahn's literature predicted back in 1964, the responsibility of a system that never precisely defined what it was asking for before it started asking.

When the Guessing Game Wears the Costume of Rigor: Cultural Fit

Lauren Rivera (Pedigree: How Elite Students Get Elite Jobs, 2015) documented that "cultural fit" at elite firms functions as a proxy for social class, school, and shared cultural capital — not for genuinely shared values. The correction: from cultural fit to cultural add — not "do they resemble us?" but "what do they add to the ecosystem, and what would the ecosystem need to change for their real capability to deliver?"

An image is worth using here that sums up the cost of not asking that question: a profile of the highest capability — epistemic authority, say — placed in the wrong ecosystem isn't lost to incompetence. It sits unused, like a Ferrari covered in dust in a parking lot, engine intact, with no road designed for it to run on. The question that falls to whoever's hiring isn't just "is this person competent?" — it's "am I willing to invest in building the ecosystem where that capability actually delivers?" That turns cultural add from an evaluation question into an investment question.

And here a distinction is worth making that avoids falling into the same trap this article has been flagging all along: there is no legitimate fit of origin — resembling the group already, before joining, is exactly the class bias Rivera documented, and nothing that follows contradicts that. But there can exist, and it's worth actively pursuing, a fit of destination: genuine social integration into a group, deliberately built after the person is already inside, not demanded as a condition of entry. Gregory Walton and Geoffrey Cohen, in a randomized experiment published in Science (2011), demonstrated this with a precision rare in psychological interventions: a brief intervention, designed to reduce first-year college students' uncertainty about whether they belonged in their new environment — framing social adversity as common and temporary, not as a sign they didn't fit — raised African American students' GPA in a sustained way over three years, cut the achievement gap with white students in half, and improved their reported health and well-being. None of the participants, by the end of the study, was aware they'd received the intervention.

The difference from fit of origin couldn't be clearer: no one was filtered for already resembling the rest of the campus — they were, in fact, the ones who least matched the institution's dominant social profile. What the intervention deliberately and measurably built was belonging itself, as a destination earned through real investment, not as an entry requirement verified in an interview. That's the fit worth pursuing inside any organizational ecosystem: not the kind that decides who gets through the door, but the kind an organization actively builds, with real resources and real time, once the person is already inside — the exact same investment logic the dust-covered Ferrari called for a paragraph back.

And the image has a flip side worth naming with the same clarity: that same Ferrari, once given the right road, doesn't just stop sitting idle — it improves the road itself. An exemplar with epistemic authority, placed in the generative crossing described in Part VIII's matrix, doesn't just deliver their own capability: they raise the standard of rigor and judgment for the entire team around them, because their way of leading — showing the reasoning, not just the conclusion — builds transferable capacity in the people around them, not dependence on them. Investing in the right ecosystem for this profile isn't just avoiding wasted talent. It's one of the few investments documented in this article that grows the whole pie, not just divides it better.

The Uncomfortable Fact About How We Actually Predict Performance

Frank Schmidt and John Hunter, in Psychological Bulletin, 1998, synthesized the most complete meta-analysis ever conducted on the predictive validity of selection methods. The unstructured interview — the one that dominates the vast majority of real hiring processes — has a predictive validity of barely ~0.38 with subsequent performance. The method with the highest validity of all those evaluated is an actual work sample, ~0.54: evaluating what the person actually produced, not how they carried themselves socially for half an hour under artificial pressure.

And That Error Hits Unevenly

Prioritizing work evidence over interview performance turns out, almost as a side effect, to be the most inclusive way to evaluate talent — it doesn't penalize unconventional communication styles or different ways of processing a question under pressure. Holly White and Priti Shah found, across two studies (2006, 2011), a specific, non-uniform pattern: adults with ADHD-associated traits outperformed adults without ADHD on divergent thinking — generating more original ideas — and, in the 2011 study, on real, world-verifiable creative achievement, though the 2006 study also found that same group performed worse on convergent thinking tasks — arriving at a single correct answer. It's not that neurodivergence guarantees more creativity across the board; it's that it excels specifically at divergent idea generation, precisely the kind of thinking a structured interview built to evaluate "the correct answer" tends not to know how to read.

This isn't limited to ADHD, or to architecture. Simon Baron-Cohen, along with Sally Wheelwright, Carol Stott, Patrick Bolton, and Ian Goodyer, documented in 1997 that parents and grandparents of children with autism showed up, in the sample studied, more than twice as often in the engineering profession as those of other children — evidence of a real cognitive link between autism-spectrum traits and the capacity to systematize complex patterns, the same capacity underlying engineering, data science, and related technical disciplines. In architecture, industrial design, engineering, data science, academic research, and other high-systemic-demand disciplines, multiple studies report a prevalence of ADHD, autism, and giftedness traits above the general population — often, the same cognitive profile that sustains complex systems thinking and genuine tolerance for multivariable ambiguity, the natural terrain of epistemic authority.

And it's worth naming something almost never mentioned together: these three conditions — ADHD, autism, and giftedness — aren't mutually exclusive categories. James Webb and colleagues, in their reference work on dual diagnosis in gifted populations (2005), documented that these three conditions overlap and get confused with each other with enormous frequency — a gifted child or adult can show the same pattern of motor restlessness, intense hyperfocus, or social difficulty that an evaluator without specific training automatically labels ADHD or autism, when in reality they're manifestations of an intellectual capacity the environment failed to accommodate — to the point that misdiagnosis, in either direction, is, according to their research, more the rule than the exception. A selection process that evaluates through social interview instead of work evidence doesn't just fail to distinguish ADHD from epistemic authority: it fails to distinguish any of these three conditions, frequently overlapping in the same person, from the social noise a traditional interview format is designed to reward.

The Adjustment Period Almost No One Considers Anymore

There's one more piece rarely part of this conversation, and it directly worsens the problem of misreading short tenure: the practical disappearance, in many current processes, of a genuine probationary period — understood not as a legal formality, but as an honest window of mutual adjustment between person and ecosystem. John Van Maanen and Edgar Schein, in their foundational work on organizational socialization (1979), documented that integrating into a new culture — understanding its real underlying assumptions, not just its stated values — follows a predictable curve that takes time, almost never less than several months in a genuinely responsible role. Michael Watkins formalized that same curve in more operational terms in The First 90 Days (2003): a new executive needs, almost universally, a period of learning the system before their performance reflects their real capability, not their initial unfamiliarity with the terrain.

When that period is shortened or disappears — when performance on unfamiliar terrain is expected from week one, or when a tenure of a few months gets read as failure without first asking whether there was even a real opportunity to complete the adjustment curve — the system is, once again, measuring wrong: it's evaluating the orientation phase as if it were the stable-performance phase, and penalizing the person for a curve the organization itself never gave them time to travel. It's the same root error Schmidt and Hunter document on the interview side — judging with the wrong instrument — now applied to time instead of method.

Short Tenure, Read the Opposite of How It's Usually Read

Michael Arthur and Denise Rousseau documented the concept of the boundaryless career: frequent movement between organizations isn't, in itself, evidence of instability. It's often the signal of someone who reads, faster than average, when an ecosystem is incompatible with what they need to deliver — and acts accordingly, instead of spending years absorbing an already-identified dysfunction. ("La Ilusión del Camino Recto" develops this same tension in more detail: why the linear career path stopped being the standard against which to measure competence, and what gets lost when we keep evaluating talent as though it still were.)

But "it isn't instability" is, on its own, an insufficient defense — it leaves the conversation on the ground of reasonable doubt, not evidence. What actually allows short tenure to be evaluated rigorously is separating three questions almost always merged into one:

  • Evidence of impact — what did they produce, verifiable beyond their own account? A project delivered, a documented efficiency gain, a system that keeps working after they left.

  • Judgment — did the departure reflect a reasoned reading of the ecosystem — a mismatch of values, an absence of governance, a structural ceiling — or a reactive pattern with no diagnosis?

  • Integrity and congruence — not the tenure itself, but whether the same values, the same judgment, and the same way of working hold up recognizably across different ecosystems. This is verifiable with considerably more precision than it seems: professional relationships the person keeps active beyond the organization that originated them, projects former colleagues or former bosses call them back to by their own choice, references that describe the same behavioral pattern across completely different companies. Congruence isn't staying — it's whoever crossed paths with that person in more than one place independently describing the same person.

And when those traditional references don't exist — the increasingly common case of someone who built an independent practice, their own platform, or documented achievements outside the structure of an employer who could vouch for them — the evidence doesn't disappear, it just changes form: public deliverables, their own publications, clients who return by choice, a verifiable body of work with no need for a former boss to confirm it.

It's also worth naming, plainly, something almost never said out loud within a hiring process: the generic suspicion that someone with a non-linear career "must have something wrong with them" — the same prejudice of assuming they're a bad person before even meeting them — has, in the vast majority of cases, no statistical foundation whatsoever. Genuinely toxic profiles in the clinical sense are, according to available epidemiology, a real minority within any population: Frank Stinson and colleagues, in the reference NESARC study (2008), documented a lifetime prevalence of narcissistic personality disorder of 6.2% in the general population; William Compton and colleagues, in a comparable representative sample (2005), documented a prevalence of antisocial personality disorder of 3.6%. Even adding both disorders without overlap, more than ninety percent of any group of candidates fits neither one. Treating the ambiguity of a career path as presumptive evidence of pathology is, statistically, almost always an error — and is, almost always, the same mechanism as the fixed mental prototype of the leader described in Part IV, in darker vocabulary that's harder to question out loud.

And there's a documented irony that makes this error worse still: evaluation schemes that rely on impression and charisma instead of work evidence aren't just blind to the genuinely toxic minority of profiles — they frequently prefer them actively, precisely because those profiles project the confidence, quick decisiveness, and apparent capacity to "bring order" that the same scheme confuses with leadership. Paul Babiak and Robert Hare, in their reference research on corporate psychopathy — synthesized in Snakes in Suits: When Psychopaths Go to Work (2006) — documented that traits like superficial charm, unshakeable confidence, and a willingness to make drastic decisions without hesitation, which a conventional interview process reads as "leadership material," are exactly the same traits that best predict a high-risk psychopathic profile in corporate settings — and that same charisma facilitates, rather than hinders, advancement within structures that reward impression over evidence. What that profile brings, consistently documented, isn't order: it's relational violence, the destruction of entire teams, and measurable waves of negative consequences the organization pays for over years after the selection process read it as strength. It's the same underlying error again: confusing the signal — confidence, charisma, decisiveness — with the substance that signal is supposed to represent.

Evaluated separately, these almost always tell a different story than the start and end dates on a résumé — and often, in fact, tell the opposite story of what that résumé suggests at first glance. A profile with verifiable impact at every stop, with reasoned judgment behind every departure, and with integrity recognizable by independent third parties across those stops, doesn't describe someone unstable. It describes, more precisely, exactly the previous section's Ferrari: a person of high capability and sustained integrity, who up to now simply hasn't found the ecosystem with the right road for them — and whose résumé, read only by dates, disguises as risk what is, in reality, accumulated evidence of good judgment.

When Technical Truth Clashes with the Commercial Goal

There's one more tension almost never part of the leadership conversation, and one that in technical professions — architecture, engineering, medicine, auditing — carries a weight other fields don't share: the moment correct professional judgment and the commercial goal of whoever's in charge don't line up. A site resident who spots that a spec cut compromises structural safety; an architect who knows a delivery date promised to a client is technically unrealistic; an auditor who sees a figure that doesn't add up. The literature on professions has had a name for this clash for decades: professional-organizational conflict (extensively documented since the 1970s in research on professionals inside bureaucracies), the structural friction between the technical standard a profession requires and the pressure of the organization employing whoever practices it.

Robert Jackall, in Moral Mazes: The World of Corporate Managers (1988), the fruit of years of direct ethnographic observation inside American corporations, documented starkly how corporate hierarchies, under sustained commercial pressure, end up producing a moral logic of their own, distinct from the individual professional ethics of those inside them — one in which telling an inconvenient technical truth to whoever holds authority becomes, over time, the decision with the highest career cost, and softening or withholding it becomes the "reasonable" decision. It is, in essence, Flight 801 again, but with the power variable in place of the cultural one: the correct information existed, the person with the technical knowledge had it, and the cost of stating it clearly exceeded, within that specific system, the apparent cost of staying silent.

This connects back, precisely, to Part V of this article: a Kelley exemplar in a technical profession doesn't just need independent judgment and commitment — they need an ecosystem where flagging technical truth doesn't compete, in daily practice, against the commercial goal of whoever's evaluating their performance. When that competition exists and isn't systematically resolved in favor of technical truth, it isn't the person who has a loyalty or attitude problem — it's the system's design that's asking, without saying so in any official document, for technical integrity to lose to commercial convenience every time the two conflict.

The Treatment Received, Specifically, by Candidates for Leadership Roles

It's worth closing this section by pointing out something that follows directly from the status-threat mechanism described in Part IV, and that has an almost mathematical logic: the higher the level of the role in play, the greater the status the evaluator has to defend if the candidate demonstrates a capability that surpasses their own — and, following Exline and Lobel, the greater the likelihood that threat translates into hostile treatment rather than genuine recognition. This predicts, uncomfortably but consistently with the evidence, that selection processes for senior leadership and top management shouldn't be, on average, more civilized than those for operational roles — they should be, if the outperformance mechanism operates with the force the literature suggests, more prone to the image clash described in this article, not less, precisely because that's where the mismatch between the real candidate and the implicit prototype has the most at stake for whoever's evaluating.

What This Means for Every Actor in the System

For HR — if the selection process doesn't explicitly distinguish between styles, the default filter ends up being the interviewer's implicit prototype, not a deliberate criterion set by the organization.

For headhunters — the work starts before the first call with the candidate, and it has three layers rarely separated consciously. First: confirm that the job description the client handed over doesn't carry the role ambiguity from Part IX — "Principal" when they're really looking for "Job Captain," "Project Manager" when they're really looking for a shop head — because no later diagnosis fixes a posting that was mis-named at the source. Second: anticipate the image clash from Part IV — ask the client, with the same directness, what mental prototype they carry fixed in their head before meeting anyone, so it can be named and flagged before it operates unnoticed during the interview. Third, once those two are handled: the question posed to the client shouldn't just be "what profile are you looking for?" but "what style does this team need, given its maturity level and its current ecosystem?" Placing epistemic authority into an ecosystem of conformists, or positional command in front of a team of exemplars, isn't a matching error — it's a diagnostic error that happened before the matching even began, and that prior diagnosis has to cover all three layers at once, not just the last one.

For the Board of Directors, owners, and shareholders — the mental image of "what a competent leader looks like" the board carries internalized determines, without anyone declaring it as formal policy, who the organization promotes and who it protects for years, even when actual performance says the opposite. A board that never examines its own implicit prototype governs by aesthetic instinct, not by evidence.

For senior executive leadership — the executive level that designs strategy, structure, and culture: vice presidencies, functional general management, the immediate circle around the CEO, distinct both from the Board, which governs from outside day-to-day operations, and from middle management, which executes without having designed anything it executes — the most direct responsibility for the five ecosystem layers described in Part VIII falls here. This is, with enormous frequency, whoever delegates downward the responsibility for a result without delegating, along with it, the real authority to change the structure producing that result. When Kanter documents that middle management gets blamed for execution failures that are really design failures, senior leadership is, almost always, who designed — or never finished designing — whatever later failed below. The question that falls to this level isn't just what leadership to hire or promote beneath itself, but what structure, what culture, and what real distribution of power it is itself building so that leadership even has somewhere to function.

For middle managers — from front-line supervisor up through functional management, deliberately grouped here as one because they share the same structural position — this is where almost everything described in this article ends up breaking hardest. It's the only position on the org chart where all three leadership directions from Part VI occur simultaneously, every day, in the same person: leading downward, following upward, and coordinating laterally with peers who don't report to them either. Floyd and Wooldridge, in The Strategic Middle Manager (1996), documented the structural tension this produces: middle managers translate strategy downward without having taken part in designing it, and translate operational reality upward without real authority to change a decision already made above them. Rosabeth Moss Kanter, going back to The Change Masters (1983), already pointed out that this layer is, with enormous frequency, held responsible for execution failures that are actually senior leadership's design failures — precisely because it's the most visible and least protected layer when something breaks.

It's also, exactly, the point where the absence of structure and process described in Part VIII — the WBS/PBS/EPS chain — hits hardest: middle management rarely designs the system it has to execute, but is the one expected to answer for its results as though it had. And it is, almost by structural definition, the point of greatest exposure to Part IV's image clash in both directions at once: it can suffer backlash from above if its judgment contradicts its superior's prototype, and lateral violence from its own team if its authority is perceived as illegitimate or poorly earned. No other level of the org chart absorbs, simultaneously, as many of the frictions this article has named one by one. If an ecosystem is going to visibly fail somewhere, the structural evidence suggests it's going to fail there first — not because middle management is the weakest link, but because it is, without having chosen it, the link carrying the greatest simultaneous tension in the whole system.

Turnover Isn't the Other Level's Fault — and It Almost Never Has a Consequence for Whoever Produces It

There's an indicator almost no organization reads correctly, and it's worth naming separately because it isn't a problem specific to middle management — it's a problem that runs through every level of the org chart equally, including the very top: the voluntary turnover rate. Bennett Tepper, in a landmark study published in the Academy of Management Journal (2000), documented with data that abusive supervision directly increases voluntary turnover, and that those who stay under that same supervision show lower affective and normative commitment, often remaining not out of loyalty but because they feel trapped, with no real alternative. This doesn't respect rank: it happens the same way to a team under a front-line supervisor as to an entire division under a CEO whose style no one on the Board dares to name.

The pattern that keeps this going is almost always the same, and it has a name. Albert Bandura, in his reference work on moral disengagement (1999), documented that two of the most common mechanisms for evading responsibility for harm caused are displacement of responsibility — attributing blame to a higher authority who "made" someone act that way — and diffusion of responsibility — spreading blame across so many people or levels that no one ends up responsible for anything in particular — and that both mechanisms are, per his own analysis, rooted precisely in organizational and authority structures. Inside a real hierarchy, this looks like: middle management blames senior leadership for impossible targets designed without its input; senior leadership blames middle management for not knowing how to "retain talent"; the CEO blames the labor market, the younger generation, or culture in general — "people just don't commit like they used to" — and each level, looking somewhere else, avoids the uncomfortable question of how much turnover its own specific leadership style is producing. No one is consciously lying by doing this — Bandura is clear that moral disengagement rarely feels, from the inside, like deliberate evasion. But the result is the same whether or not it's intentional: responsibility dilutes until it disappears, and with it, any real consequence for whoever actually produced the departure.

The cost of sustaining this pattern isn't only human: the Society for Human Resource Management estimates, as a standard industry benchmark, that replacing an employee costs between six and nine months of their salary. Each departure gets filed and explained separately, as if it were an isolated coincidence, with no one connecting the dots between accumulated turnover and the real quality of the leadership producing it at each specific level — the same error of analytical level this article has been naming since Guam, now expressed in a figure that already sits in every HR report, just almost never read for what it is, and almost never reaching, without diluting along the way, whoever actually produced it.

And there's a twist that complicates the diagnosis even further, one rarely named: not all turnover is evidence of bad leadership. It's often evidence of good leadership doing the most uncomfortable part of its job. Will Felps, Terence Mitchell, and Eric Byington, in their landmark study on group dysfunction (2006), documented that a single destructive member within a team — what's colloquially called a "bad apple" — can reduce the entire group's performance by thirty to forty percent, and that this damage frequently persists precisely because those with the authority to act lack the real power or institutional will to do so. When a middle manager does act — removing that bad apple, or the weeds that have been choking the rest of the garden for a while — that departure enters exactly the same turnover statistic as any other, with no distinction made. And the middle manager who made the correct call can end up penalized not because they were wrong, but because that correction generated real, visible friction elsewhere: a replacement search that costs time and budget, the annoyance of other departments that depended on the person removed, the short-term operational disruption any change produces. The correct correction and the original problem look, in the quarterly report, exactly the same.

This reveals something no turnover figure alone can show: what the organization actually prioritizes, beyond what it declares. If the short-term cost of removing a bad apple — the friction, the search, the disruption — systematically weighs more, in the real evaluation of that middle manager, than the long-term cost of tolerating it, the organization is revealing, through its own behavior, that it prefers the convenience of a garden with tolerated weeds over the visible cost of keeping it healthy. And when that happens, the entire work culture ends up training its middle managers not to act — not because they lack judgment, but because the system, without saying so in any document, punishes correct action almost as hard as it would punish inaction. Breaking that pattern demands something more uncomfortable than a better dashboard: it demands that every level of the hierarchy, starting with the highest, stop looking at the level below when asked why people are leaving — and learn to tell, within that same figure, how much turnover is a symptom of harm, and how much is evidence that someone, finally, decided to heal the garden.

And there's something that complicates this scenario even further, worth naming precisely because it isn't the same mechanism as the leader prototype from Part IV, even though the two frequently clash at the same time. Charles O'Reilly, Jennifer Chatman, and David Caldwell, in a landmark study (1991) that gave rise to the Organizational Culture Profile, documented that organizations also operate with a culture mold — a profile of expected values, just as precise as the implicit leadership prototype, though almost never written down anywhere — against which every person gets measured, and that fit or misfit with that mold measurably predicts job satisfaction a year later and actual turnover two years later. It's, under a different name, the same kind of template Rivera documented as a class filter disguised as "cultural fit" — except this one operates even after hiring, not just at the door.

This means that in the bad-apple scenario, there are potentially two molds clashing at once, not one. The middle manager who decides to act may not fit the expected leader mold — too direct, too willing to generate visible friction instead of maintaining the surface harmony the implicit prototype rewards. And whoever ends up replacing the bad apple may, in turn, not fit the existing culture mold — precisely because that mold was built, over time, around what the weeds tolerated, not around what a healthy garden would actually need. When both clashes happen at once, the whole system can end up reading a correct intervention as a double anomaly: a leader who "doesn't fit" and a replacement who "doesn't fit either," when in reality both are signaling, from different angles, that the original mold was never designed to sustain health — it was designed, without anyone saying so, to sustain whatever was already there.

The Exercise That Falls to the CEO — Not as Audience, But as Subject

After mapping seven leadership archetypes, five followership patterns, three directions, and five ecosystem layers, the final question can't stay at "what leadership are you hiring?" It has to turn toward whoever has the most power in the room:

What kind of leadership do you exercise, right now, with the people in front of you — the leadership you believe you exercise, or the leadership your team would describe if no one was listening to them?

Almost no executive sees themselves as pure positional command — almost all of them narrate themselves internally as visionaries, mentors, or transformational leaders, even when the real pattern they produce is systematic alienation or silent conformity. And almost no executive who talks more than everyone else in every meeting sees themselves, either, as a beneficiary of the mechanism described in Part IV — the one that confuses volume with judgment, the one from the babble hypothesis documented by MacLaren and colleagues. The honest question that follows from that finding is uncomfortable and specific: how much of your perceived authority in the room depends on genuinely having the best available read of the problem, and how much simply depends on the fact that you spoke first, spoke most, and no one stopped to check whether you were right? The gap between the style a CEO believes they exercise and the one their own matrix of followers actually experiences is, probably, the most expensive and least audited blind spot in any organization. It is, literally, the same mechanism as the captain of Flight 801: he didn't see the accident coming, not because he didn't care, but because his own self-image as a competent pilot left no room to process the signal two people, with less formal authority than him, were already giving him.

After such a long journey, it's worth returning with its full weight to the phrase that opened this article: leadership is not a problem of character. It never was. Saying it at the start was a thesis. Saying it here, after seven styles, five followership patterns, five ecosystem layers, fifty thousand years of evolution, and an algorithm that amplifies all of it, is something else — it's the conclusion that has now earned the right to state itself without hedging. It doesn't mean character doesn't matter: integrity, congruence, the value of holding on to an uncomfortable technical truth against commercial pressure, all of that is real, and this article has defended it in every one of its parts. It means something more precise, and harder to accept: that no character, however solid, is enough on its own to produce leadership that works, if the system around it — structure, culture, power, psychological safety, the implicit prototype, the algorithm — isn't designed to recognize it in time. The captain of Flight 801 didn't have bad character. Most people who exercise backlash against whoever outperforms them don't have it either, most of the time. What they all have is a system that never taught them to distinguish between the signal that confirms what they already expected to see and the signal that should have made them stop. That's why it isn't enough to ask people to have better character, more discipline, or more self-awareness — even though none of that hurts. What's needed is to design systems that don't depend on people having it.

And there's a specific, widespread confusion this statement has to dismantle explicitly, so it doesn't remain half-finished: the idea that "strong" and "soft" character are the two ends of a single gradient, and that more leadership naturally lives toward the strong end. It is, in other words, exactly the same error already documented by name and evidence in Part IV: confusing volume, perceived firmness, or assertiveness — what's colloquially called strong character — with greater real leadership capacity, when Anderson and Kilduff, the babble hypothesis from MacLaren and colleagues, and the Grant, Gino, and Hofmann study itself already showed that firmness predicts perceived competence, not real competence, and that it performs worse, not better, than a quieter style with proactive teams. A soft character isn't a character with less leadership in it — it can be, just as often and with better documented business results, epistemic authority exercised without needing to impose itself. And a strong character isn't, by itself, evidence of more leadership — it can be, just as often, positional command or pure theatricality, guaranteed smoke with no substance behind it. The strength or softness of a character describes a temperament style. It doesn't describe, in any sense this article's evidence supports, how much real leadership stands behind it.

The best captains in the world crash too — not for lack of character, and not because someone else was right and they weren't. They crash for lack of a system capable of recognizing, in time, the otherness of whoever's in front of them — even, especially, when that otherness doesn't wear the same rank.

X. An Open Ending — When the Treatment Itself Becomes the Message

Everything above describes a diagnostic problem: the system measures wrong, recognizes wrong, hires wrong. But there's a final consequence worth naming here, even if it doesn't get fully resolved in this article, because it follows directly from the image clash described in Part IV — and because its cost doesn't stay contained in the interview room.

When a hiring process treats the mismatch between who a candidate really is and the fixed image an interviewer carries in their head not as useful information about the role, but as grounds for hostile treatment, rule changes without notice, or prolonged silence with no feedback, that treatment has a name in the literature and a measured consequence. Christine Porath, in research developed over nearly two decades with Christine Pearson, documented the cost of workplace incivility in "The Price of Incivility" (HBR, 2013): those who experience it don't just bear a personal cost, they deliberately reduce their effort, and they talk about the experience with others, with a measurable effect on reputation. Porath extended this specifically to the hiring process: candidates treated with incivility — unannounced scope changes, hostile treatment, no feedback — tend to share that experience publicly, whether or not they get the job. There is, in fact, a formal applied field within human resources dedicated exactly to this — candidate experience — which measures how treatment during hiring affects the employer brand, regardless of the process's final outcome.

Stated at the level of pattern, without naming anyone: when the fixed mental image of whoever's interviewing becomes the real evaluation criterion — when the candidate who doesn't fit the prototype is treated not with neutrality but with active hostility, following exactly the mechanism described with Eagly, Karau, Exline, and Lobel — that treatment doesn't stay contained in that room. It becomes, over time, with enough candidates going through the same thing, the public image of how that organization treats people even before it has any formal authority over them. And that image is precisely what HR and whoever leads hiring exist institutionally to protect, not to erode without realizing it.

This deserves its own development, with more room than this article has left — a separate piece, not a fourth part of the already-announced three-part series, but an introduction to another, entirely different conversation about what happens when the image clash between recruiter and candidate stops being a design flaw and becomes a method. It's announced here, deliberately left unresolved — because the goal of this article was never to close the conversation, but to catch, along the way, whoever is genuinely willing to look at their own system with the same rigor they demand of others.

The question this piece opened with still doesn't have a closed answer. How many times we've been the captain and not the co-pilot. How many times someone told us, with all the clarity their position allowed, something that didn't fit the idea we had of ourselves — and we didn't listen. Not every article needs to close with a formula; context decides the form, not a fixed rule. But this one's purpose was always concrete, and it's worth stating once more, plainly: it matters because the next time someone, with less formal authority than you, says something worth hearing, the system — not just the character of whoever's listening — should be designed so that, this time, it actually gets heard.

Nine thousand flight hours weren't enough that night over Guam. Not because the captain didn't know how to fly. Because no one, in that entire system, had designed a way to recognize, in time, that the person with less rank had a different and valid reading of the same reality — not a mistaken reading waiting to be corrected, but another legitimate way of seeing what he, from his own instrument and his own position, couldn't see. The best captains in the world crash too — not for lack of character, and not because someone else was right and they weren't. They crash for lack of a system capable of recognizing, in time, the otherness of whoever's in front of them — even, especially, when that otherness doesn't wear the same rank. We come hardwired not to recognize one another. The question left, after this whole journey, isn't how to eliminate that wiring — it's whether we're willing to design, on top of it, a system capable of doing what that wiring alone never would.

And there's one more question this article can barely open without fully resolving: recognizing one another honestly requires, first, a certain courage — the courage not to live according to the pattern the prototype, the algorithm, or the room expects of you. Before you can lead anyone else, you have to have exercised that same independent judgment on yourself — to be, in some sense still worth exploring, a leader of your own life before being one for someone else's.

And that courage carries a predictable cost, one this article has already documented without calling it by its full name: the genuinely authentic person almost always gets pushed back. Canceled, in the most literal, everyday sense of the word — not necessarily with a firing or a public scandal, but with the silent wear of Rudman and Glick's backlash effect, with the discomfort Exline and Lobel measured toward whoever simply performs better without asking permission, with Griffin's lateral violence wielded by peers who'd rather not have such a demanding mirror nearby. And this doesn't distinguish between leadership and followership: it happens the same way to the leader who directs with epistemic authority instead of imposing, and the same way to Kelley's exemplary follower, who questions with independent judgment instead of simply obeying. Both get punished for exactly the same reason — because their authenticity, without meaning to, sets a standard the rest of the system never agreed to sustain, and no system takes kindly to a standard it didn't ask for. Perhaps governing, in the most honest sense of the word, isn't imposing yourself on others — it's holding that standard long enough, despite the cost, for the whole system to eventually catch up to it. It's left announced here, deliberately open, for the third piece in this series.

References

Anderson, C., Brion, S., Moore, D. A., & Kennedy, J. A. (2012). A status-enhancement account of overconfidence. Journal of Personality and Social Psychology, 103(4), 718–735.

Adler, P. S., Goldoftas, B., & Levine, D. I. (1998). Stability and change at NUMMI. In R. Boyer, E. Charron, U. Jürgens, & S. Tolliday (Eds.), Between imitation and innovation: The transfer and hybridization of production systems in the international automobile industry (pp. 128–160). Oxford University Press.

Anderson, C., & Kilduff, G. J. (2009). Why do dominant personalities attain influence in face-to-face groups? The competence-signaling effects of trait dominance. Journal of Personality and Social Psychology, 96(2), 491–503.

Arthur, M. B., & Rousseau, D. M. (Eds.). (1996). The boundaryless career: A new employment principle for a new organizational era. Oxford University Press.

Bandura, A. (1999). Moral disengagement in the perpetration of inhumanities. Personality and Social Psychology Review, 3(3), 193–209.

Baron, R. A., & Neuman, J. H. (1996). Workplace violence and workplace aggression: Evidence on their relative frequency and potential causes. Aggressive Behavior, 22(3), 161–173.

Babiak, P., & Hare, R. D. (2006). Snakes in suits: When psychopaths go to work. HarperCollins.

Baron-Cohen, S., Wheelwright, S., Stott, C., Bolton, P., & Goodyer, I. (1997). Is there a link between engineering and autism? Autism, 1(1), 101–109.

Bass, B. M. (1985). Leadership and performance beyond expectations. Free Press.

Bass, B. M., & Avolio, B. J. (1994). Improving organizational effectiveness through transformational leadership. Sage.

Boehm, C. (1999). Hierarchy in the forest: The evolution of egalitarian behavior. Harvard University Press.

Branch, S., Ramsay, S., & Barker, M. (2013). Workplace bullying, mobbing and general harassment: A review. International Journal of Management Reviews, 15(3), 280–299. 

 

Brady, W. J., Wills, J. A., Jost, J. T., Tucker, J. A., & Van Bavel, J. J. (2017). Emotion shapes the diffusion of moralized content in social networks. Proceedings of the National Academy of Sciences, 114(28), 7313–7318.

Bronfenbrenner, U. (1979). The ecology of human development: Experiments by nature and design. Harvard University Press.

Burns, J. M. (1978). Leadership. Harper & Row.

Cain, S. (2012). Quiet: The power of introverts in a world that can't stop talking. Crown.

Compton, W. M., Conway, K. P., Stinson, F. S., Colliver, J. D., & Grant, B. F. (2005). Prevalence, correlates, and comorbidity of DSM-IV antisocial personality syndromes and alcohol and specific drug use disorders in the United States. Journal of Clinical Psychiatry, 66(6), 677–685.

Dweck, C. S. (2006). Mindset: The new psychology of success. Random House.

Eagly, A. H., & Karau, S. J. (2002). Role congruity theory of prejudice toward female leaders. Psychological Review, 109(3), 573–598.

Edmondson, A. C. (2019). The fearless organization: Creating psychological safety in the workplace for learning, innovation, and growth. Wiley.

Exline, J. J., & Lobel, M. (1999). The perils of outperformance: Sensitivity about being the target of a threatening upward comparison. Psychological Bulletin, 125(3), 307–337.

Feigenbaum, A. V. (1991). Total quality control (3rd ed.). McGraw-Hill.

Felps, W., Mitchell, T. R., & Byington, E. (2006). How, when, and why bad apples spoil the barrel: Negative group members and dysfunctional groups. Research in Organizational Behavior, 27, 175–222.

Floyd, S. W., & Wooldridge, B. (1996). The strategic middle manager: How to create and sustain competitive advantage. Jossey-Bass.

Gabarro, J. J., & Kotter, J. P. (1980, January). Managing your boss. Harvard Business Review.

Gladwell, M. (2008). Outliers: The story of success. Little, Brown and Company.

Goffman, E. (1959). The presentation of self in everyday life. Doubleday.

Goleman, D. (2000, March–April). Leadership that gets results. Harvard Business Review.

Greenleaf, R. K. (1970). The servant as leader. Robert K. Greenleaf Center for Servant Leadership.

Griffin, M. (2004). Teaching cognitive rehearsal as a shield for lateral violence: An intervention for newly licensed nurses. Journal of Continuing Education in Nursing, 35(6), 257–263.

Gronn, P. (2002). Distributed leadership as a unit of analysis. The Leadership Quarterly, 13(4), 423–451.

Grant, A. M., Gino, F., & Hofmann, D. A. (2011). Reversing the extraverted leadership advantage: The role of employee proactivity. Academy of Management Journal, 54(3), 528–550.

Hearn, A. (2008). Meat, mask, burden: Probing the contours of the branded self. Journal of Consumer Culture, 8(2), 197–217.

Heifetz, R. A. (1994). Leadership without easy answers. Harvard University Press.

Helper, S., & Henderson, R. (2014). Management practices, relational contracts, and the decline of General Motors. Journal of Economic Perspectives, 28(1), 49–72.

Hofstede, G. (1980). Culture's consequences: International differences in work-related values. Sage.

House, R. J. (1976). A 1976 theory of charismatic leadership. In J. G. Hunt & L. L. Larson (Eds.), Leadership: The cutting edge (pp. 189–207). Southern Illinois University Press.

Jackall, R. (1988). Moral mazes: The world of corporate managers. Oxford University Press.

Jones, E. E., & Pittman, T. S. (1982). Toward a general theory of strategic self-presentation. In J. Suls (Ed.), Psychological perspectives on the self (Vol. 1, pp. 231–262). Lawrence Erlbaum Associates.

Jost, J. T., & Banaji, M. R. (1994). The role of stereotyping in system-justification and the production of false consciousness. British Journal of Social Psychology, 33(1), 1–27.

Judge, T. A., & Cable, D. M. (2004). The effect of physical height on workplace success and income: Preliminary test of a theoretical model. Journal of Applied Psychology, 89(3), 428–441.

Judge, T. A., Bono, J. E., Ilies, R., & Gerhardt, M. W. (2002). Personality and leadership: A qualitative and quantitative review. Journal of Applied Psychology, 87(4), 765–780.

Kahn, R. L., Wolfe, D. M., Quinn, R. P., Snoek, J. D., & Rosenthal, R. A. (1964). Organizational stress: Studies in role conflict and ambiguity. John Wiley.

Kanter, R. M. (1983). The change masters: Innovation and entrepreneurship in the American corporation. Simon & Schuster.

Kelley, R. E. (1988, November–December). In praise of followers. Harvard Business Review.

Klofstad, C. A., Anderson, R. C., & Peters, S. (2012). Sounds like a winner: Voice pitch influences perception of leadership capacity in both men and women. Proceedings of the Royal Society B: Biological Sciences, 279(1738), 2698–2704.

Kotter, J. P. (1990). A force for change: How leadership differs from management. Free Press.

Kristof-Brown, A. L., Zimmerman, R. D., & Johnson, E. C. (2005). Consequences of individuals' fit at work: A meta-analysis of person–job, person–organization, person–group, and person–supervisor fit. Personnel Psychology, 58(2), 281–342.

Kruger, J., & Dunning, D. (1999). Unskilled and unaware of it: How difficulties in recognizing one's own incompetence lead to inflated self-assessments. Journal of Personality and Social Psychology, 77(6), 1121–1134.

Lewin, K., Lippitt, R., & White, R. K. (1939). Patterns of aggressive behavior in experimentally created "social climates." The Journal of Social Psychology, 10(2), 271–299.

Lord, R. G., Foti, R. J., & De Vader, C. L. (1984). A test of leadership categorization theory: Internal structure, information processing, and leadership perceptions. Organizational Behavior and Human Performance, 34(3), 343–378.

Lord, R. G., & Maher, K. J. (1991). Leadership and information processing: Linking perceptions and performance. Unwin Hyman.

MacLaren, N. G., Yammarino, F. J., Dionne, S. D., Sayama, H., Mumford, M. D., Connelly, S., & Martin, R. W. (2020). Testing the babble hypothesis: Speaking time predicts leader emergence in small groups. The Leadership Quarterly, 31(5), 101409.

Meindl, J. R., Ehrlich, S. B., & Dukerich, J. M. (1985). The romance of leadership. Administrative Science Quarterly, 30(1), 78–102.

Murphy, M. C., & Dweck, C. S. (2010). A culture of genius: How an organization's lay theory shapes people's cognition, affect, and behavior. Personality and Social Psychology Bulletin, 36(3), 283–296.

Ohno, T. (1988). Toyota production system: Beyond large-scale production. Productivity Press.

Ostrom, E. (1990). Governing the commons: The evolution of institutions for collective action. Cambridge University Press.

O'Reilly, C. A., Chatman, J., & Caldwell, D. F. (1991). People and organizational culture: A profile comparison approach to assessing person-organization fit. Academy of Management Journal, 34(3), 487–516.

Pittenger, D. J. (1993). Measuring the MBTI... and coming up short. Journal of Career Planning and Employment, 54(1), 48–52.

Porath, C., & Pearson, C. (2013, January–February). The price of incivility. Harvard Business Review.

Porter, M. E., & Kramer, M. R. (2011, January–February). Creating shared value. Harvard Business Review.

Rivera, L. A. (2015). Pedigree: How elite students get elite jobs. Princeton University Press.

Rudman, L. A., & Glick, P. (2001). Prescriptive gender stereotypes and backlash toward agentic women. Journal of Social Issues, 57(4), 743–762.

Schein, E. H. (1985). Organizational culture and leadership. Jossey-Bass.

Schmidt, F. L., & Hunter, J. E. (1998). The validity and utility of selection methods in personnel psychology: Practical and theoretical implications of 85 years of research findings. Psychological Bulletin, 124(2), 262–274.

Stinson, F. S., Dawson, D. A., Goldstein, R. B., Chou, S. P., Huang, B., Smith, S. M., Ruan, W. J., Pulay, A. J., Saha, T. D., Pickering, R. P., & Grant, B. F. (2008). Prevalence, correlates, disability, and comorbidity of DSM-IV narcissistic personality disorder: Results from the wave 2 national epidemiologic survey on alcohol and related conditions. Journal of Clinical Psychiatry, 69(7), 1033–1045.

Society for Human Resource Management (SHRM). (n.d.). Employee replacement cost benchmark: 6–9 months of salary. Industry benchmarking reference cited in SHRM leadership development and turnover research.

Tepper, B. J. (2000). Consequences of abusive supervision. Academy of Management Journal, 43(2), 178–190.

Uhl-Bien, M., Marion, R., & McKelvey, B. (2007). Complexity leadership theory: Shifting leadership from the industrial age to the knowledge era. The Leadership Quarterly, 18(4), 298–318.

Van Maanen, J., & Schein, E. H. (1979). Toward a theory of organizational socialization. Research in Organizational Behavior, 1, 209–264.

Van Vugt, M., Hogan, R., & Kaiser, R. B. (2008). Leadership, followership, and evolution: Some lessons from the past. American Psychologist, 63(3), 182–196.

Walton, G. M., & Cohen, G. L. (2011). A brief social-belonging intervention improves academic and health outcomes of minority students. Science, 331(6023), 1447–1451.

Watkins, M. (2003). The first 90 days: Critical success strategies for new leaders at all levels. Harvard Business School Press.

Webb, J. T., Amend, E. R., Webb, N. E., Goerss, J., Beljan, P., & Olenchak, F. R. (2005). Misdiagnosis and dual diagnoses of gifted children and adults: ADHD, bipolar, OCD, Asperger's, depression, and other disorders. Great Potential Press.

Weber, M. (1978). Economy and society: An outline of interpretive sociology (G. Roth & C. Wittich, Eds.). University of California Press. (Obra original publicada póstumamente en 1922)

Weick, K. E. (1995). Sensemaking in organizations. Sage.

White, H. A., & Shah, P. (2006). Uninhibited imaginations: Creativity in adults with attention-deficit/hyperactivity disorder. Personality and Individual Differences, 40(6), 1121–1131.

White, H. A., & Shah, P. (2011). Creative style and achievement in adults with ADHD. Personality and Individual Differences, 50(5), 673–677.

Zaleznik, A. (1977, May–June). Managers and leaders: Are they different? Harvard Business Review.

Ricardo Serrano Ayvar. 

All rights reserved. 2025.

bottom of page