Personality tests, competency models, performance reviews, and assessment centres are four different instruments that get shopped as if they were one. Boards and CHROs that treat them as interchangeable end up owning a lot of data and no reading of judgment. Here is how to tell them apart.
Every Mexican board that has commissioned an evaluation in the last five years has been handed a report that felt rigorous, arrived on time, and did not, in the end, tell the room what it needed to know. The most common cause is not a bad instrument. It is the wrong instrument, elegantly executed. Four methods travel under the same commercial label — executive assessment, or assessment de directivos in Mexican corporate practice — and they answer four different questions. A CHRO or a chair who cannot say which is which will buy the wrong one at least half the time.
This is the plain-language taxonomy we walk clients through, before we scope any evaluation. It is not a ranking. Each method has a legitimate use. What matters is the fit between what the instrument reads and what the decision needs.
The four instruments, at a glance
| Method | What it reads | What it answers well | What it cannot answer |
|---|---|---|---|
| Personality-based assessment | Traits, style, preferences, motivations | Cultural fit, working style, coaching topics | Whether the person can hold the complexity of a specific role |
| Competency-based assessment | Observed behaviours against a defined model | Development gaps, role-fit against explicit criteria | Depth of judgment under conditions the competency model does not describe |
| Performance review | Past results in the current role | Delivery, reliability, contribution to date | Readiness for a role that runs on a different horizon |
| Assessment centre / capability assessment | Judgment observed under structured, role-relevant conditions | Depth at which the person can hold complexity through judgment | The full richness of long-form performance history (uses it as one input, not the reading) |
A competent evaluation practice can build a case for any of these. The problem is not that they exist. The problem is that they are shopped and sold under overlapping labels — sometimes deliberately, sometimes because the buyer did not ask the right question at the front of the process.
Personality-based assessment — useful, and not the reading a board needs
Personality-based instruments read stable traits and preferences: how a person tends to think, decide, respond to conflict, work with others, and process ambiguity. Well-constructed ones are psychometrically sound and produce reliable readings on the dimensions they are designed for. They are useful for team composition, coaching topics, and cultural fit conversations.
Where they are misused is at the moment of a governance decision. A rich personality profile handed to a board considering a CEO or business-unit leader appointment feels like it is answering the succession question. It is not. It is answering "who is this person, in general?" Boards need the answer to a different question: "can this person hold what we are about to hand them?" A personality profile is one input into that answer. It is not the answer.
One clear tell: if the instrument produces the same reading regardless of the role under consideration, it is a personality-based instrument. Role is not a variable in what it reads. That is fine — as long as the buyer knows what they bought.
Competency-based assessment — useful when the model is right, and hazardous when it is not
A competency-based assessment reads observed behaviours against a defined competency model — a set of named behaviours the organisation has decided matter, usually derived from strategy, values, or a job-family framework. Well-executed, it is useful for development planning, role-fit reads against explicit criteria, and internal talent reviews.
The method has two structural weaknesses that show up in senior-role decisions. First, the reading is only as good as the model, and models age. A competency framework built for the last operating context often keeps producing green ratings while the enterprise's actual demands drift somewhere else. Second, competency observation compresses a person's judgment into behavioural exemplars, which is a legitimate reading of what someone has done but a partial reading of what they can hold. The distinction matters at the level of appointments where the next role is not the same shape as the last one.
When an organisation asks us to evaluate for a role that materially differs from the incumbent's, we treat competency data as one input, not as the reading. The reading is elsewhere.
Performance review — history, not readiness
A performance review is the most familiar of the four and the most systematically misused in appointment decisions. It reads past results in the current role: what the person has produced, at what quality, over what period, judged against the goals the role was set. Well-run, it is the primary instrument for reward, promotion within the same shape of role, and calibration across a peer group.
What it does not answer is readiness for a role that runs on a different horizon. A person who delivers beautifully at a five-quarter operating rhythm may or may not hold the seven-year horizon of the next role. Nothing in the performance review can tell the board which, because it did not observe the person under the conditions the next role imposes. Using it as the primary evaluation instrument for a step-change appointment produces the pattern boards know too well: a highly rated executive appointed into a role they cannot hold, discovered eighteen months later.
Assessment centres and capability assessment — reading judgment under conditions close to the real ones
The fourth method is different in kind. An assessment centre — or, more precisely, a capability assessment — reads judgment observed under structured conditions built to approximate the complexity of the target role. The instrument does not ask what the person has done. It asks what they can do when the frame is unfamiliar, the signals are weak, the decision must be made under uncertainty, and the horizon runs beyond the visible.
A rigorous capability reading names three constructs separately: habilidad (ability — the practical skills the person has), capacidad (capability — the depth at which they can hold complexity through judgment), and alcance (capacity — the scope, scale, and load their contribution can carry). Ability and capacity are readable from history. Capability requires observation under load, which is what the method is designed to provide. The three readings are not interchangeable, and mistaking one for another is the most common failure of appointment decisions in the Mexican mid-market — brilliant operators appointed to roles that need longer-arc judgment, promoted on capacity when the mandate needed capability.
The method is also the one that most closely fits the question a board or an owner actually needs to answer at a governance decision: not "who is this person, in general?" and not "how has this person performed to date?" but "can this person hold this role, at this horizon, in this operating reality?" A leadership capability assessment built inside the Anker Bioss advisory model sits here.
How to tell which one a firm is actually selling you
Four questions distinguish the methods in practice, regardless of the label on the cover of the report.
First: does the reading change when the target role changes, and how? A personality-based instrument produces the same reading for any role. A competency-based instrument changes only if the competency model is role-specific. A performance review does not change. A capability assessment names the target role's timespan and complexity at the front of the engagement and returns a reading calibrated to it.
Second: what conditions was the person observed under, and how close are those conditions to the role's real ones? Personality instruments observe under self-report and third-party rating. Competency reviews observe under prior behaviour. Performance reviews observe under past goals. Capability assessment observes under structured conditions designed to approximate the target role's actual complexity, using dilemmas built from the enterprise's real environment.
Third: how does the report separate ability, capability, and capacity? A firm that answers this cleanly — showing where each reading comes from and where it does not overlap with the others — is doing capability work. A firm that produces a single composite score, a numeric grade, or an overall "potential" number is offering something else, however elegant.
Fourth: what does the firm refuse to say? Rigorous evaluation practices refuse to compress judgment into a single number, refuse to rank people against each other outside the context of a specific role, and refuse to trespass into stewardship territory that belongs to appreciation rather than evaluation. A firm that will produce any number you ask for is telling you something important about how they think about the reading.
Where the money actually leaks
Mexican boards spend real amounts on evaluation. The waste is not usually in the price of the instrument. It is in the mismatch between what was purchased and what the decision needed. A rigorous personality profile bought for a CEO appointment is not a bad instrument used badly — it is the wrong instrument, and buying more of them will not solve the problem. A competency review used to answer a step-change succession question is the same failure with a different vendor's letterhead.
The governance discipline is small, procedural, and quiet: name the question the room needs to answer before the instrument is chosen. If the question is "who is this person, in general?" a personality instrument is legitimate. If the question is "development gaps against our named competencies?" a competency instrument is legitimate. If the question is "contribution to date?" a performance review is legitimate. If the question is "can this person hold what we are about to hand them?" the room needs a capability reading. Confusing the four is what produces the pattern of appointments the market keeps regretting.
Where family-controlled and founder-anchored enterprises need to be most careful
In family-controlled enterprises, the mix of the four methods is especially consequential. Personality-based readings are read as "is this person like us?" Competency-based readings are read as "has this person done what our people do?" Performance reviews are read as "has this person delivered inside our system?" None of those answers the question the family is actually asking, which is "can this person hold the enterprise across the horizon we care about?" The gap is not a data problem. It is a stance problem — and the remedy is a reading built for the question the room is asking.
One pattern in particular: internal candidates get more of the wrong reading and external candidates get less of it. Internal candidates arrive with years of performance review data and often a competency profile; the room already knows their personality. What the room does not have — for either side — is a rigorous, role-calibrated reading of judgment. That is the gap capability assessment closes.
Frequently asked questions
Which method should we start with when we do not know? Start with the question, not the method. Name what the room actually needs to answer, and the method follows. If the question is a governance decision — CEO succession, business-unit leader appointment, board-seat readiness — the reading is a capability assessment, and the other instruments are inputs.
Can we combine methods? Yes, and it is usually the right answer at senior levels. The rule is that each method reads its own construct and none of them substitutes for the others. A capability assessment can use competency, performance, and personality data as inputs, but the reading — the answer the board decides against — is the capability reading.
How do we know a firm is offering capability assessment and not a competency instrument in different packaging? Ask them to name the target role's timespan and complexity before they design the assessment. Ask them how they separate ability, capability, and capacity in what they report. Ask them what they will not tell you. If any of those questions produces evasion or a scored composite, you are being sold a different instrument.
Is there a role for numerical scores in senior evaluation? Inside the instrument, yes — rigour has to be built somewhere. Outside the instrument, in what the board reads, no. Human judgment does not compress into a single number, and any report that offers one is doing something else. Rigour inside, prose outside. That is capability that complexity can't break.
Anker Bioss builds capability assessments as one part of an installed architecture — evaluation, capability, succession, and organisational design as one system. If the board is preparing to decide, or preparing not to have to decide under pressure, we can help.
