Skip to main content

Talent reviews

What is the 9-box grid? How to use it for talent reviews

The grid is only ever as good as the conversation around it. What the two axes measure, what each of the nine boxes means, and how a placement becomes a succession plan.

By the CLEAR Talent team14 min read

Organisation-wide performance report in CLEAR Talent

Key takeaways

  • The grid exists to separate two questions a single rating blurs: how someone performs now, and how much headroom they have for a different role.
  • Performance is evidenced from work you have already done. Potential is a judgement — define it in observable terms, in writing, before anyone rates it.
  • A box is a prompt for a conversation, not a verdict. Nothing useful happens until every placement has an action, an owner and a date.
  • A well-populated top-right box is not a succession plan: “ready for more” is not the same as “ready for this role”.
  • Calibrate placements across managers and revisit them on a cycle, or the grid preserves last year’s opinions indefinitely.

The 9-box grid has outlived several confident predictions of its demise, and the reason is straightforward: it asks two different questions about a group of people, on one page, at the same time. How well is this person doing the job they hold — and how much headroom do they have for a different one?

It also fails in predictable ways. Potential gets rated as though it were a second performance score, boxes harden into labels people carry for years, and the grid that took an afternoon to fill in is never opened again. This guide covers what each axis actually measures, what all nine boxes mean, a template you can copy, and how to get from a placement to a succession plan somebody owns.

What the 9-box grid is, and what it is for

A 9-box grid is a three-by-three matrix. One axis carries current performance, usually rated low, moderate or high. The other carries potential — the judgement about whether someone could take on a larger or different role, and how soon. Everyone discussed lands in one of the nine boxes, and the grid shows the whole group at once. You will also see it called the nine-box talent matrix or the performance-potential matrix; the mechanics are the same.

The value is in the separation. A single performance rating collapses two things that behave quite differently: the specialist who is excellent in a role they have no wish to leave, and the person who is merely solid today because they are six months into a job that stretches them. Rank both on one scale and they look similar. Place them on two axes and they sit in opposite corners, which is a far more useful starting point for deciding what each of them needs next.

The grid is a discussion tool, and it is at its most dangerous when treated as though it produced decisions. A placement is a summary of what a room believed on a particular afternoon. That is a reasonable place to begin a conversation and a poor place to end one.

A box is where a conversation starts. It is not a verdict, and it is not a permanent property of a person.

The two axes: one is evidenced, the other is judged

Treating both axes as the same kind of rating is the most common structural mistake, and everything downstream inherits it. Performance is retrospective and evidenced: goal results, review ratings, competency assessments and feedback from the period. Most organisations already hold all of it, and the grid should read that evidence rather than reinvent it on the day.

Potential is a prediction, and predictions about people are where bias does most of its work. Left undefined, potential quietly defaults to visibility — proximity to decision-makers, confidence in a meeting, resemblance to the people doing the rating. The answer is not to drop the axis. It is to say, in writing and before the session, what this organisation means by it.

Ability to learn
Does this person pick up unfamiliar work quickly, and do they change how they operate when the evidence says they should? The useful examples are the ones where somebody was wrong and adjusted, not only the ones where they were right.
Aspiration
Does the person want a larger or different role? Willingness to relocate, to lead or to step up belongs on the grid as a stated fact rather than an assumption — and it is a question to ask the person, not about them.
Behaviours that would scale
How someone works under pressure, how they treat people who cannot help them, whether colleagues seek them out. This is where competency assessments and 360° input earn their place, because the manager is not the only person who has seen it.
Evidence of stretch
Has the person done something meaningfully harder than their job description, and how did it go? A potential rating with no stretch assignment behind it is a judgement about someone nobody has yet seen tested.

Whatever definition you settle on, put it in front of the room during the session. A shared definition is what turns nine boxes into a common language rather than nine labels each manager reads slightly differently.

It is worth being explicit about one more thing: potential describes suitability for a different role, not the worth of a person. Someone placed in the low-potential column is not failing. They may be doing exactly the work the organisation needs from them, at a standard that would be expensive to replace.

What each of the nine boxes means

The labels below are the conventional ones — use whatever language your organisation already understands. What matters is that every box carries an agreed action, because a grid where six of the nine boxes have no response attached is a picture rather than a plan.

Template

9-box grid template: what each box means and what follows it

Rows are grouped by performance band; the first column is the potential rating. Each line pairs a placement with the action it should trigger.

9-box grid template: what each box means and what follows it
BoxPotentialWhat the placement saysAction that follows
High performance
StarHighDelivers now and has headroom for a bigger roleStretch assignment, succession candidate, retention plan
High performerModerateStrong in role with room to grow within or near itBroaden scope, develop towards a named next role
Trusted professionalLowExcellent where they are, no appetite or headroom to moveRecognise, retain, use as mentor and subject expert
Moderate performance
Emerging talentHighNew or stretched, showing clear signs of headroomCoach closely, set a defined stretch task, review in six months
Core playerModerateDependable — most of the organisation lives hereKeep developing in role; do not treat the middle as a problem
Solid performerLowMeets the standard, unlikely to move beyond itMaintain performance, keep skills current, reassess if the role changes
Low performance
EnigmaHighSigns of headroom, results not there yetDiagnose the cause: role fit, clarity, support or motivation
Inconsistent performerModerateFalling short of a standard they could meetName the gap specifically, agree support and a date to review
UnderperformerLowNot meeting the standard, no current evidence of headroomAddress performance directly through your normal process

The labels are conventional rather than prescriptive — many teams use numbers or their own wording instead. Select the table to paste it into a spreadsheet, or download it below.

Download this template as a CSV file

Two boxes cause most of the argument. Low performance with high potential is a diagnosis waiting to happen: the same placement fits someone in the wrong role, someone promoted without support and someone who has simply stopped trying, and each needs a different response. High performance with low potential is the box organisations handle worst — it usually holds the people who would be hardest to replace, and treating “low potential” as a polite way of saying “no longer of interest” is how they end up taking a recruiter’s call.

Resist the pull of the top-right corner more generally. Most organisations run on the middle of the grid, and a talent process that spends all of its attention on one box in nine is not really a talent process.

Running the talent review around the grid

The grid is the artefact; the session is the work. Most of what makes a placement defensible happens before anyone is put in a box.

  1. Agree the population and both definitions

    Decide which group is being discussed and confirm what each axis means in writing. If the definition of potential is still being debated during the meeting, placements made before the debate are not comparable with the ones made after it.

  2. Bring performance in — do not re-rate it

    Performance should arrive from the completed review cycle, ideally after calibration. Re-rating it in a talent review quietly undoes the calibration you already ran and introduces a second, undocumented standard.

  3. Place people individually, then compare

    Ask each manager for a first placement and their reasoning before the room hears anyone else’s. Placements offered sequentially converge on whoever spoke first and most confidently.

  4. Work the disagreements, not the whole grid

    Spend the time on the people whose placements differ between raters, who have moved a box since the last review, or who sit on a boundary. Everyone else can be confirmed quickly.

  5. Record the reason next to the box

    A placement with no rationale attached is indistinguishable from a guess six months later — and the rationale is the first thing anyone asks for when a promotion decision is questioned.

  6. Leave with actions, not a picture

    Everyone discussed should leave the session with a next step, a named owner and a review date. The exercise is finished when the actions are, not when the boxes are full.

A talent review has the same failure modes as a calibration session, for the same reasons: the most confident advocate, the most recent memory, the person discussed last. If you already calibrate performance ratings, run the talent review with the same discipline and much of this comes for free.

How to prepare and run a calibration session

Common pitfalls, and how to avoid them

Potential rated as a second performance score
If the two axes correlate almost perfectly, the room is rating the same thing twice and the grid has collapsed into a ranked list. A neat diagonal of placements is a warning sign, not a result.
The box becomes a label
A placement describes a moment. Once “she’s a box nine” enters the vocabulary it follows people into decisions they were never placed for, and the annual refresh becomes a formality that confirms last year’s language.
Distribution used as a quota
Spread is a useful diagnostic: if four in five people land in the top row, the performance standard is the problem. Imposed as a quota, it turns placement into an allocation exercise and costs you the managers’ trust in the whole process.
No written definition of potential
Without one, potential defaults to visibility and nobody in the room can explain why two similar people were placed differently. That is also the version of the grid least likely to survive a challenge.
A one-off exercise
A grid built once and never revisited preserves a set of opinions well past their expiry date. Refresh it on a cycle and look at movement between refreshes — the direction of travel usually tells you more than the box does.
A process nobody can describe
Organisations differ on whether individuals are told their placement, and both positions are defensible. What is hard to defend is secrecy about the process itself: people should know that talent reviews happen, what the axes mean and how the output is used, even where the grid stays confidential.

From a box to a succession plan

The grid answers a question about a group. Succession planning answers a question about a role, and conflating the two is where most talent reviews lose their value. A well-populated top-right box is not a succession plan, because “ready for more” is not the same as “ready for this”.

Getting from one to the other means changing the unit of analysis. Start from the critical roles rather than from the people: for each one, who are the candidates, how ready is each of them against what that role actually requires, and what would close the distance? Readiness bands — ready now, ready in one to two years, ready in three or more, not a successor for this role — do more work here than the grid, because they are anchored to a specific job rather than to general headroom.

Then compare the candidates for a role like with like: readiness against that target role, performance trajectory rather than last cycle alone, multi-rater input on the behaviours the role needs, and whether the person actually wants it. Whatever the evidence set is, it must be the same for everyone in the pool. A comparison where one candidate arrives with competency data and another with a manager’s narrative measures preparation rather than readiness.

Finally, treat a role with a single name against it as a finding rather than a plan. Single-successor roles, and critical roles with no credible candidate at all, are the output of the exercise that most deserves a leadership team’s attention.

Guide: comparing succession candidates side by side

What actually changes: development actions

Everything above is preparation for the only part that changes an outcome — what people do differently afterwards. For most of the grid that means development actions specific enough to work on: attached to a named capability, with an owner, a date and an agreed idea of what progress would look like.

The most useful actions are usually work rather than training. A stretch assignment, a project with real consequences, deputising while someone is on leave, a piece of the manager’s job handed over for a quarter. These also generate the evidence the next talent review would otherwise be missing, because a potential judgement made after someone has been tested is a different kind of judgement from one made before.

Where the gap is genuinely a capability gap, size it against the role being planned for rather than the role being held, prioritise the two or three that matter most, and reassess on the same scale next cycle. That is what makes it possible to say whether any of this worked.

Explore employee development and IDP software

The 9-box grid in CLEAR Talent

CLEAR Talent’s succession planning includes an interactive 9-box grid with configurable axes, live filtering by role family, region or diversity dimension, and drag-and-drop placement where every move is written to an audit log. It draws on the competency assessments, 360° feedback and performance scores the review cycle already produces, rather than a separate dataset assembled for the meeting.

From there the grid connects to the rest of the plan. Readiness ratings — ready now, ready in one to two years, ready in three or more, or not a successor — each carry the specific competency gaps that would need to close to move up a category, and an assessment drafts an individual development plan from the competencies that matter most, ranked by the size of the gap. Candidates for the same role can be compared side by side against the same evidence. Dashboards flag critical roles with fewer than two successors, single-point-of-failure positions and diversity gaps in the pipeline, and the pack exports to PDF or Excel for a board or nomination committee.

None of that decides who gets the job. What it removes is the part of the decision that was being made from memory, from a deck assembled the week before, and from whoever happened to be discussed last.

Explore succession planning software

Frequently asked questions

Is the 9-box grid still worth using?
It is still widely used, and most criticism of it is really criticism of how it gets applied. A grid filled in from memory, used as a label and never revisited deserves its reputation. A grid built from review and competency evidence, with a written definition of potential and an agreed action for every box, is a straightforward way to look at a group of people and decide what each of them needs next.
What are the nine boxes called?
There is no standard naming. A common set is star, high performer, trusted professional, emerging talent, core player, solid performer, enigma, inconsistent performer and underperformer, and plenty of organisations use numbers or their own wording instead. The labels matter far less than the action agreed for each box — and names that sound like a verdict tend to stick to people longer than they should.
How do you measure potential?
You do not measure it, you judge it, which is exactly why the definition has to be written down before anyone rates anyone. Most workable definitions combine ability to learn, aspiration for a larger or different role, behaviours that would scale, and evidence from a stretch assignment. Ask people directly about aspiration and mobility rather than inferring it: assuming someone wants a job they do not want is a common cause of a wrong placement.
Should employees be told which box they are in?
Practice differs and both approaches are defensible, provided the organisation is consistent and can explain what it does. Many teams share the development and readiness conversation without sharing the grid itself. Whatever you decide, apply your data protection and access rules to the record, and make the existence and purpose of the process known rather than running something nobody can describe.
How often should a 9-box grid be refreshed?
Once a year, aligned to the review cycle, is the common rhythm, with a lighter check mid-year where something has changed materially — a new role, a significant stretch assignment, a change in aspiration. What matters more than frequency is looking at movement: a placement that has not changed in three years is either a genuinely stable picture or a conversation nobody has reopened.
Can you run a 9-box grid in a spreadsheet?
Yes, and it is a sensible way to run the first one. The limits appear afterwards: the placements sit apart from the evidence behind them, nobody can see what moved since last time, and the file is rebuilt by hand before every session. The point to move is usually when the grid has to connect to review data, development plans and succession on the same record.

See how this works in practice

Book a walkthrough focused on your review cycle, your goal structure, and the decisions your managers actually have to make.