Skip to content
What You Are Measuring

All notes · Running it

Training Assessors

The highest-return intervention available to most organisations, and the one most often skipped. What a half-day should cover.

Running it · Procedure

Assessors are usually managers who were handed a rubric and a candidate. Training them is cheap and improves consistency more than any change of instrument.

For a recurring programme built around “Training Assessors”, improvement depends on seeing how the work is actually carried out between review cycles. daily work tracking can help teams organise assessor time and compare planned effort with delivery, but the resulting records should support calibration conversations rather than become an automatic judgement about an individual.

For an independent benchmark, compare the process with U.S. OPM assessment resources; the useful test is whether the local method remains job-related, proportionate and explainable.

What a half-day covers

What each stage is for and what it is measuring.

How to apply the rubric, with worked examples.

Note-taking: evidence rather than impression.

The common judgement errors, named and illustrated.

Practice: score three recorded or written answers, then compare.

And what not to ask, which is both a fairness matter and a legal one.

The errors worth naming

First impressions forming in the opening minutes and everything after being confirmation.

Similarity: rating people like yourself higher, which is well documented and invisible from inside.

Halo: one strong answer raising every other score.

Contrast: scoring against the previous candidate rather than against the criteria.

And leniency or severity drift, where an assessor's whole scale shifts over a day.

Note-taking

Write what was said, not what you thought of it.

"Described leading a migration, named the specific rollback decision and why" is evidence.

"Seemed confident and knowledgeable" is an impression, and it is what most interview notes contain.

This single change improves calibration and is what you will need if a decision is challenged.

The practice exercise

Three sample answers: strong, borderline, weak.

Everybody scores independently, then the group compares and discusses the differences.

The discussion is the training. People discover their criteria differ, which no amount of explanation achieves.

Refreshers

Annually, and before any high-volume round.

Drift is real and invisible: assessors who were calibrated a year ago are not now.

Fifteen minutes before a round, re-reading the rubric and one worked example, is most of the benefit.

Who should not assess

Anybody with a relationship to the candidate, declared and handled.

Anybody who has not been trained, however senior.

And anybody assessing alone on a decision that matters, because a single assessor's errors have nothing to correct them.

The resistance

Senior people frequently believe they do not need this.

The argument that works is consistency and defensibility rather than accuracy, which is heard as a criticism.

And the practice exercise usually persuades better than the explanation: discovering that two experienced managers scored the same answer three points apart is more convincing than any statistic.

What to check

Have your assessors been trained, or only briefed?

Do interview notes contain evidence or impressions?

When was the last calibration exercise?

And is anybody assessing alone on decisions that matter?

The point

Train assessors on judgement rather than briefing them on the form.

The practice exercise persuades better than any explanation.

Underlying all of this

Almost everything in this collection reduces to one discipline: write down what the job requires, assess that thing directly, record the evidence, and look at your own outcomes afterwards. None of it requires buying anything, and organisations that do those four things consistently outperform ones running longer processes built from instruments chosen before the requirements were known.

The recurring pattern

The recurring failure across every section here is the same: measuring what is convenient rather than what matters, then never checking whether it predicted anything. The check is an afternoon of work once a year, and it is the step that separates a process that improves from one that merely persists.

Independent guidance on skills assessment, selection design and fair hiring practice. External tools are included for practical comparison; evidence from the job remains the basis for decisions.