Skip to content
What You Are Measuring

All notes · Fairness

Assessment Anxiety and Its Effects

Anxiety degrades performance on assessments more than on the job. What that means for what you are measuring, and what reduces it.

Fairness · Analysis

General orientation rather than clinical advice.

Once the criteria described in “Assessment Anxiety and Its Effects” are fixed, the practical challenge is running them consistently across assessors. A coordinator can use this practical resource to plan review time, compare the effort spent at each stage and spot avoidable delays, without confusing administrative efficiency with the quality of the evidence collected from candidates.

For an independent benchmark, compare the process with WHO guidance on mental health at work; the useful test is whether the local method remains job-related, proportionate and explainable.

Somebody who performs well at work may perform badly in an assessment, and the gap is not random: it falls hardest on people with least experience of such situations.

Why it matters for validity

If anxiety degrades performance, the assessment is measuring composure under observation as well as capability.

For most roles, composure under observation is not a requirement.

Which means anxiety is measurement error, and it is error distributed unevenly — by experience with assessments, by condition, and by how much is riding on the outcome.

What raises it

Unclear expectations: not knowing what will happen, how long, or how it is scored.

Observation, particularly a panel.

Artificial time pressure beyond what the job involves.

Recorded one-way video, which candidates consistently report as the worst format.

And high stakes for the candidate, which you cannot remove.

What reduces it, cheaply

Telling candidates what will happen, in detail, in advance. The format, the length, the criteria, who will be present.

Sending questions or the exercise brief ahead where it does not compromise the assessment — which for behavioural questions it does not.

A short warm-up question that is not scored.

Letting candidates ask questions before starting.

And keeping panels small. Three people is plenty; five is intimidating and adds nothing.

Stereotype threat

Documented effect: reminding somebody of a group identity before a test can depress their performance on it.

Practically: avoid framing assessments as measures of raw ability, avoid demographic questions immediately before a test, and describe assessments as job-relevant tasks rather than as tests of aptitude.

Small wording changes, measurable effects.

Time pressure

Only apply it where the job applies it.

A generous limit still constrains, and it removes a source of error that has nothing to do with the work.

Where speed genuinely matters, say so and assess it deliberately.

What not to do

Deliberate stress interviews. They have no demonstrated validity, they damage your reputation, and they select for a trait most roles do not need.

Surprise exercises sprung on the day.

And long silences intended to see how somebody copes, which is not a technique.

The limit of accommodation

You cannot remove the stakes and should not pretend to.

What you can do is remove the avoidable additions: uncertainty, artificial pressure, unnecessary observation.

Those are most of it, and removing them improves the accuracy of the measurement as well as the experience.

What to check

Do candidates know the format, length and criteria in advance?

How many people sit on your panels?

Is time pressure job-relevant or inherited?

And does anything in your process resemble a deliberate stress test?

The point

Anxiety is measurement error distributed unevenly.

Telling candidates the format, length and criteria in advance removes most of the avoidable part.

Underlying all of this

Almost everything in this collection reduces to one discipline: write down what the job requires, assess that thing directly, record the evidence, and look at your own outcomes afterwards. None of it requires buying anything, and organisations that do those four things consistently outperform ones running longer processes built from instruments chosen before the requirements were known.

The recurring pattern

The recurring failure across every section here is the same: measuring what is convenient rather than what matters, then never checking whether it predicted anything. The check is an afternoon of work once a year, and it is the step that separates a process that improves from one that merely persists.

Independent guidance on skills assessment, selection design and fair hiring practice. External tools are included for practical comparison; evidence from the job remains the basis for decisions.