Quick takeaways
- What it is: A half-day to full-day selection event combining multiple exercises, scored by trained assessors against a fixed competency framework.
- Typical exercises: In-tray/e-tray, group exercise, presentation, role-play, written exercise, plus a psychometric test session and interviews.
- How scoring works: Competency-based, not exercise-based — the same competency is assessed across 2-3 different exercises and combined, not scored in isolation.
- Where it's used: Graduate schemes, management hiring, public-sector and Civil Service recruitment, and senior roles across most industries.
- Biggest trap: Treating each exercise as isolated instead of carrying a consistent, professional version of yourself through the whole day.
The invitation email is the most under-read document in the process
Somewhere in the email that got you here is a schedule, a list of exercises, and — surprisingly often — the actual competency framework you will be scored against. Most candidates skim it for the start time and the dress code, then go and practise numerical reasoning for a week. That is the wrong allocation of effort, and it is the single most common reason strong candidates come out of a day with a mediocre score sheet.
Here is the thing worth internalising before anything else: an assessment centre is not a series of tests you pass or fail. It is an evidence-gathering exercise. A group of trained assessors spends 6 hours to 8 hours trying to fill in a grid. Down one axis are competencies — usually 5 to 8 of them. Across the other are the exercises of the day. Their job is to put a defensible, evidence-backed score in each cell. Your job, whether anyone tells you this or not, is to make those cells easy to fill in favourably.
That reframe changes almost every tactical decision you will make on the day. It is why talking the most in a group exercise scores badly. It is why a quiet, well-structured 6 minute presentation beats a charismatic 14 minute one. And it is why recovering visibly from a bad exercise is worth more than most candidates realise.
What the day is actually built to measure
The method is older than most people assume. Structured multi-exercise selection has been in corporate use since the 1950s — AT&T's Management Progress Study, which began in 1956, is the usual reference point for the modern form, and it has been adapted continuously since. The underlying logic has not changed: triangulation. One good answer in one interview is weak evidence. The same behaviour appearing in three unrelated contexts, observed by different assessors who have not yet compared notes, is strong evidence.
So the day is designed to give the same competency multiple independent chances to show up. In the UK public sector this is formalised — the Civil Service Success Profiles framework has 5 elements (Behaviours, Strengths, Ability, Experience and Technical) and publishes 9 named behaviours, which means Civil Service candidates can read their own scoring rubric before they arrive. Private-sector employers are less consistent about publishing, but graduate schemes very often list their competencies on the careers site under some variant of "what we look for."
If your employer publishes theirs, stop reading general advice and go read that instead. It is the rubric.
The competencies themselves cluster fairly predictably: communication (written and verbal), analysis and problem-solving, decision-making under incomplete information, commercial or organisational awareness, working with others, influencing without authority, planning and prioritisation, and resilience. Different employers name them differently and weight them differently, but the underlying set is stable enough that preparing against it generically is not wasted effort.
A realistic full day: graduate scheme, 8 candidates, 1 site
The abstraction above becomes much easier to act on when you can see the shape of the day. Here is a representative itinerary for a graduate-scheme assessment centre — the kind run by a large professional services, banking, retail or public-sector employer for a cohort of 8 candidates in a single day. Exact timings vary, but the rhythm and the density are typical.
08:45 — Arrival, ID check, coffee. Nothing is scored. Something is observed. You will be in a room with the other 7 candidates and at least one member of the recruitment team, and how you behave in unstructured space is not on the grid but does reach the room.
09:00 — Welcome and briefing, 20 minutes. Someone senior explains the day, the exercises, and — usually — the competencies. Write them down. Candidates who can name the framework by 09:20 behave differently for the rest of the day than candidates who cannot.
09:25 — Psychometric test session, 60 minutes. Typically 2 tests back to back: a numerical reasoning test of roughly 20 questions in 25 minutes, and a verbal or logical reasoning test of similar length. Often these were sat online at an earlier stage and are being re-sat under supervision to verify the earlier result.
10:30 — Break, 15 minutes.
10:45 — In-tray or e-tray exercise, 45 minutes. An inbox of 12 to 20 items — emails, a voicemail transcript, a policy memo, a spreadsheet extract — with instructions to prioritise, action and justify. Increasingly delivered on a platform that timestamps every decision you make.
11:35 — Group exercise, 40 minutes. The cohort of 8 splits into 2 groups of 4, or runs as one group of 8 with 3 assessors in the corners. Typically 5 minutes of individual reading time, then 30 minutes of discussion, then a 5 minute group summary delivered to the assessors.
12:20 — Lunch, 45 minutes. Usually with current employees, often recent graduates from the previous intake. Treated by candidates as a break. Treated by the employer as a soft-signal collection window and, more importantly, as a selling opportunity — they are trying to convince the good candidates to accept.
13:10 — Presentation exercise, 30 minutes prep then 10 minutes delivery plus 5 minutes of questions. Sometimes the topic was sent in advance; increasingly it is handed over on the day, drawn from the same case pack used elsewhere.
14:30 — Role-play exercise, 15 minutes. One to one with a trained actor and 1 observing assessor. You get 10 minutes of briefing time beforehand.
15:15 — Competency-based interview, 45 minutes. Two interviewers, typically 5 to 6 questions, each mapped to a named competency.
16:15 — Close, next steps, 15 minutes. Decision communicated within 5 working days to 10 working days at most graduate employers.
Look at the density. That is 5 scored exercises plus an interview inside about 7 hours, with roughly 75 minutes of genuine break across the whole day. The fatigue is not a side effect of bad scheduling. Consistency under sustained load is itself one of the things being measured, and assessors do compare your 15:15 self to your 09:25 self.
The exercises, and what each one is really for
Each exercise exists because it generates evidence the others cannot.
In-tray / e-tray. Prioritisation and decision-making under time pressure, in writing, alone. The only exercise where your reasoning is captured verbatim and can be re-read after the fact.
Group exercise. Influence without authority, listening, and how you behave when other people's behaviour is outside your control. The only exercise where the input is genuinely unscripted.
Presentation exercise. Structured verbal communication to a passive audience, then to an adversarial one during Q&A. The only exercise that tests whether you can hold a line under direct challenge.
Role-play exercise. Interpersonal skill in a live, emotionally loaded one to one. The only exercise with a trained actor deliberately pushing you off your prepared approach.
Written exercise. Sustained analysis and written recommendation on one substantial case. The only exercise that rewards depth over speed.
Most centres also run a psychometric session and at least 1 competency interview alongside these.
One competency, three exercises: how "commercial awareness" actually gets scored
This is the part almost nobody shows candidates, so here it is concretely. Take a single competency — commercial awareness — and follow it through 3 exercises on the day above. The employer is a national retail chain. The case material across the day concerns a regional distribution centre running over budget.
Assessors are not writing "good commercial awareness." They are writing observable behaviour against indicators. Something close to this ends up on the sheets:
Exercise 1 — In-tray, 10:45. Item 7 in the inbox is a supplier email offering a 6 percent discount on packaging in exchange for moving from 30 day to 14 day payment terms. Item 11, three items later, is a finance memo noting the division's cash position is tight through Q3.
Assessor note (in-tray): "Declined the 6 percent packaging discount and referenced the Q3 cash constraint from item 11 in the written justification. Explicitly linked 2 unrelated inbox items. Flagged that the discount would be worth revisiting in Q4. — meets indicator: 'considers financial impact of decisions beyond the immediate task.'"
The behaviour being scored is not "knew the right answer about payment terms." It is connected two documents the exercise deliberately separated, and said so in writing. A candidate who took the discount because a discount is good money scores low here. So, importantly, does a candidate who declined it for the right reason but never wrote down why.
Exercise 2 — Group exercise, 11:35. The group has to recommend which of 4 cost-reduction options to put to the board. Two candidates are arguing for the option with the largest headline saving — closing a shift.
Assessor note (group): "Asked the group what the recruitment cost would be if volumes recovered in Q4 and the shift had to be restaffed. Reframed the comparison from headline saving to net saving over 12 months. Group changed its ranking as a result. — meets indicator: 'considers financial impact beyond the immediate task'; also evidences 'influences others' under a different competency."
Same competency, completely different behaviour, and note that one contribution has now generated evidence for 2 competencies at once. That is the highest-leverage thing you can do in any exercise, and it happens when you say the reasoning, not just the conclusion.
Exercise 3 — Competency interview, 15:15. The question is "tell me about a time you had to make a decision with incomplete financial information."
Assessor note (interview): "Used a part-time retail role. Described reducing a weekly stock order by 20 percent after 3 weeks of declining footfall rather than waiting for the monthly report. Quantified the outcome — waste down, no stockouts. Acknowledged the risk taken. — meets indicator: 'considers financial impact beyond the immediate task.'"
Three cells in the grid, three different exercises, three different assessors, one consistent competency. That is what a strong score sheet looks like — and it is why a single weak exercise rarely sinks a strong candidate. The reverse is also true and less comfortable: a candidate who is fluent and likeable but never once connects a decision to its consequences will have 3 empty cells in the same row, and that is the pattern that ends applications.
The actionable version: in every exercise, once per exercise, say the part most people leave in their head. "I'd deprioritise this because —". "Before we rank these, the thing I'd want to know is —". Assessors cannot score an inference they did not hear.
How the day changes with seniority and sector
The exercise names stay the same; the weighting and the difficulty move a lot.
Graduate and early careers. Heaviest emphasis on potential rather than demonstrated skill. Group exercises are near-universal because they are cheap to run at cohort scale. Psychometric tests carry real weight, and are often a hard sift before anyone reads anything you wrote. Expect 4 to 5 exercises in 1 day.
Experienced hire and management. Group exercises become rarer — running one requires a cohort, and mid-career hiring is usually not cohorted. Role-play and written exercises get heavier. The case material stops being generic and starts looking like the actual business. Expect a half day with 3 exercises, often spread over 2 separate visits.
Senior and executive. Frequently a bespoke day built around a single realistic scenario, with a business simulation, a stakeholder role-play with a professional actor briefed on a specific conflict, and a presentation to a panel that includes people who would be your peers. Psychometrics shift from ability testing toward personality and leadership-style questionnaires, often followed by a feedback conversation with an occupational psychologist.
Public sector. More procedurally rigid, more transparent, and more literal about the framework. If the framework names 9 behaviours, the scoring will follow those 9 behaviours closely, and evidence that does not map to one of them may simply not be recorded. Use their vocabulary.
Consulting and finance. Case-heavy. The written and presentation exercises dominate, and commercial reasoning is weighted far above interpersonal warmth relative to other sectors.
How to approach the day
Name the framework before 09:30. If it is published, learn it beforehand. If it is not, capture it during the briefing.
Say your reasoning out loud, once per exercise, deliberately. See the worked example above. This is the highest-return habit available to you.
Treat the transitions as part of the day. The 15 minutes between exercises is when assessors are writing. It is also when candidates who have decided the last exercise went badly start behaving like it.
Assume every exercise is independently scored, because it mostly is. Assessors are usually trained to score their own exercise before any group discussion, precisely to stop a halo or horns effect carrying between exercises. A bad 10:45 does not have to become a bad 11:35 — unless you carry it there yourself.
Be the same person at 16:00 as at 09:00. Consistency is not a soft virtue here. It is an explicit scoring dimension in most frameworks.
Common mistakes
The biggest is uneven preparation: 10 hours on numerical reasoning, 0 hours on the group exercise. Candidates prepare for what they can practise alone, which is exactly backwards — the exercises you cannot easily rehearse are the ones where preparation has the largest marginal effect, because so few of your competitors will have done any.
The second is visible competitiveness. Modern frameworks score collaborative and influencing behaviour; several explicitly score dominance negatively. Cutting across another candidate to make your point is not assertiveness on the sheet, it is a missed indicator for listening.
The third is treating lunch as either off-record or as another exercise. Both are errors. Nobody is scoring your sandwich, but the recent graduate hosting your table will be asked for an impression, and candidates who audibly perform through lunch stand out in the wrong direction.
The fourth is silent recovery. If an exercise goes badly, the damage is one cell. Candidates who then disengage, apologise repeatedly, or go quiet for the next 90 minutes convert one weak cell into a row of them.
The fifth is answering without showing the reasoning — covered at length above because it is the one that costs the most.
How to prepare
- Read the invitation properly, twice. Extract: the exercise list, the timings, the competencies, and whether preparation materials are permitted.
- Practise each exercise type separately. They reward genuinely different skills. The linked guides below go deep on each.
- Build 6 real examples, not 6 scripts. Use a structure such as STAR (Situation, Task, Action, Result) so that each one can be re-cut for different questions. Six flexible examples cover a 45 minute competency interview better than 15 memorised answers.
- Do 1 timed run of each written format. A 45 minute in-tray and a 90 minute written exercise both feel completely different once you have actually run the clock down on one.
- Rehearse recovery explicitly. Deliberately practise a presentation where you lose your thread, and practise continuing. The skill being built is not fluency, it is resumption.
- Learn the employer. Not the marketing copy — the last annual report, the last 2 pieces of press coverage, and one genuine question about the direction of the business. Commercial awareness is the competency most often scored and least often prepared.
- Sleep. A 7 hour day with 75 minutes of break is an endurance event, and the exercises that matter most are usually after lunch.
Related skill hubs
Provider guides that use these exercises
Frequently asked questions
If I bomb one exercise, is the day over?
No — and the structure is deliberately built so that it is not. Because scoring is competency-based across multiple exercises, a competency assessed in 3 places survives 1 weak showing. What does not survive is a consistent gap: the same indicator missing in all 3 cells. Practically, one bad exercise out of 5 is recoverable; the same weakness appearing in 3 exercises is the pattern that ends applications.
Are the other candidates my competition?
Usually not directly. Most graduate employers score against a fixed standard, not a curve, which means in principle all 8 candidates in a cohort can be made offers, and in practice more than 1 frequently is. Group exercises are scored individually even though the task is shared. Behaving as though it is a knockout is both inaccurate and, because dominance is scored negatively, actively costly.
How much of an assessment centre is online now?
A large share, and the split is often hybrid rather than all-or-nothing: psychometric sessions and in-tray exercises delivered remotely, then a shorter in-person day of 4 hours to 5 hours for the group, presentation and interview elements. Role-plays run perfectly well over video call, which is why they were among the first exercises to move.
Do assessors talk to each other during the day?
Typically not about scores. The standard design has each assessor score their own exercise independently before a wash-up meeting at the end of the day, specifically so impressions do not leak between exercises. They will, however, rotate — the assessor watching your group exercise is often not the one interviewing you 4 hours later, which is exactly why consistency across the day is visible to the panel even when no single person watched you all day.
What do assessors actually write down?
Behavioural evidence, in something close to the format shown in the worked example above: what you did, what you said, and which indicator it maps to. Not impressions. This is why concrete, observable action beats a general sense of going well — an assessor who liked you but has 2 blank lines under an indicator cannot score you on it.
Is lunch scored?
Not formally, at almost any employer. But you are in the building with employees who will be asked what they thought, and lunch is genuinely also a sales pitch aimed at the candidates the employer wants. The workable posture is ordinary professional warmth: ask real questions, eat something, do not perform.
How long until I hear back?
Most graduate employers communicate within 5 working days to 10 working days, because a cohort day produces all the evidence at once and the wash-up meeting usually happens the same afternoon. Longer silences more often reflect an internal approval step than a borderline score.
See TestSolve in action first
Watch a 2-minute walkthrough, then download when you're ready. No subscription, no signup.
TestSolve is independent and not affiliated with any test provider or employer named on this page. All product names and trademarks belong to their respective owners.