How Managing a Quality Service is scored
Managing a Quality Service statements are usually scored on the Success Profiles 1–7 scale. This guide explains what 4, 5, 6 and 7 look like for this behaviour, and how moderation keeps scoring consistent.
The 1–7 scale in plain English
Assessors award one score per behaviour, normally from 1 (no evidence) to 7 (exceptional evidence). Most campaigns treat 4 as the benchmark: a clear, relevant example, competently evidenced. Scores of 5 and 6 reward depth; 7 is rare and reserved for statements that leave the assessor wanting nothing.
You are scored against the level advertised, not against other candidates. A precise, well-evidenced AO statement can score 6; a vague senior one can score 3.
What each score looks like for this behaviour
| Score | What the statement shows |
|---|---|
| 4 — pass | Relevant example; you organised work to a standard and gave a result, though checks or customer awareness may be thin |
| 5 — good | Clear standard, customers' needs shaped your choices, you verified quality, result measured |
| 6 — very good | All of the above plus judgement: priorities weighed, root causes fixed, quality sustained over time |
| 7 — exceptional | Complete, reflective evidence at the right level; standard, customers, checks, prevention and measured impact all explicit |
Below 4, statements usually miss the level (too junior), miss the behaviour (all pace, no quality), or miss the evidence (claims instead of actions).
How assessors decide between adjacent scores
The 4–5 boundary is about completeness: standard, customer, check and result all present, or not. The 5–6 boundary is about judgement: did you make reasoned choices about priorities and risks, and fix causes rather than symptoms? The 6–7 boundary is about polish and reflection — nothing generic, nothing missing.
For example, two SEO statements both cut queue times by 40%. One says how; the other also explains why the fix was chosen over cheaper options, how vulnerable customers were protected, and how the gain was locked in. Same result, different score.
Moderation: why scoring stays consistent
Departments do not leave scoring to one person's taste. A sample of sift scores is typically second-marked or moderated, and assessors are trained against the same level descriptors. Borderline and outlier scores get extra scrutiny, as of July 2026 this is standard practice in most campaigns.
The practical lesson: write for the descriptors, not for an imagined sympathetic reader. Evidence that matches the published indicators survives moderation; impressively written vagueness does not.
Tip
Score your own draft honestly. Give yourself a point for each of: named standard, customer needs, quality check, measured result, level-appropriate scope. Five points means a credible 5; anything missing shows you exactly what to add.
Frequently asked questions
Usually, but not guaranteed. Many campaigns set 4 as the benchmark per behaviour, yet competitive rounds may shortlist above the benchmark, and some campaigns score the application as a whole rather than per behaviour. The job advert or candidate pack should state the approach — read it before assuming a 4 is enough.
In principle, yes — the scale rewards completeness of evidence, not drama. In practice, 7s are rare because most statements leave something implicit: the standard, the check, or the prevention step. An ordinary example told with every element explicit and honest reflection can reach the top band. Do not chase it at the cost of clarity.
That is exactly what moderation exists for. Sifts typically second-mark a sample and review scores that sit at boundaries or look inconsistent, so a single harsh or generous read is usually corrected. Your protection is the same either way: evidence that plainly matches the level descriptors, because moderated scoring always returns to the published indicators.
Next step
Get the free STAR template — the same structure used in every example on this site.
Download the template pack