Absolute against relative
| Absolute | Relative | |
|---|---|---|
| Compares against | A defined standard | Other employees |
| Possible outcome | Everyone can be rated highly, or poorly | Someone is always at the bottom |
| Small teams | Works | Arbitrary by construction |
| Main risk | Inflation and drift between managers | Damage to collaboration and to the record |
| Answers | Is this person performing the role well? | Who should get the money? |
The bottom row is the crux. Absolute ratings answer the question a manager and an employee actually care about. Relative ranking answers the question the compensation budget poses. Organisations that need both, which is most of them, keep drifting between the two and settling on a compromise that does neither well.
The stable arrangement is to rate absolutely and allocate separately: assess each person against the standard, then distribute a fixed pot against those assessments. The rating stays honest and the budget stays controlled, at the cost of admitting openly that a high rating does not guarantee a large increase.
Building a scale that works
An absolute rating is only as good as the standard it measures against, and most scales fail because the levels are labels rather than descriptions.
- Describe behaviour at each level, not adjectives. Exceeds expectations means nothing; delivers work that others rely on without rework, and improves how the team does it, can be assessed.
- Keep the scale short. Five levels are argued over; four or three are usually enough and reduce the pretence of precision.
- Avoid a middle that is treated as failure. Where the central rating denotes solid performance and the organisation treats it as a disappointment, managers stop using it and the scale collapses upward.
- Separate what from how where the organisation cares about both, and be explicit about the weight each carries.
- Anchor the levels with examples from the actual work, refreshed periodically.
The test is whether two managers reading the descriptions would place the same employee at the same level. Where they would not, the scale is a container for opinion.
Drift and inflation
The characteristic failure of absolute rating is that standards move without anyone deciding to move them.
- Managers avoid difficult conversations, so ratings creep upward year on year.
- A manager under pressure to retain someone rates them highly to support a counter-offer.
- Different functions develop different implicit bars, so a rating in engineering does not mean what it means in operations.
- Ratings become linked to increments so tightly that the rating is chosen from the desired increase and worked backwards.
Calibration is the standard remedy and it has to be genuine: managers comparing evidence, not negotiating numbers. The diagnostic for whether drift has taken hold is to look at the distribution over three years. If the proportion in the top band has risen steadily and organisational performance has not, the scale is inflating.
The other remedy is to decouple rating from increment enough that the rating is not worth gaming. Where the two are welded together, no amount of calibration will hold the standard.
When absolute ratings are the right choice
They fit where the work is not zero-sum and where the standard is describable.
- Small teams, where relative ranking is arbitrary and known to be.
- Roles with clear output standards, where meeting the standard is the point rather than beating a colleague.
- Collaborative work, where relative ranking actively discourages the behaviour the organisation wants.
- Organisations that need a defensible record, since a rating against a described standard explains itself in a way that a rank does not.
They fit poorly where the budget genuinely requires differentiation and the organisation is unwilling to make that constraint explicit. In that situation absolute ratings will be quietly overridden by allocation, and the visible result is a rating system nobody trusts.
The honest version, in that case, is to say plainly that the assessment establishes performance and a separate, budget-constrained decision establishes pay. Employees accept that far more readily than a rating that was adjusted after the fact for reasons nobody will name.
Frequently asked questions
What are absolute ratings?
Performance ratings assessed against a defined standard rather than against colleagues, so every employee can in principle receive the same rating. They contrast with relative methods that rank people against each other.
Are absolute ratings better than forced ranking?
They are fairer in small teams, where relative ranking is arbitrary by construction, and they support collaboration better. Their weakness is that standards drift upward without calibration, whereas ranking holds a distribution by force.
How do you stop absolute ratings inflating?
Calibrate on evidence rather than numbers, describe behaviours at each level rather than using adjectives, and loosen the link between rating and increment enough that the rating is not worth gaming. Check the distribution over three years to see whether drift has set in.
How many levels should a rating scale have?
Four or three is usually enough. Five invites argument about the boundaries and implies a precision the assessment does not have. What matters more than the count is that each level describes observable behaviour.
Can absolute ratings work with a fixed pay budget?
Yes, if the two are separated: assess against the standard, then allocate a fixed pot against those assessments. The failure mode is adjusting ratings to fit the budget, which corrupts the assessment and is quickly noticed.
How Engage handles ratings
Engage holds the rating scale with its described levels alongside the evidence recorded against each employee through the period, so a rating can be traced to what it was based on rather than to what was remembered at the end. Distributions can be read over several cycles and across functions, which is how rating drift becomes visible before it has to be corrected all at once.
See performance management in Engage