What a classic 360° actually measures
Most 360° instruments are competency models with a questionnaire: they define the leadership behaviour the organisation wants to see and ask supervisors, peers and direct reports how often they observe it. That is not a flaw; it is the definition. Bracken, Rose and Church (2016) define 360° feedback as a process for collecting, quantifying and reporting co-worker observations about a person – raters’ perceptions of specific behaviours. Observation is the raw material. Nothing more was ever promised.
- DefinitionBehaviour-only 360° feedback
- A multi-rater process in which several groups of raters score a leader’s observable behaviour against a competency model, usually alongside a self-assessment on the same items. It answers one question: what leadership behaviour do others observe? Why it occurs and when it flips are not questions it asks.
A good 360° describes this outer ring reliably, given enough raters. What such a report says is rarely wrong. What it does not say is the problem: it shows a pattern and is silent about its cause. "Delegates too little" is an observation. Whether a missing skill sits behind it, a belief, or a pressure reflex that only appears in crunch weeks – the report has no data on that, and that is where development starts. What is 360° feedback? covers the basics.
Three things a behaviour report cannot show
1. The belief behind the pattern
Leadership behaviour follows beliefs – about the role, the team, what happens when you let go. Kegan and Lahey call these competing commitments: a leader wants to delegate and is at the same time, usually unspoken, committed to never losing control of an outcome. The second commitment wins when it matters. Others cannot rate this layer: beliefs are not observable from outside. You have to ask the leader – separately from behaviour, or they rate their own behaviour a second time. More in Beliefs versus behaviour.
2. The capability that exists but is not used
A behaviour report cannot tell "cannot" from "does not"; both appear as low frequency. A leader who lacks a skill needs training. A leader who has it and does not call on it under certain conditions needs an understanding of those conditions. Bandura’s self-efficacy work offers the way in: whether a leader believes they can show a behaviour is a question of its own. Only comparing that confidence with what others observe reveals whether this is a training, a calibration or a retrieval problem – see The gap between capability and behaviour.
3. The pressure condition under which the pattern flips
Almost every pattern that shows up as a weakness in a 360° is conditional: it appears when the deadline closes in or the client escalates – and disappears when the pressure lifts. A frequency question over the past months averages that condition away: "controls everything in crunch weeks" and "otherwise gives plenty of room" become an unremarkable mean. The clinical, psychodynamic approach to leadership that Kets de Vries stands for takes exactly these reaction patterns seriously – with the assumption that under pressure people fall back on older strategies. Which ones they are cannot be observed – but it can be asked; see Leadership under pressure.
Why this matters for development
Feedback without a why produces resolutions. The leader reads "delegates too little", resolves to delegate more, and does – until the next crunch week. Then the pattern returns, and the next report looks like the last. That is not a character flaw but the predictable outcome of an intervention that names a symptom and leaves the cause untouched.
The research supports this scepticism more clearly than the industry would like. Kluger and DeNisi (1996) showed across 607 effect sizes that feedback interventions improve performance on average (d = .41) but that more than a third made performance worse. Their feedback intervention theory explains when: effectiveness declines as attention shifts from the task towards the self. A behaviour-only report is easily read as a verdict on the person: it offers no mechanism to work on.
Smither, London and Reilly (2005) meta-analysed 24 longitudinal studies of multisource feedback. Ratings from others improved only modestly afterwards: corrected effect sizes of .15 for direct reports, .05 for peers, .15 for supervisors and −.04 for self-ratings, with confidence intervals including zero for every source except supervisors. The authors warn against expecting large, widespread improvement. Change becomes more likely when the feedback signals that change is needed and recipients see it as necessary and feasible, set goals and take action. Two studies they cite found that managers who worked with an executive coach improved more.
Five questions instead of one
If behaviour is the end of a chain, an instrument has to ask along the whole chain. For us that means five questions; only the leader can answer the first and the third.
| Question | Layer | Who can answer it | Behaviour-only 360° |
|---|---|---|---|
| What do I believe? | Belief | The leader only | Not asked |
| What can I do? | Capability | Leader (confidence) + raters (estimate) | Not separated from behaviour |
| What changes under pressure? | Trigger | The leader only | Not asked; averaged away |
| What do others experience when I lead? | Behavior | Raters + self-assessment | The one question it asks |
| What effect does that have? | Outcome | All rater groups | Partly, as an overall rating |
What a whole-system 360° asks differently
LEADBeyond 360° is a consultant-led 360° leadership assessment built along this chain – Belief → Capability → Trigger → Behavior → Outcome, with Well-being as a sixth layer and energy buffer. Only the leader answers the Belief and Trigger items. The Trigger layer asks in ten items which reaction patterns take hold under pressure, grouped into five defence patterns: Over-Control, Pleasing, Over-Performance, Distancing and Avoidance. Capability is captured in the self-view as confidence and compared with what raters observe; the difference is read as a calibration question, not as a verdict on untapped potential. Details on the methodology page.
Where even a whole-system 360° reaches its limits
- Self-report stays self-report. Only the leader knows their beliefs and triggers – and they can be wrong, self-protective or aspirational. These layers are not an X-ray but a structured self-description that gains weight only against the view from outside.
- Links are hypotheses. That a belief blocks a behaviour is an assumption of the model, tested in the conversation – not an empirical finding about this one leader.
- Self–other agreement has inconsistent findings. Fleenor and colleagues (2010) show that interest in self–other agreement stems from its purported links to self-awareness and leader effectiveness, but that the field has used inconsistent metrics and produced discrepant findings. A gap is a prompt for conversation, not a metric – see Self-image and the view of others.
- More layers do not cure a bad process. Too few raters, distrusted anonymity or a report without a conversation remain problems – see How a 360° assessment works.
- Consultant-led means: not for every use case. A whole-system 360° is built for development – not for unaccompanied mass roll-outs, not for selection or promotion decisions. Free-text comments are not analysed by AI in the current version.
Is it worth asking about the why, if all it yields is hypotheses? Yes – because the conversation that follows them is a different one. A report that says "you delegate too little" ends in a resolution. A report that asks "what do you believe happens when you let go – and exactly when do you step back in?" is where development begins.
Frequently asked questions
What are the limitations of behaviour-based 360 feedback?
It measures observed behaviour – reliably, if the process is sound. But it cannot tell "cannot" from "does not", it averages pressure conditions away, and it does not ask about the beliefs behind a pattern. The result is often a report that triggers resolutions but no change.
Does that make behaviour-only 360° feedback worthless?
No. The outer ring – what others experience – is indispensable, and a whole-system 360° contains it in full. The problem is not what a behaviour report shows but what it leaves out.
Why does only the leader answer the questions about beliefs and pressure?
Because beliefs and inner pressure reactions are not observable from outside. Raters can score what they see – not what a leader believes or what throws them off under pressure. Those self-reports gain their weight only when set against the view from outside.
Is the link between beliefs and behaviour scientifically proven?
The model is theory-grounded – it draws on the work of Kegan and Lahey, Bandura and Kets de Vries. The links in the report are hypotheses for the conversation. To our knowledge no study has tested a whole-system 360° against a behaviour-only one, and we claim none. Details on the methodology page.
What does research say about whether 360° feedback works?
Kluger and DeNisi (1996) found an average improvement across 607 effect sizes, but more than a third of feedback interventions made performance worse. Smither, London and Reilly (2005) found only modest average improvements in ratings from others after multisource feedback – and change above all where recipients discussed the feedback, set goals or worked with a coach.
How can I tell whether a 360° provider really captures the why?
Three questions: are there separate belief items that only the leader answers? Is capability captured separately from behaviour? Are reactions under pressure asked about rather than averaged away? And a fourth: are the links presented as hypotheses or as verdicts? Further criteria in the buyer checklist.
Is a whole-system 360° suitable for promotion decisions?
No. It is built for development: the Belief and Trigger layers are self-reports, the links are hypotheses, and raters tend to answer differently once their answers help decide careers. LEADBeyond 360° is not built for selection or promotion decisions – that calls for an assessment centre or a selection-oriented potential analysis.
Sources
- Bracken, D. W., Rose, D. S., & Church, A. H. (2016). The evolution and devolution of 360° feedback. Industrial and Organizational Psychology, 9(4), 761–794., Cambridge University Press (2016) — Peer-reviewed; only the definition of 360° feedback given there is cited.
- Kluger, A. N., & DeNisi, A. (1996). The effects of feedback interventions on performance: A historical review, a meta-analysis, and a preliminary feedback intervention theory. Psychological Bulletin, 119(2), 254–284., American Psychological Association (1996) — Peer-reviewed meta-analysis (607 effect sizes); source for d = .41, "more than a third" and feedback intervention theory.
- Smither, J. W., London, M., & Reilly, R. R. (2005). Does performance improve following multisource feedback? A theoretical model, meta-analysis, and review of empirical findings. Personnel Psychology, 58(1), 33–66., Wiley (2005) — Peer-reviewed meta-analysis of 24 longitudinal studies; source for the corrected effect sizes .15 / .05 / .15 / −.04 and the conditions under which change becomes more likely.
- Fleenor, J. W., Smither, J. W., Atwater, L. E., Braddy, P. W., & Sturm, R. E. (2010). Self–other rating agreement in leadership: A review. The Leadership Quarterly, 21(6), 1005–1034., Elsevier (2010) — Peer-reviewed literature review; does not support any single headline statistic.
- Kegan, R., & Lahey, L. L. (2009). Immunity to Change: How to Overcome It and Unlock the Potential in Yourself and Your Organization. Harvard Business Press., Harvard Business Press (2009) — Cited only as the source of the competing-commitments concept; no statistics.
- Bandura, A. (1997). Self-Efficacy: The Exercise of Control. W. H. Freeman., W. H. Freeman (1997) — Cited only as the canonical reference for self-efficacy theory; no statistics.
- Kets de Vries, M. F. R. (2006). The Leader on the Couch: A Clinical Approach to Changing People and Organizations. Jossey-Bass (John Wiley & Sons)., Jossey-Bass (John Wiley & Sons) (2006) — Cited only as an example of the clinical, psychodynamic approach to leadership; no statistics.
Related reading
Perspectives
Leadership beliefs vs leadership behaviour: why development starts with what leaders believe
Why leaders rarely change behaviour after feedback: beliefs as the operating assumptions behind patterns, former success strategies – and how a debrief works with them.
Read more →Perspectives
The capability–behaviour gap: when leaders can, but don't
Why a gap between what a leader can do and what others observe is worth more diagnostically than a low score: four causes, cross-source measurement and how to debrief it.
Read more →Perspectives
Leadership under pressure: why leadership behaviour changes under stress – and why that is normal
Under pressure, leaders fall back on old patterns – a mechanism, not a character flaw. The five defense patterns, their typical triggers, and what leaders and firms can do.
Read more →