Skip to content

Perspective

Why behaviour-only 360° feedback misses part of the picture

Author: Frederic C. RupprechtPublished: 10 min read

In short

A behaviour-only 360° measures what supervisors, peers and direct reports observe. Precise, but incomplete. It cannot show the belief behind a pattern, the unused capability, or the pressure that flips it. Feedback without that why produces resolutions, not change. A 360° leadership assessment across all six layers – a whole-system 360° – asks five questions instead of one: what do I believe, what can I do, what changes under pressure, what do others experience, what effect does that have?

Key takeaways

  • Competency-based 360° feedback describes the outer ring – observable behaviour – reliably, but says nothing about why that behaviour occurs.
  • Three things stay invisible: the belief behind the pattern, the capability that exists but is not used, and the pressure condition under which the pattern flips.
  • The feedback research is sobering: feedback interventions improve performance on average, yet more than a third make it worse (Kluger & DeNisi 1996); ratings from others improve only modestly after multisource feedback (Smither et al. 2005).
  • A whole-system 360° asks five questions instead of one and adds the leader’s own account of Belief, Capability and Trigger to what others observe.
  • Honest limit: self-report stays self-report, and every link between belief, pressure and behaviour is a hypothesis for the conversation – not a diagnosis.

What a classic 360° actually measures

Most 360° instruments are competency models with a questionnaire: they define the leadership behaviour the organisation wants to see and ask supervisors, peers and direct reports how often they observe it. That is not a flaw; it is the definition. Bracken, Rose and Church (2016) define 360° feedback as a process for collecting, quantifying and reporting co-worker observations about a person – raters’ perceptions of specific behaviours. Observation is the raw material. Nothing more was ever promised.

DefinitionBehaviour-only 360° feedback
A multi-rater process in which several groups of raters score a leader’s observable behaviour against a competency model, usually alongside a self-assessment on the same items. It answers one question: what leadership behaviour do others observe? Why it occurs and when it flips are not questions it asks.

A good 360° describes this outer ring reliably, given enough raters. What such a report says is rarely wrong. What it does not say is the problem: it shows a pattern and is silent about its cause. "Delegates too little" is an observation. Whether a missing skill sits behind it, a belief, or a pressure reflex that only appears in crunch weeks – the report has no data on that, and that is where development starts. What is 360° feedback? covers the basics.

Three things a behaviour report cannot show

1. The belief behind the pattern

Leadership behaviour follows beliefs – about the role, the team, what happens when you let go. Kegan and Lahey call these competing commitments: a leader wants to delegate and is at the same time, usually unspoken, committed to never losing control of an outcome. The second commitment wins when it matters. Others cannot rate this layer: beliefs are not observable from outside. You have to ask the leader – separately from behaviour, or they rate their own behaviour a second time. More in Beliefs versus behaviour.

2. The capability that exists but is not used

A behaviour report cannot tell "cannot" from "does not"; both appear as low frequency. A leader who lacks a skill needs training. A leader who has it and does not call on it under certain conditions needs an understanding of those conditions. Bandura’s self-efficacy work offers the way in: whether a leader believes they can show a behaviour is a question of its own. Only comparing that confidence with what others observe reveals whether this is a training, a calibration or a retrieval problem – see The gap between capability and behaviour.

3. The pressure condition under which the pattern flips

Almost every pattern that shows up as a weakness in a 360° is conditional: it appears when the deadline closes in or the client escalates – and disappears when the pressure lifts. A frequency question over the past months averages that condition away: "controls everything in crunch weeks" and "otherwise gives plenty of room" become an unremarkable mean. The clinical, psychodynamic approach to leadership that Kets de Vries stands for takes exactly these reaction patterns seriously – with the assumption that under pressure people fall back on older strategies. Which ones they are cannot be observed – but it can be asked; see Leadership under pressure.

Why this matters for development

Feedback without a why produces resolutions. The leader reads "delegates too little", resolves to delegate more, and does – until the next crunch week. Then the pattern returns, and the next report looks like the last. That is not a character flaw but the predictable outcome of an intervention that names a symptom and leaves the cause untouched.

The research supports this scepticism more clearly than the industry would like. Kluger and DeNisi (1996) showed across 607 effect sizes that feedback interventions improve performance on average (d = .41) but that more than a third made performance worse. Their feedback intervention theory explains when: effectiveness declines as attention shifts from the task towards the self. A behaviour-only report is easily read as a verdict on the person: it offers no mechanism to work on.

Smither, London and Reilly (2005) meta-analysed 24 longitudinal studies of multisource feedback. Ratings from others improved only modestly afterwards: corrected effect sizes of .15 for direct reports, .05 for peers, .15 for supervisors and −.04 for self-ratings, with confidence intervals including zero for every source except supervisors. The authors warn against expecting large, widespread improvement. Change becomes more likely when the feedback signals that change is needed and recipients see it as necessary and feasible, set goals and take action. Two studies they cite found that managers who worked with an executive coach improved more.

Five questions instead of one

If behaviour is the end of a chain, an instrument has to ask along the whole chain. For us that means five questions; only the leader can answer the first and the third.

The five questions of a whole-system 360° against the one question of a behaviour-only 360°.
QuestionLayerWho can answer itBehaviour-only 360°
What do I believe?BeliefThe leader onlyNot asked
What can I do?CapabilityLeader (confidence) + raters (estimate)Not separated from behaviour
What changes under pressure?TriggerThe leader onlyNot asked; averaged away
What do others experience when I lead?BehaviorRaters + self-assessmentThe one question it asks
What effect does that have?OutcomeAll rater groupsPartly, as an overall rating

What a whole-system 360° asks differently

LEADBeyond 360° is a consultant-led 360° leadership assessment built along this chain – Belief → Capability → Trigger → Behavior → Outcome, with Well-being as a sixth layer and energy buffer. Only the leader answers the Belief and Trigger items. The Trigger layer asks in ten items which reaction patterns take hold under pressure, grouped into five defence patterns: Over-Control, Pleasing, Over-Performance, Distancing and Avoidance. Capability is captured in the self-view as confidence and compared with what raters observe; the difference is read as a calibration question, not as a verdict on untapped potential. Details on the methodology page.

Where even a whole-system 360° reaches its limits

  • Self-report stays self-report. Only the leader knows their beliefs and triggers – and they can be wrong, self-protective or aspirational. These layers are not an X-ray but a structured self-description that gains weight only against the view from outside.
  • Links are hypotheses. That a belief blocks a behaviour is an assumption of the model, tested in the conversation – not an empirical finding about this one leader.
  • Self–other agreement has inconsistent findings. Fleenor and colleagues (2010) show that interest in self–other agreement stems from its purported links to self-awareness and leader effectiveness, but that the field has used inconsistent metrics and produced discrepant findings. A gap is a prompt for conversation, not a metric – see Self-image and the view of others.
  • More layers do not cure a bad process. Too few raters, distrusted anonymity or a report without a conversation remain problems – see How a 360° assessment works.
  • Consultant-led means: not for every use case. A whole-system 360° is built for development – not for unaccompanied mass roll-outs, not for selection or promotion decisions. Free-text comments are not analysed by AI in the current version.

Is it worth asking about the why, if all it yields is hypotheses? Yes – because the conversation that follows them is a different one. A report that says "you delegate too little" ends in a resolution. A report that asks "what do you believe happens when you let go – and exactly when do you step back in?" is where development begins.

Frequently asked questions

What are the limitations of behaviour-based 360 feedback?

It measures observed behaviour – reliably, if the process is sound. But it cannot tell "cannot" from "does not", it averages pressure conditions away, and it does not ask about the beliefs behind a pattern. The result is often a report that triggers resolutions but no change.

Does that make behaviour-only 360° feedback worthless?

No. The outer ring – what others experience – is indispensable, and a whole-system 360° contains it in full. The problem is not what a behaviour report shows but what it leaves out.

Why does only the leader answer the questions about beliefs and pressure?

Because beliefs and inner pressure reactions are not observable from outside. Raters can score what they see – not what a leader believes or what throws them off under pressure. Those self-reports gain their weight only when set against the view from outside.

Is the link between beliefs and behaviour scientifically proven?

The model is theory-grounded – it draws on the work of Kegan and Lahey, Bandura and Kets de Vries. The links in the report are hypotheses for the conversation. To our knowledge no study has tested a whole-system 360° against a behaviour-only one, and we claim none. Details on the methodology page.

What does research say about whether 360° feedback works?

Kluger and DeNisi (1996) found an average improvement across 607 effect sizes, but more than a third of feedback interventions made performance worse. Smither, London and Reilly (2005) found only modest average improvements in ratings from others after multisource feedback – and change above all where recipients discussed the feedback, set goals or worked with a coach.

How can I tell whether a 360° provider really captures the why?

Three questions: are there separate belief items that only the leader answers? Is capability captured separately from behaviour? Are reactions under pressure asked about rather than averaged away? And a fourth: are the links presented as hypotheses or as verdicts? Further criteria in the buyer checklist.

Is a whole-system 360° suitable for promotion decisions?

No. It is built for development: the Belief and Trigger layers are self-reports, the links are hypotheses, and raters tend to answer differently once their answers help decide careers. LEADBeyond 360° is not built for selection or promotion decisions – that calls for an assessment centre or a selection-oriented potential analysis.

Sources

  1. Bracken, D. W., Rose, D. S., & Church, A. H. (2016). The evolution and devolution of 360° feedback. Industrial and Organizational Psychology, 9(4), 761–794., Cambridge University Press (2016)Peer-reviewed; only the definition of 360° feedback given there is cited.
  2. Kluger, A. N., & DeNisi, A. (1996). The effects of feedback interventions on performance: A historical review, a meta-analysis, and a preliminary feedback intervention theory. Psychological Bulletin, 119(2), 254–284., American Psychological Association (1996)Peer-reviewed meta-analysis (607 effect sizes); source for d = .41, "more than a third" and feedback intervention theory.
  3. Smither, J. W., London, M., & Reilly, R. R. (2005). Does performance improve following multisource feedback? A theoretical model, meta-analysis, and review of empirical findings. Personnel Psychology, 58(1), 33–66., Wiley (2005)Peer-reviewed meta-analysis of 24 longitudinal studies; source for the corrected effect sizes .15 / .05 / .15 / −.04 and the conditions under which change becomes more likely.
  4. Fleenor, J. W., Smither, J. W., Atwater, L. E., Braddy, P. W., & Sturm, R. E. (2010). Self–other rating agreement in leadership: A review. The Leadership Quarterly, 21(6), 1005–1034., Elsevier (2010)Peer-reviewed literature review; does not support any single headline statistic.
  5. Kegan, R., & Lahey, L. L. (2009). Immunity to Change: How to Overcome It and Unlock the Potential in Yourself and Your Organization. Harvard Business Press., Harvard Business Press (2009)Cited only as the source of the competing-commitments concept; no statistics.
  6. Bandura, A. (1997). Self-Efficacy: The Exercise of Control. W. H. Freeman., W. H. Freeman (1997)Cited only as the canonical reference for self-efficacy theory; no statistics.
  7. Kets de Vries, M. F. R. (2006). The Leader on the Couch: A Clinical Approach to Changing People and Organizations. Jossey-Bass (John Wiley & Sons)., Jossey-Bass (John Wiley & Sons) (2006)Cited only as an example of the clinical, psychodynamic approach to leadership; no statistics.

Author

Frederic C. Rupprecht

Co-founder & Managing Director, LEADBeyond GmbH

Frederic C. Rupprecht is co-founder and Managing Director of LEADBeyond GmbH and responsible for the methodology and product behind LEADBeyond 360°. A former Associate Partner at McKinsey & Company, he works with leaders in consulting and professional-services firms on diagnostics and development.

Related reading