InsightsWhat multi-rater feedback measures
🔁 Multi-rater method

360 feedback does not measure performance. It measures vantage point.

Most leaders go straight to their overall score. There is a more useful way to read a 360 feedback report, and it starts with understanding that 360-degree feedback does not measure performance — it measures where each person was standing when they watched you work.

Multi-rater feedback: a leader observed from several different vantage points by colleagues
Every rater group stands somewhere different — and each sees a different slice of the same person.

Your raters aren’t disagreeing with each other. They’re describing different behavior in the same person.

There’s a more useful way to think about raters than as judges. Think of them as witnesses — each standing in a different spot, each with a different view of the same person.

Someone who reports to a leader watches how that person allocates attention on an ordinary Tuesday. They notice what happens in the hour after a decision goes badly, or how the leader responds when contradicted in a meeting. A peer sees something else entirely: whether commitments made across team boundaries actually get honored, or how that person behaves in a room where they hold no authority over the outcome. A supervisor is watching something different still — how judgment holds under scrutiny, and how the work gets represented upward.

When multi-rater feedback surfaces differences of opinion across these groups, the instinct is often to look for who’s right. A more productive instinct is to ask: what situation is each person describing? The same leader can be patient in front of their manager and short with a peer on a deadline. That’s not a contradiction — it’s information.

Because each rater group observes from a different proximity, 360Score.me combines individual ratings into a single weighted peer score that reflects how well-positioned each group is to observe the subject day to day. By default, that weighting is based on org chart proximity — the people who work closest to someone carry more weight in the final score — and administrators can adjust the model before a cycle runs. The intent is not to declare one perspective more valid than another. It’s to ensure no single vantage point speaks louder than its position warrants.

The gap between how you see yourself and how others see you — and why it matters more than either number alone

There is one data point in a 360Score.me report that almost no other management tool produces: a self-evaluation score. Most feedback tools tell you how someone is perceived by their peers. A leadership 360 assessment in 360Score.me also tells you how accurately that person understands their own perception — because it collects a self-rating alongside every other rater group’s. Decades of multi-rater research point consistently to the same finding: self-ratings and observer ratings rarely align closely, and the direction of that gap tends to predict what happens next more reliably than either number does on its own.

Three patterns show up in practice. When a leader’s self-rating sits above their peer scores, the gap often points toward a blind spot — something others are observing that the leader hasn’t registered. When the self-rating lands close to what others see, the self-view is accurate, and developmental conversations can focus on the substance of the feedback rather than on reorienting the person first. When the self-rating sits below the peer scores, the capability is already there; what’s often missing is the confidence to claim it.

Over
Self above peers
Usually the blind spot worth working
Align
Self near peers
Accurate self-view — coach on content
Under
Self below peers
Capability present; confidence is the constraint

Of the three, overestimation is the one worth spending the most coaching attention on — and it’s also the most resistant to shifting, for a straightforward reason: it tends to protect itself. A leader who rates themselves highly on giving clear direction is unlikely to take a report saying otherwise at face value. Nothing in their own experience contradicts that self-image, and the very behavior drawing the lower ratings — being difficult to interrupt, perhaps, or hard to read — is the same behavior that would normally invite correction from outside. Underestimation, by contrast, is far more tractable. The capability exists; the work is showing someone the evidence until they believe it.

A leader facing their own reflection, illustrating leadership blind spots in self-assessment
Self-view and observed behavior rarely line up exactly. The offset is the useful part.
The offset between self-view and observed behavior is often the most useful number on the page.

Read the scores and the comments together — neither tells the full story alone

A 360Score.me report comes in two halves: a set of scores and a set of written comments. In practice, one tends to get more attention than it deserves.

The weighted peer score is very good at telling you how significant something is and roughly where it sits relative to everything else in the report. Because it pools ratings from multiple observers, any single rater’s unusually generous or unusually critical instincts get absorbed into a larger signal. What scores can’t convey is the texture of the behavior — the specific situations in which it appears, the impact it leaves behind, the moments that stuck.

That’s what comments are for. Each one is a single colleague describing a single experience. Workplaces are complex, and any given person will have been experienced differently across different circumstances — under pressure and under calm, in a formal meeting and an informal one, during a strong quarter and a difficult one. Comments reflect that range, and they’re worth holding lightly as the individual data points they are.

The most productive way to hold the two halves together: when scores and comments point in the same direction, you have both the magnitude of a pattern and its content, and you can act on it with confidence. When they diverge — a competency scoring comfortably while the surrounding comments carry a note of hesitation — that divergence is itself the finding. It usually signals that a behavior holds in most situations and not in one that matters. The same logic extends to the self-other gap. The gap tells you a misperception exists. The comments tell you what everyone else was watching while the subject was looking somewhere else.

A short field guide to reading your 360 peer review results

Reading your own 360 feedback well is a skill, and it gets easier once you know what patterns to expect. What follows is a practical orientation to some of the most common things you’ll notice in a multi-rater report — not as flaws in what your colleagues said, but as natural features of how people observe and describe each other at work. Understanding them helps you find what’s genuinely useful in your results, rather than reacting to what’s most visible.

When your scores feel warmer than you expected.

If your ratings land on the higher end, that’s worth receiving with some nuance. It’s entirely natural for colleagues to score the people they work alongside generously — candor carries social cost in close working relationships, and warmth costs very little. This doesn’t make the scores inaccurate; it makes them human. The more useful thing to examine isn’t the height of any individual score but the shape of your profile overall. Because warmth tends to lift scores fairly evenly, your relative strengths and growth areas remain visible in relation to each other — you’ll still see where you’re strongest and where there’s distance from that.

For raters:

The most useful thing you can offer isn’t a uniformly high set of scores — it’s an honest read that reflects your actual experience working with this person. Your identity is never revealed in 360Score.me results, so you can respond candidly without concern about how your colleague will receive it. Specific, behavioral responses are almost always more useful to the person receiving them than generalized praise.

When your profile looks unusually even.

Occasionally a report comes back with very little variation — every competency in roughly the same range, no contrast anywhere. If you see this in your results, it’s less likely to mean that you perform identically across every dimension, and more likely to reflect how one or more raters worked through the questions. When someone forms an overall impression of a person and answers every question from that impression, the profile flattens. This is worth knowing mostly so you don’t over-interpret an even profile as precision. Look for any variation that does exist — even a modest contrast can be meaningful — and let the written comments add texture that the scores alone can’t provide.

For raters:

Try to answer each question from a specific memory rather than from your general sense of the person. Questions about observable behavior — how someone gives direction, how they handle disagreement — are asking you to check the question against something you actually watched. Your identity is never shared, so you can respond from genuine experience without hesitation.

When a comment feels like it’s about one specific moment.

A memorable project, a high-pressure deadline, a visible stumble — any of these can show up in feedback carrying more weight than the broader context of someone’s work. If you read a comment and recognize the situation it’s describing, it’s worth asking yourself whether it reflects a pattern or a moment. Both are worth understanding, but they point to different things. A single moment described in strong terms is worth noting; a theme that surfaces across multiple comments is worth acting on.

For raters:

Try to think across the full review period rather than anchoring on something recent. Your anonymity is fully protected, so there’s no need to soften or obscure what you observed — but grounding your response in the full span of your experience together will make your feedback more useful to the person receiving it.

When you wonder who your raters actually were.

360Score.me derives your rater group from your position in the org structure — direct reports, peers at your level, your supervisor — rather than from a list you selected yourself. That structural starting point means your results reflect a representative cross-section of the people who experience your work most directly. Administrators can adjust the group before a cycle runs, but individual rater identities are never revealed. That’s by design: the anonymity that protects your raters’ candor is also what makes anonymous peer feedback worth trusting.

When not enough responses come in.

Sometimes a cycle closes with fewer responses than expected. 360Score.me requires a minimum of five rater responses before peer results are released — three for skip-level results. If your results aren’t available yet, it’s because that threshold hasn’t been met. This isn’t a procedural detail; it serves two purposes at once. Statistically, a very small group stops producing a reliable average — one strong opinion can shift a competency score by a full rating band. And practically, a very small group stops being truly anonymous: it becomes possible to connect specific comments to specific people simply by knowing who was asked. The floor exists to protect the integrity of your results and the people who contributed to them.

5 raters minimum for peer results · 3 for skip-level

Where 360 feedback fits alongside a performance evaluation

A 360 review and a performance evaluation are answering different questions, and the distinction matters enough to hold clearly.

A performance evaluation asks what was delivered — against a stated standard, with a named evaluator and usually a decision attached. A 360 review asks how someone works, and how that lands on the people around them. Neither can substitute for the other. A leader can hit every target on the sheet while steadily depleting the team that produced those results. In a standard annual cycle, a 360Score.me review is often the only instrument in the process that will surface that dynamic before it becomes costly.

That distinction also supports one principle worth holding firmly, even when there’s pressure to blur it. Connecting 360 results directly to compensation decisions is tempting — the data is already collected, the logic feels efficient. But raters are perceptive. When they understand that a candid answer might affect a colleague’s pay, the safer answer becomes the generous one. Scores compress toward the top, the meaningful spread between competencies collapses, and within a cycle or two the survey is still running while producing very little that anyone can act on. Keep 360Score.me reviews developmental, and they keep telling you the truth.

See a multi-rater report read the way it should be read.

A walkthrough of our 360 feedback software with anonymized data — rater-source breakdown, self-other gaps, and competency profiles.

See It Live Next: reading open text →