HR tools built by a senior HR executive · United States

Published standard  ·  Assessments  ·  Contact

Manager capability and performance calibration

The manager is the single biggest factor in whether a team performs, and the hardest part of the job is judging people fairly. Most managers were never taught how, and the ratings they produce drift apart from one another. This note covers what the research says about manager readiness, the specific errors that distort performance reviews, and the calibration process that keeps ratings honest.

How this note is governed

Research synthesis. Not jurisdictional.

Applies to Employers running manager-led performance ratings across more than one team. Research synthesis and practice guidance, not a legal rule.

Short answer

70%. Managers account for about 70% of the variance in team engagement, yet only 44% have had formal management training. Calibration aligns ratings across managers before they go out, so a rating means the same thing everywhere.

Published Last verified

Refreshed against Gallup's State of the Global Workplace 2026. The 70% manager variance and 44% training figures were unchanged.

44%
of managers report ever receiving formal management training, so most are leading and rating people without it.
70%
of the variance in team engagement is attributable to the manager, the most consistent finding in Gallup’s research.

Most managers were never taught to manage

The promotion path in most companies rewards strong individual work with a management title. It then leaves the new manager to figure out the people part alone. The data on how common this is is stark. Gallup’s State of the Global Workplace report finds that only 44% of managers have ever received formal management training. Managers account for roughly 70% of the variance in team engagement. The same research found that even basic training cuts active disengagement among managers by about half. The gap is not just common. It is the lever with the most weight behind it.

The cost of leaving it unaddressed has been rising. Gallup recorded manager engagement falling from 31% in 2022 to 22% in 2025. That was the steepest drop of any group in the workforce, with the largest single-year fall between 2024 and 2025. Managers have lost what Gallup calls the engagement premium and are now about as engaged as the people they lead. When the person who sets the tone for a team is stretched and unsupported, the team feels it.

The starting point for someone stepping into management for the first time, with the core habits and conversations that the promotion did not come with.

The New Manager Kit, $69

Judging people is where capability gets tested

A manager does many things, but the one that carries the most weight, and the most risk, is evaluating performance. A rating drives pay, promotion, and who stays. It is also the task where an untrained manager is most exposed, because human judgment of other people runs on predictable shortcuts. These are not character flaws. They are systematic errors that every rater makes, and the people making them are usually unaware they are doing so. Naming them is the first step to controlling them.

    Calibration is the established correction

    Calibration is a structured session where managers who supervise comparable groups compare their proposed ratings with each other. HR or a neutral facilitator guides it, and it happens before any review reaches an employee. The purpose is consistency: to make a 4 from one manager carry the same weight as a 4 from another. It is, in plain terms, a review of the reviews. SHRM describes the core sequence as managers posting names and proposed ratings for all to see, then discussing each. They adjust to assure accuracy and consistency before final appraisals are prepared.

    The reason it works is that the errors above are hard to catch from inside your own head but easy to spot from outside. When one manager’s team is rated uniformly higher than a comparable team, the discrepancy is visible in the room and gets examined. Performance-management specialist Dick Grote has noted that calibration also makes it easier for managers to deliver honest but negative appraisals. The standard is shared rather than personal. Calibration also exposes strong performers to a wider set of senior leaders.

    What separates a useful session from a political one

    Calibration done badly is worse than none at all. The common failure, documented by SHRM, is that sessions defer to the loudest or highest-ranking person in the room. They end up calibrating one set of biased ratings against another. A few conditions keep a session honest.

      For a manager who has never run one, there is a gap between knowing calibration matters and being able to run one well. That is exactly the capability gap this note opened with. The skills are learnable, which is the encouraging part: structured preparation, evidence-based discussion, and a clear rubric turn a vague exercise into a defensible one.

      Six red flags to check before you fire someone

      Free, and written to the same standard

      A five minute screen to run before you act, sent to your inbox as a print-ready PDF. Every figure in it traces to a reference note like this one.

      Where these figures come from

      5 citations checked, newest check 24 June 2026
      1. Gallup, State of the Global Workplace 2026. The source for managers driving roughly 70% of the variance in team engagement, and the 44% who have received management training. It also carries the halving of active disengagement among trained managers, and manager engagement falling from 31% in 2022 to 22% in 2025. gallup.com gallup.com Checked 24 June 2026
      2. SHRM, Improving Performance Evaluations Using Calibration. The source for the calibration sequence, where managers post and discuss proposed ratings then adjust for consistency. It also carries Dick Grote’s points on honest appraisals, skilled facilitation, and bringing data rather than views. shrm.org shrm.org Checked 24 June 2026
      3. SHRM Labs, Fixing Performance Reviews. The source for the failure mode where calibration defers to the loudest or highest-ranking manager and ends up calibrating biased ratings against other biased ratings. shrm.org shrm.org Checked 24 June 2026
      4. SHRM Certified Professional, rater errors in performance measurement. The source for the taxonomy of rater errors: halo and horns, leniency and severity, central tendency, recency, and similar-to-me bias. SHRM-CP reference trustedinstitute.com Checked 24 June 2026
      5. Dartmouth College HR, Common Rater Errors. A university HR reference confirming the standard rater-error definitions and the point that observers are usually unaware they are making them. dartmouth.edu dartmouth.edu Checked 24 June 2026

      Common questions

      What percentage of managers receive training?

      Gallup’s State of the Global Workplace research finds that only 44% of managers report ever receiving formal management training. The same research finds that managers drive about 70% of the variance in team engagement. Most companies are leaving their single biggest engagement lever undeveloped.

      What is a performance calibration meeting?

      It is a structured session where managers who supervise comparable groups compare their proposed performance ratings. This happens before any review reaches an employee, guided by HR or a neutral facilitator. The purpose is to make a given rating mean the same thing across teams. A 4 from one manager carries the same weight as a 4 from another.

      What are the most common rater errors?

      The well-documented ones are recency bias, where recent events outweigh the full period. The halo or horns effect is one trait coloring every score. Leniency or severity shifts the whole scale up or down by manager, and central tendency clusters everyone in the middle. Similar-to-me bias is familiarity reading as competence. Most raters make them without realizing it.

      Does calibration mean forcing a bell curve?

      No. The goal of calibration is consistency, not a forced distribution. A good session aligns what each rating means across managers so the standard is shared. Slotting people into a predetermined curve is a different practice. Calibrating biased ratings against each other or against a curve is the common way the process goes wrong.

      Put it to work

      • The starting point for someone stepping into management for the first time, with the core habits and conversations that the promotion did not come with.

        $69
      • The fuller program for building a new manager up across their first year. It is structured so capability is developed on purpose rather than by trial and error.

        $149
      • The working tools an experienced manager reaches for: the templates and structures behind one-to-ones, feedback, and the recurring people decisions that fill the week.

        $79
      • Word-for-word openings for the conversations managers avoid, so the hardest moments do not depend on finding the right words in real time.

        $69
      • The same conversation tools shaped for a dental or medical practice, where the manager is often clinical-side and leading people is the newer skill.

        $79
      • Built for the floor, where a supervisor leads a shift while still doing the work. Feedback and accountability happen in real time rather than in a quiet office.

        $79

      This note is general information about employment practice rather than legal advice for your situation. Check the review date and the jurisdictions above, follow the source link, and confirm the rule before you act on it.

      From evidence to action

      Use the note to make the next decision.

      A reference note establishes scope and authority. The useful next move is to test the facts, install the operating method, or review the live situation.

      01 · Test

      Run a related calculator

      Put your own facts into the method instead of relying on a general example.

      Open the analysis →
      02 · Implement

      The New Manager Kit

      Move from the rule or method into an editable operating document.

      See the operating path →
      03 · Apply

      Use the matched tool

      The kit or calculator built for this issue carries the evidence into a file you can run.

      Browse the tools →