Templates

The 1–5 interview rating scale, with anchors written for every point

On this page
  1. What each of the five points means, in general
  2. The scale applied to five competencies
  3. Writing an anchor for a competency not listed above
  4. Decision rules that go with a 1–5 scale
  5. Converting scores from a different scale
  6. Keeping a 3 from becoming the default
  7. Putting the scale on a scorecard
  8. Common mistakes with a 1–5 scale
  9. When five points are the wrong choice
  10. Questions people ask

A 1–5 interview rating scale only works if every one of its five points means the same thing to every interviewer. Most teams that adopt "1 to 5" stop there and never write down what a 2 looks like versus a 3, so the scale becomes five ways of saying "I liked them" or "I didn't." Below is the full scale: five points, each with a written anchor, applied to five common competencies, plus the decision rules that keep a 3 from becoming the default answer.

This page assumes you have already decided to use five points. If you are still choosing between four, five, seven or ten, see interview rating scale for that decision; this page goes one level deeper into the scale most teams land on.

What each of the five points means, in general

Before anchoring individual competencies, interviewers need a shared definition of the five levels themselves. Write this once and put it at the top of every scorecard.

ScoreLabelWhat it means
1No evidence, or evidence of the oppositeThe candidate could not demonstrate the behavior when asked directly, or described doing the opposite of what the role needs.
2Weak evidenceSome relevant experience, but the example was thin, the candidate needed heavy prompting, or the outcome was poor.
3Solid evidenceA clear, complete example that meets the bar for the role at the level being hired. This is a real "yes," not a placeholder for "unsure."
4Strong evidenceA clear example plus something extra: handled unusual difficulty, reflected on what they would do differently, or the scale of the example exceeded the role's normal demands.
5Exceptional evidenceStrong evidence on more than one occasion, or evidence so specific and well-reasoned that it removes doubt entirely. Rare by design.

Two rules make this table usable instead of decorative. First, 3 is a genuine pass, not a midpoint interviewers pick when they are unsure — unsure evidence is a 2. Second, 5 should be rare; if a third of your scores are 5s, the anchor for 5 has drifted down to what 4 used to mean.

The scale applied to five competencies

A shared definition of the five levels is necessary but not sufficient. Interviewers also need the levels translated into what they would actually hear a candidate say, for each competency on the scorecard. Copy the rows below for the competencies your loop already uses, and write new ones the same way for anything not covered here.

Ownership

12345
Describes a failure as someone else's decision, with no mention of what they controlled. Names their role but the example stops at "I raised it" with no follow-through described. Names a problem they were responsible for, what they did about it, and the outcome. Owns a failure that was arguably not their fault, and explains what they changed anyway. Two or more examples of owning outcomes beyond their formal responsibility, each with a concrete result.

Communication

12345
Answer is disorganized or contradicts itself when the interviewer asks a follow-up. Gets to the point eventually, but needs several redirects from the interviewer. States the point first, then supports it, without needing redirection. Adjusts the explanation on the fly when the interviewer signals confusion, without being asked to. Explains a genuinely complex situation so a non-expert in the room follows it on the first pass.

Problem solving

12345
Cannot describe how they reached a decision, or the reasoning does not connect to the outcome. Reasoning is present but shallow: one option considered, no mention of tradeoffs. Names at least two options, explains why one was chosen, and the choice holds up. Revisits an earlier assumption mid-problem when new information appears, without being prompted. Solves a problem outside their stated experience by reasoning from first principles, and the interviewer can follow every step.

Collaboration

12345
Describes a disagreement in terms of who was right, with no acknowledgment of the other side's point. Mentions working with others but the example is really a solo effort with people nearby. Describes a specific disagreement, what they did to resolve it, and the working relationship afterward. Changed their own position after hearing a colleague's argument, and says what changed their mind. Repaired a relationship that had broken down, with a concrete account of what they did differently.

Domain or role-specific skill

12345
Cannot perform the core task when asked directly, or the attempt has fundamental errors. Completes a simplified version of the task, or needs substantial hints to finish. Completes the task at the level the role requires, independently. Completes the task and explains an edge case or failure mode unprompted. Completes the task, explains tradeoffs, and teaches the interviewer something about the domain.

Replace "domain or role-specific skill" with the actual skill your loop tests — a SQL query, a sales role-play, a med-surg scenario. The five-level shape stays the same: absent, thin, solid, solid-plus-something, and rare.

Writing an anchor for a competency not listed above

Use this order and each anchor writes itself in a few minutes:

  1. Write the 3 first. Describe the example that would make you comfortable moving the candidate forward on this competency alone. This is your bar.
  2. Write the 1. Describe what a candidate says when they clearly do not have this. Usually it is the opposite behavior, not just "less of the 3."
  3. Write the 2. Describe the version of the 3's example that is thinner: fewer details, weaker outcome, or the candidate needed the interviewer to supply half the reasoning.
  4. Write the 4. Take the 3's example and add one specific thing: harder circumstances, self-correction, or a second example that confirms the first.
  5. Write the 5 last, and try to avoid needing it. If you cannot describe what would be rarer than the 4, your 4 was actually a 5, and you should renumber rather than force a 5 into existence.

Decision rules that go with a 1–5 scale

The scale only settles disputes if everyone agreed on the rules before the debrief, not during it.

  • No half points. A 3.5 means the interviewer could not decide between the 3 anchor and the 4 anchor. Send them back to which one the evidence actually matches.
  • Set the pass line per competency, not as a single overall number. A common rule: no score below 2 on any must-have competency, and an average of 3 or better across the rest. Averaging alone lets a 5 on a nice-to-have cancel out a 1 on something the role cannot do without.
  • A 1 on a must-have ends the loop. Later interviewers should be told, so they do not spend a session re-testing something already settled.
  • Ties are broken by re-reading evidence, not by re-scoring from memory. If two candidates land on the same total, compare the actual notes for the competency that matters most for this role, not the numbers.
  • "Not assessed" is not a 3. If a competency was not actually covered in the interview, mark it as not assessed and adjust the loop next time rather than filling in a guess.

Converting scores from a different scale

If you are migrating from a 1–4 or 1–10 scale, do not attempt a mathematical conversion — the meanings do not map cleanly and a formula hides that. Instead, re-score a handful of past candidates' written notes against the new five anchors and see where they land. The table below is a rough starting point for how the ranges tend to relate, useful for talking to a team mid-migration, not for converting historical data automatically.

1–4 scaleRoughly maps to on 1–51–10 scaleRoughly maps to on 1–5
111–21
223–42
335–63
44 or 57–84
——9–105

Notice the 1–4 scale has nowhere to put a genuine 5. That gap is exactly why some teams move to five points in the first place: a 1–4 scale forces "strong" and "exceptional" into the same box.

Keeping a 3 from becoming the default

The most common failure of any 1–5 scale is not a bad anchor, it is interviewers defaulting to the middle because it feels safe. Three things reduce it:

  • Require a one-sentence justification with every score. "3, because they described debugging a production issue independently and the fix held" is checkable. "3" alone is not.
  • Review the score distribution monthly, not just individual scores. If one interviewer's scores cluster at 3 across every candidate and every competency, that is a training signal, not a reflection of the candidate pool.
  • Run a short calibration exercise where several interviewers score the same recorded answer independently and compare. See interview calibration for the full format.

Putting the scale on a scorecard

The scale above is the rating logic; it still needs a place to live. Pair it with an evidence column so the score and the reason sit next to each other, the way interview scorecard template lays out. A score with no evidence column next to it is the single most common reason calibration sessions run long: nobody can remember, three weeks later, what a given 3 was actually based on.

Competency: [name]
Score (1-5): [ ]
Evidence: [what the candidate said or did, specific enough to check against the anchor]
Anchor matched: [copy the sentence from the anchor table above that matches]

Common mistakes with a 1–5 scale

MistakeWhat it causesFix
Anchors only written for the 1, 3 and 5Interviewers guess at 2 and 4, which drift over timeWrite all five, even if 2 and 4 look similar to their neighbors at first
Same anchor text reused across every competencyThe scale stops distinguishing what it was built to distinguishWrite competency-specific anchors using the five steps above
5 given for "no red flags" rather than exceptional evidenceScore inflation makes every candidate look strong on paperRequire two strong examples, or one unusually well-reasoned one, for a 5
Scores averaged across competencies with no floorA single 1 on a must-have gets diluted by unrelated 4s and 5sSet a per-competency floor for must-haves, separate from the average
New interviewers given the scale with no example scored answersTheir first few candidates become the calibration exercise, at the candidate's expenseScore two or three sample transcripts together before their first real loop

When five points are the wrong choice

A 1–5 scale is not the right default for every situation. If your interview loop is short (a 15-minute phone screen) or the interviewers are new and not yet trained on anchors, a 1–4 scale that removes the middle box altogether is usually more reliable, because it forces a lean one way or the other instead of hiding uncertainty in a 3. Five points earn their keep once interviewers have enough practice to reliably tell a 3 from a 4, which usually means a handful of calibrated loops, not the first one.

Questions people ask

Why does this scale skip a 0?

Because 1 already means the candidate showed the behavior and it was weak or missing evidence. A 0 invites interviewers to score a competency they never actually asked about, which is a coverage problem, not a rating. Handle missing coverage as 'not assessed' in a separate field, not as a score.

Can I use this scale with only three competencies instead of five?

Yes. The five points and their meaning do not depend on how many competencies you score. Use as many rows as your interview loop actually covers, and give each one its own five anchors rather than reusing one generic set.

What do I do when an interviewer keeps giving everyone a 3?

Ask them to point to the sentence in the transcript or notes that matches the anchor they chose. A 3 with no evidence attached is usually an interviewer who did not ask enough follow-up questions, not a genuinely average candidate. See the calibration steps below.

Should a 5 ever be given on a first-round screen?

Rarely, and say so in your guidance. A 5 means evidence strong enough to remove doubt, and a 20-minute screen seldom produces that much evidence for anything beyond a single narrow skill. Reserve 5s for the interview stage built to test that competency in depth.