Evidence-based scoring
Evidence-Based Scoring
A great interview is wasted if the scoring is guesswork. Learn how to convert the evidence you gathered into fair, consistent, defensible scores — using anchors and facts instead of gut feeling.
Introduction
You have asked great questions, probed skillfully, and gathered observable evidence. Now comes the moment that decides the candidate's fate: turning that evidence into a score. This is where many interviews quietly fail. Even a well-run interview becomes unfair if the scoring is based on a vague overall feeling.
Evidence-based scoring is the discipline of rating each candidate against predefined criteria, using the specific evidence they demonstrated. It replaces "I think she was an 8" with "she meets the level-4 anchor because she solved the task and considered edge cases." It is the bridge between gathering evidence and making a fair decision.
Learning Objectives
- Define evidence-based scoring.
- Explain why it beats overall-impression scoring.
- Use rating anchors to score consistently.
- Link every score to specific evidence.
- Avoid common scoring errors and biases.
Why Evidence-Based Scoring Matters
A number feels objective, but a number pulled from a gut feeling is just bias with a decimal point. When scores are not tied to evidence, they vary wildly between interviewers, reward confident talkers, and cannot be defended if a decision is questioned.
Evidence-based scoring makes evaluation consistent, comparable, and defensible. Two interviewers using the same anchors and the same evidence should reach similar scores — and that consistency is the heart of fairness.
What Is Evidence-Based Scoring?
Evidence-based scoring is rating a candidate on each competency by matching the specific behaviour they demonstrated to a predefined rating scale with clear anchors — not by forming an overall impression.
Think of a Judge and the Law
A fair judge does not decide on a hunch — they weigh the evidence against written law. A score is your verdict; the anchors are your law; the evidence is what you weigh against it.
Overall Impression vs Evidence-Based Scoring
Overall-Impression Scoring
- "I'd give her about an 8."
- One vague number for the whole interview.
- Driven by feeling and recency.
- Varies hugely between interviewers.
Evidence-Based Scoring
- Each competency scored separately.
- Every score tied to specific evidence.
- Matched to clear anchors.
- Consistent across interviewers.
Rating Anchors: The Backbone of Fair Scoring
A rating anchor describes what each score level actually looks like in observable terms. Anchors turn a vague scale into a shared standard everyone applies the same way.
| Score | Anchor (Example: Problem-Solving) |
|---|---|
| 1 – Poor | Could not approach the problem even with prompts. |
| 2 – Developing | Made progress only with significant help. |
| 3 – Solid | Reached a working solution with minor prompting. |
| 4 – Strong | Solved it independently and considered edge cases. |
| 5 – Excellent | Solved, optimized, tested, and explained trade-offs. |
How to Score, Step by Step
Review the Evidence
Look at your notes for what the candidate actually said and did on each competency.
Match to an Anchor
Find the anchor whose description best fits that evidence.
Record the Score and Reason
Write the score and the specific evidence that justifies it.
Score Each Competency Separately
Rate every competency on its own, so one strength or weakness doesn't spill over.
Worked Example: Scoring With Evidence
A candidate is assessed on three competencies. Notice how each score is justified:
| Competency | Evidence | Score |
|---|---|---|
| Problem-solving | Solved the task independently, handled one edge case after a prompt. | 4 |
| Communication | Explained each step clearly and checked understanding. | 5 |
| Fundamentals | Explained arrays well but could not describe time complexity. | 3 |
Score Promptly While Evidence Is Fresh
Memory fades fast, and the last few minutes of an interview tend to dominate a delayed judgment (recency bias). Score each competency as soon as possible — ideally during or immediately after the interview — while the evidence is still clear in your mind and your notes.
Common Scoring Errors to Avoid
| Error | What Happens |
|---|---|
| Halo / horn effect | One strong or weak area colours every score. |
| Central tendency | Rating everyone in the safe middle (all 3s). |
| Leniency / severity | Scoring everyone too high or too low. |
| Recency bias | The last answer dominates the score. |
| Impression scoring | Assigning a number from a gut feeling. |
Best Practices for Evidence-Based Scoring
- Score each competency separately against its anchors.
- Tie every score to specific, observable evidence.
- Match evidence to the closest anchor, not a gut number.
- Score promptly while the evidence is fresh.
- Use the full scale — avoid clustering everyone in the middle.
- Add a short justification note for every score.
Common Mistakes
Avoid These
- Giving one overall gut-feeling score.
- Scoring without any anchors.
- Letting one area colour all the others.
- Rating everyone a safe "3".
- Scoring days later from memory.
Do These Instead
- Score each competency with evidence.
- Use clear, predefined anchors.
- Keep competencies independent.
- Use the full range honestly.
- Score promptly with fresh notes.