How to run an interview debrief.
The debrief is where good interviews go to die. Four people interview a candidate. Everyone forms a view. Then the meeting opens with “so, what did we think?”, the most senior or most confident voice sets the tone, and the information everyone collected quietly stops mattering. The interviews were structured; the decision was not.
A debrief meeting is an open discussion where the hiring team shares what each person learned during the interviews, and it is where the collective wisdom of the panel gets applied to the decision. This page shows how to run one: what has to happen before the meeting, the questions that kick it off, and a worked example of a panel working through a real disagreement.
Before the meeting: opinions are formed independently.
Google's re:Work hiring research, the finding behind its "Rule of Four," showed that four interviews predicted a hire with about 86% confidence, and each additional interview improved the decision by only about 1%. But that only holds if each interviewer forms their opinion independently. That means no sharing notes, scores, or impressions with each other before the interviews are done. The debrief is the time to come together, compare notes, and share what the team learned. If interviewers talk to each other along the way, the panel stops being several independent reads and becomes one opinion with several signatures.
Independence needs a few things in place:
- Every interviewer had an assignment. Each person covered specific areas, with questions prepared for them, so the panel’s coverage adds up instead of overlapping on first impressions.
- Every candidate was scored on the same scorecard for the role. One set of criteria, one scale, with a written description of what each score level looks like. If you do not have a scorecard yet, start with the complete interview scorecard template and build one before the interviews, not after.
- Scores and notes were submitted before the debrief. Written down, in the system, before anyone hears anyone else’s view.
The scorecard is not there to replace anyone’s judgment. It puts every candidate on the same scale, so the panel can compare what happened rather than compare impressions. Each interviewer still brings their read; the scorecard gives that read something to attach to.
Running the meeting.
The debrief is usually led by the hiring manager or the recruiter. Open by reviewing the role requirements and the key competencies the panel was evaluating, so the discussion runs against the role rather than against a general impression of the candidate. Then work through a consistent set of kickoff questions:
- Does anyone have questions for the other interviewers about the candidate? Each interviewer only saw a piece. This is where the pieces get connected.
- Any additional comments about the candidate? Things that did not fit a scorecard row still belong in the room.
- Is there anything further we need to investigate before making a decision? The discussion may surface an issue worth probing with the candidate directly, or a specific question to take into reference calls.
- Has anyone changed their hire or no-hire recommendation? Interviewers should feel free to update their view from what they learned in this meeting. A recommendation changing here is the meeting working, not someone caving.
- If the consensus is no hire, is the candidate a fit for another role? Worth asking at a company with other openings, or one growing fast enough to have them soon.
- What are the next steps? If there is a hire consensus, planning the offer. If there is no consensus, it may be that the hiring manager makes the call, or that the panel gathers specific missing information first.
Two rules of conduct make the questions work:
The leader speaks last. If the most senior person in the room shares their view first, people without strong opinions will follow it, and people with mild doubts will keep them quiet. Your read may be right, but you will make a better decision after hearing everyone else’s uninfluenced version first. The meeting leader’s other job is to make it safe to disagree: interviewers should be able to voice a view that differs from the consensus or from leadership without it costing them anything.
Scores triangulate; they do not decide. The best hiring decisions are human decisions made by adding objectivity to parts of the process. The process does not kick out a score that makes the hiring decision, so stay away from decision criteria like “anyone over a 3.5 is a hire.” Everyone has had a blind spot with a candidate at some point, and the debrief is where a blind spot gets caught: when the discussion shows your scores were wrong, change them. Use the scores as an important piece of data you triangulate with everything else the panel learned.
Watch out for the failure modes: no structure to the discussion, one overdominant voice, interviewers anchoring on each other, misaligned expectations about what the role needs, and a meeting that ends without clear next steps.
A worked example: a 4 and a 2 on the same competency.
Here is an illustrative debrief for a customer success manager role, with a panel of three. The company and candidate are invented; the mechanics are the point.
The panel: Maya (hiring manager, covering role outcomes and ownership), Tomás (peer CSM, covering the skills: renewal conversations and account planning), Priya (support lead, covering collaboration across teams). All three submitted scores and notes independently, and nobody compared notes beforehand. On the collaboration competency, Maya scored a 4 and Priya scored a 2.
Maya leads the meeting but holds her own read for last. She opens with the role’s key competencies, then the first kickoff question: does anyone have questions for the other interviewers?
Priya does. “Maya, your notes mention the candidate kept teams aligned on a renewal save. Did they say what they actually did after looping support in?” Maya checks her notes: the candidate said “I looped in support early and kept everyone aligned,” and Maya had not pushed further. Priya shares what happened when she did push: she asked twice for a case where the candidate stayed involved after a handoff, and did not get one. Every collaboration example the candidate gave was about routing the problem to someone else.
That takes the meeting to the fourth question before Maya has argued for her 4 at all: has anyone changed their recommendation? Maya has. “You asked the follow-up I didn’t. My 4 was resting on a claim, and your interview tested the claim.” She updates her score, and says so out loud, because a hiring manager visibly changing her mind on new information is what makes it safe for everyone else to.
Next steps: the panel’s picture is now coherent. Strong renewal skills from Tomás’s role-play, real ownership of a number from Maya’s interview, thin evidence on cross-team collaboration. Rather than rejecting on an area the panel only probed once, Maya decides on a final conversation targeting collaboration directly, with the specific probe named in advance. If it happens again after that conversation, the reference calls ask about it too.
Notice that no score decided anything. The 4 and the 2 did their job by pointing at the exact spot where two interviewers learned different things, the discussion resolved which read was better evidenced, and the decision-maker ended up with something better than a feeling to decide on.
The kickoff questions, ready to use.
- Does anyone have questions for the other interviewers about the candidate?
- Any additional comments about the candidate?
- Is there anything further we need to investigate before making a decision?
- Has anyone changed their hire or no-hire recommendation?
- If the consensus is no hire, is the candidate a fit for the company in another role? What are possible roles?
- What are next steps?
Alter them for the circumstances, but keep the sequence: connect what the panel saw first, decide last, and let the leader speak at the end.
Where Yardstick fits.
Everything above works with a shared document and discipline. What makes it routine instead of heroic is having the inputs ready without anyone assembling them. In Yardstick, the interview guide for the role is designed with AI before anyone is scheduled: which areas each interviewer covers, the questions they ask, what to listen for, and a scorecard with written anchors, edited by you before the first interview. Every interviewer scores against the same scorecard for the role, and those ratings roll up into one grid, so the debrief opens with the evidence already lined up instead of scattered across six documents. You can compare candidates on the same evidence side by side, and the built-in chase missing scorecards agent sends configured reminders to interviewers, escalating to the hiring manager at the final step, so the debrief does not start with two scorecards missing.
Common questions about interview debriefs.
Should interviewers compare notes before the debrief?
No. Each interviewer submits scores and notes before hearing anyone else's view, and nobody shares impressions between interviews. The value of a panel is several independent reads, and that value evaporates if opinions get shared along the way. The debrief is the place where the notes come together.
What if the panel still disagrees after the discussion?
Then the disagreement is real, and that is a finding. It may mean there is something further to investigate before deciding, with the candidate directly or in the reference calls. If a decision has to be made without consensus, the hiring manager makes it, with the disagreement visible. Do not average it away; an average of a 4 and a 2 is a 3 that nobody observed.
Does the highest total score win?
No. The scores organize the discussion; they are not a cutoff and they do not auto-decide anything. Two candidates with the same total can be very different hires, and the debrief is where that difference gets said out loud. The hiring manager makes the call and owns it.
How long should a debrief take?
Usually 30 minutes when the scores and notes are in beforehand. The meetings that run long are the ones doing the scoring and the remembering live, which is exactly the work that should have happened before the room convened.
Is this just for big panels?
No. Even with two interviewers, committing scores independently and comparing against written anchors changes what the conversation is about. The failure mode this fixes is not panel size; it is the unstructured “so, what did we think?” opener, which fails at any size.
Run your next debrief on evidence.
The full version of what this page describes is an interview guide built for the specific role you’re hiring: every interviewer with an assignment, questions prepared for them, and a scorecard with written anchors, so two interviewers hearing the same answer land on the same score and the debrief opens with every rating already in one grid. Yardstick designs that guide with AI from the role you describe, and you edit it before anyone is scheduled. Your first 3 Jobs are free.
.webp?dpl=dpl_UKtdTx42sfkvBwnduSu8q6aoznXB)