How to Evaluate Candidates After an Interview
Top Picks
Submit PRStay Ahead of the Market
Get the latest startup funding, hiring trends and global opportunities delivered to your inbox every week.
To evaluate candidates after an interview, have each interviewer record evidence against the same role criteria before the group talks, compare those records criterion by criterion, look into any disagreement, then make and document the decision. That sequence ties the final choice to what candidates actually said and showed, not to the most confident voice in the meeting.
Below: each step, a scorecard and a decision meeting agenda you can reuse. This is the second supporting guide in [LINK: How to Select the Right Candidate], and it sits near the end of the hiring process.
Key takeaways
Write individual feedback before any group discussion.
Use the same criteria and rating definitions for every candidate.
Record evidence and uncertainty next to every rating.
Compare essential criteria first. One strong answer does not cover the whole role.
When interviewers disagree, ask what evidence sits behind each view.
If nobody clearly fits, try a focused follow-up before changing the role or the search.
Document the decision, the reasons and who owns each next action.
How to evaluate candidates after an interview, step by step
Record: each interviewer writes down evidence and uncertainties alone.
Compare: line the records up criterion by criterion.
Investigate: look into disagreements and gaps in the evidence.
Decide: choose, then write down why.
What should you evaluate after an interview?
Evaluate the evidence this interview stage was built to produce: answers to the planned questions, plus the result of any assessment or work sample in the stage. Not everything the candidate is. Not everything the role will eventually need.
Match each criterion to the stage meant to test it. If a technical round never touched stakeholder communication, leave that criterion unrated for that round. A blank is honest. A low score for something nobody asked about is not.
Next, separate what you observed from what you concluded. "She seemed sharp" is a conclusion. "She described changing the approval order on supplier invoices and said the backlog cleared within a month" is an observation. The U.S. Office of Personnel Management (OPM) tells interviewers to take notes that summarise what the candidate actually said, and to keep out judgments about the candidate or their personality.
Collect interviewer feedback before the group discussion
Ask every interviewer to write their record alone, before the debrief. For each criterion: what the candidate said or did, the rating that supports, and what the interviewer could not tell. That last part is easy to skip. Insist on it.
The reason is independence. Daniel Kahneman, in a 2022 interview with Issues in Science and Technology, named independence as the central principle of what he and his co-authors call decision hygiene: pieces of information should be as independent of one another as possible. Once one interviewer announces a verdict, the rest stop being independent witnesses.
OPM's training has interviewers rate the candidate individually, and it calls for a consensus discussion only if the ratings do not already agree. In that discussion, each interviewer explains their rating from their notes.
Use a consistent interview scorecard
Use one scorecard for every interviewer and every candidate in the same role. Five fields do the job:
Criterion: a capability the role needs, named in the same words each time.
Evidence: what the candidate said or did, taken from the interviewer's notes.
Rating definition: what each level on the scale means for this criterion.
Uncertainty: what the interviewer could not tell.
Interviewer and stage: who recorded it and in which round.
OPM's training says to pick one rating range for all criteria, usually three to seven levels, and to label at least three of them. It also says a structured interview typically assesses four to six competencies. A four-level scale with a separate "not enough evidence" option is one workable setup:
1, Clear gap: the evidence shows the need is not met.
2, Partly meets: some relevant evidence, with gaps that matter.
3, Meets: the evidence matches what the role needs.
4, Exceeds: the evidence goes beyond what the role needs.
NE, Not enough evidence: the criterion was not covered, or the answer was too vague to rate.
Write these meanings down before the interviews begin. A rating without a definition is a private opinion with a number on it.
Even then, the number looks more exact than it is. A 2022 meta-analysis in the Journal of Applied Psychology ranked structured interviews as the top selection procedure, yet its authors concluded that many selection methods predict job performance considerably less well than previously thought. Kahneman made a similar point: structured interviews are clearly better than unstructured ones, but neither predicts job success very well. Treat each rating as a label for the evidence beside it, not a measurement.
Some employers bring in specialist interviewers for one round, from inside the company or outside it. Give them the same scorecard, criteria and rating definitions. Their input then arrives in the same format as everyone else's and can be compared on equal terms. WorkNation's Hiring Solutions page describes bringing in interviewers who know the role for technical, functional, leadership and culture rounds, and handling panels, scheduling, scorecards and feedback loops.
Compare candidates against the role requirements
Sort the criteria into essential and desirable before you meet the finalists. Then compare on the essentials first, one criterion at a time, with each candidate's rating and evidence side by side. Strengths and tradeoffs come after: say plainly what each person is strong at and what the role would have to absorb.
Watch for the halo effect, as OPM's training calls it: a rating on one competency bleeding into the ratings for others. Kahneman went further. Past a small core of information, he said, people tend to overweight minor details, with one exception: a real deal-breaker should count. Let the essential criteria decide. Interesting details can wait.
What should you do when interviewers disagree?
Start with the evidence, not the ratings. Ask each interviewer to point to what the candidate said that supports their view. OPM's consensus step works the same way, with each person explaining a rating from their notes.
Then check that everyone rated the same thing. One interviewer may have scored communication with clients, another communication inside the team. The numbers differ because the questions did.
Last, name what is still unknown and write it down as a follow-up question.
Decide in advance how a final rating is set. OPM lists consensus, majority and average as options. Pick one and record it. An average of a 4 and a 2 is a 3, which describes neither interviewer's evidence.
What if no candidate clearly meets the criteria?
First decide whether the gap is one missing piece of evidence or a real shortfall. If the panel never saw a candidate handle a budget, a short follow-up on that criterion alone can settle it. Rate it on the same scale.
If the evidence stays thin, or the finalists fall short on an essential criterion, look at the role and the hiring process rather than at the bar. Are the essentials realistic for the scope and pay? Did the search reach the right people? Relaxing an essential criterion is allowed, but write down what changed and why, so the record shows the new standard.
Run and document the final decision meeting
Keep the meeting short and ordered. Every item has an owner who leaves with an action.
Evidence review. Owner: meeting chair. Show each criterion with every interviewer's rating and evidence. Output: one shared view of the record.
Gaps. Owner: chair. List every criterion marked "not enough evidence". Output: follow-up questions, each with a named owner.
Disagreements. Owner: the interviewers who differ. Each gives the supporting evidence. Output: agreed ratings, or a recorded split.
Decision. Owner: hiring manager. State the choice and the reasons, tied to the essential criteria. Output: a recorded decision.
Approval. Owner: whoever signs off on the hire. Output: approval, or stated conditions.
Communications. Owner: recruiter or HR lead. Decide who tells the chosen candidate and who tells the others, and by when. Output: dated communication tasks.
Afterwards, file the final ratings, the evidence behind them, open uncertainties, any criterion that changed, and the approver. OPM's training has interviewers sign and date their rating forms and initial later changes. What records you must keep, and for how long, depends on your country and state, so check the rules for your hiring location.
Post-interview scorecard and worked example
Illustrative example: a 12-person company hiring its first operations manager, continuing the role from [LINK: Article 1.1 title] and [LINK: Article 3.1 title]. Candidate A and Candidate B are invented for this example. There are three essential criteria: process ownership, vendor and budget handling, and clear communication with non-specialists. Interviewer 1 is the founder and Interviewer 2 is the finance lead, both in the same stage 2 interview. Ratings below are listed in that order.
Application-stage evidence is what a candidate claims on paper. Interview-stage evidence is what they show when asked for specifics. The two often differ.
Criterion 1: Process ownership
Candidate A. Application stage: lists order-processing improvements as a past responsibility. Interview stage: gave the starting problem, the change and the measured result when probed. Ratings: 3 and 3. Uncertainty: the work was at a much larger company.
Candidate B. Application stage: lists a wider scope of process work. Interview stage: answered mostly in "we" terms, and probing did not isolate the candidate's own role. Ratings: 2 and NE. Uncertainty: personal contribution unclear.
OPM's training advises asking for the candidate's own role whenever an answer comes in "we did" terms. That is what the probing was for.
Criterion 2: Vendor and budget handling
Candidate A. Application stage: no budget work listed. Interview stage: described renegotiating one supplier contract and the reasoning behind it. Ratings: Interviewer 1 gave 3, Interviewer 2 gave 4. Uncertainty: one example only.
Candidate B. Application stage: lists vendor management. Interview stage: described tracking three suppliers' deliveries but no spending decisions. Ratings: 2 and 2. Uncertainty: budget ownership not discussed.
Criterion 3: Clear communication with non-specialists
Candidate A. Interview stage: explained a delay to a hypothetical customer in plain language but took a while to reach the point. Ratings: 3 and 3.
Candidate B. Interview stage: gave a short, plain explanation with next steps. Ratings: 4 and 4.
Reading the result: Candidate B leads on communication, and Candidate A has better evidence on the other two essentials. The interviewers split on A's budget handling, 3 versus 4. Their notes show one contract, so the panel records a 3 and adds a question about a second example. B's process ownership is unresolved and essential, so the panel books a short follow-up on that one point before deciding. Picking B on communication alone would let the best answer replace the full evaluation.
If you change one thing first, change the order: no discussion until every written record is in.
Need specialist interview support or a consistent decision process? Explore WorkNation's hiring solutions.
FAQ
How soon after an interview should you evaluate a candidate?
Write your record straight after the interview, while the details are fresh. OPM's structured interview training has interviewers review their notes and rate the candidate immediately after each interview. If you interview finalists over several days, record each one the same day, then hold the comparison meeting once every finalist has been seen.
Should interviewers discuss candidates before scoring them?
No. Score first, then discuss. Once one person shares a verdict, the others stop judging independently, and the meeting becomes agreement or argument instead of a review of evidence. Written records submitted beforehand also show who saw what, which makes later disagreements easier to settle.
Can you use the same scorecard for every interview stage?
Use the same role criteria and rating definitions at every stage, and mark which criteria each stage was designed to test. Rate only those, and mark the rest "not enough evidence" rather than low. That way a technical round and a leadership round feed one scorecard, and gaps show up as gaps, not weaknesses.
Sources
U.S. Office of Personnel Management, "Structured Interviews" training presentation (undated): https://www.opm.gov/policy-data-oversight/assessment-and-selection/structured-interviews/structured-interviews.pdf
Daniel Kahneman and Sara Frueh, "Try to Design an Approach to Making a Judgment; Don't Just Go Into It Trusting Your Intuition," Issues in Science and Technology, Spring 2022: https://live-issues-asu.ws.asu.edu/daniel-kahneman-interview-noise-judgment-decisionmaking
Paul R. Sackett, Charlene Zhang, Christopher M. Berry and Filip Lievens, "Revisiting meta-analytic estimates of validity in personnel selection," Journal of Applied Psychology 107(11), 2022: https://doi.org/10.1037/apl0000994
WorkNation, Hiring Solutions page (read 6 October 2026): https://worknation.buzz/hiring-solutions
Have a Funding Round or Exciting Update to Share?
Submit your press release to get featured on WorkNation and reach founders, investors, and tech leaders.

