All assessments
Communication & Professional Skills Interview & Hiring Skills Assessments Hub: Tests, Rubrics & Templates

Interview scoring template: rubric, question bank, and scorecard for hiring

A ready-to-use template for scoring the interviews you run: behavioral questions with look-fors, anchored 1-5 rubric, weights, and red flags.

Most small-business interviews are unscored conversations, which means the job goes to whoever the interviewer clicked with. Then the charming candidate with few specifics beats the substantial one, and you find out three months in. This page is the fix: a complete template for scoring the interviews you run. You get a question bank with look-fors, an anchored 1-to-5 rubric, role weights, red flags, and a short calibration routine for when more than one person interviews. It works whether your "panel" is you and your GM or just you.

What to score

Score observable behaviors, not personality. Six competencies cover most roles:

  • Structure and clarity: logical sequencing, stays on point, appropriate detail
  • Active listening and responsiveness: answers the question asked, clarifies ambiguity, adapts when redirected
  • Evidence orientation: specific examples, numbers, outcomes, and learnings instead of vague claims
  • Judgment and trade-offs: names constraints, prioritizes, anticipates second-order effects
  • Professionalism and rapport: respectful tone, composure, time awareness
  • Depth under probing: can go deeper on request; the story survives follow-up questions

What this template deliberately doesn't score: technical competence (use a role-specific work sample for that) and "culture fit" as a vague feeling. If team match matters, define the specific values-aligned behaviors and score those.

The question bank, with what good answers sound like

Behavioral questions (STAR format)

1. Influence without authority. "Tell me about a time you had to influence a decision without being in charge. What was the context, what did you do, and what changed?"
Look for: named stakeholders, a clear approach, real objections, a measurable outcome, and a learning. Weak answer: "I'm good at getting people on board" with no specific case.

2. Failure with accountability. "Describe a project or decision that didn't go as planned. What was your role, what did you learn, and what would you do differently?"
Look for: ownership without excuses, a specific failure mode, prevention steps. Weak answer: a failure that was entirely someone else's fault, or a humblebrag ("I care too much").

3. Prioritization under constraints. "Tell me about a time you had too many priorities and not enough time. How did you decide what to do?"
Look for: explicit criteria, trade-offs, communication with affected people, results. Weak answer: "I just worked harder and got it all done."

4. Conflict and alignment. "Share an example of a disagreement on a team. How did you handle it and what was the outcome?"
Look for: respectful language, genuine understanding of the other view, a resolution approach. Weak answer: the other person was simply wrong and eventually came around.

Situational questions, with the aligned answer marked

5. The ambiguous urgent request. "Your manager says 'we need this done as soon as possible,' but the scope is unclear and two other tasks are blocked."
A) Start immediately and fill in scope as you go. B) Ask for 15 minutes to clarify success criteria and priority, then propose a minimum viable plan. C) Wait for written requirements. D) Escalate that the request is unreasonable.
Most teams score B highest: it combines speed with clarity. What matters is whether the candidate's reasoning matches how you want your team to operate.

6. The rambling answer, reversed. Ask the candidate: "A colleague in a meeting talks for four minutes per question and drifts off topic. What do you do?"
A) Let them finish; interrupting is rude. B) Interrupt and say time is short. C) Politely redirect: restate the question, ask for a 60-second summary, then go deeper on one point. D) Stop engaging.
C is the aligned answer for most teams: respectful and structured. This question doubles as a communication sample.

Role-play prompt

7. Clarify and respond. "How would you improve our customer onboarding in 30 days?" Give no data. Tell them they may ask up to five clarifying questions first.
Look for: whether they actually ask clarifying questions, state assumptions, prioritize, and name measurable outcomes. Candidates who launch straight into a generic plan score low on listening and judgment.

Probing follow-ups that separate signal from story

When a candidate claims "I led a major process improvement that saved significant time," don't nod. Ask: "What was the baseline, and what was it after?" "What was your part versus the team's?" "How was it measured?" A real story gets more specific under probing. An inflated one gets vaguer. This single habit fixes more hiring mistakes than any other change you can make.

The rubric: anchored 1-to-5 scale

Rate each competency against these anchors, not against your impression:

  • 5, benchmark: consistently demonstrates the competency under pressure; answers are structured, specific, and adaptable; strong judgment and self-correction.
  • 4, proficient: demonstrates it with minor gaps; evidence is mostly specific; handles probing well.
  • 3, mixed signal: some structure and evidence, but inconsistent; over-indexes on storytelling or struggles with trade-offs.
  • 2, limited: frequently vague or reactive; limited evidence; difficulty staying aligned to the question.
  • 1, unsatisfactory: disorganized, non-responsive, or unprofessional; no evidence behind claims.

Example, active listening: a 5 clarifies ambiguity, answers the asked question, and adapts when redirected. A 3 answers generally but misses parts of the question. A 1 is repeatedly off-topic and ignores clarifications.

Weights and the total score

Default weights: structure and clarity 20%, active listening 15%, evidence orientation 20%, judgment and trade-offs 20%, professionalism 10%, depth under probing 15%. Shift weight toward what the role actually demands: customer-facing roles up-weight listening and professionalism; analytical roles up-weight judgment and evidence.

To calculate: rate each competency 1 to 5, convert to a percentage (score divided by 5, times the weight), and sum for a 0-to-100 total. Bands: 85 to 100, strong alignment across the board. 70 to 84, solid with specific development areas; hireable when the gaps are coachable and named. 55 to 69, mixed; advance only with a specific reason you can write down. Below 55, the interview evidence doesn't support the role as defined.

Red flags that trigger a review regardless of score

  • Repeated inability to provide evidence despite probing. The stories never get specific.
  • Dismissive or disrespectful behavior, discriminatory comments, or badmouthing identifiable former colleagues.
  • Claims that contradict the resume or application. Note it, verify it, and ask directly.

Running it consistently

  1. Pick the competencies and weights for the role before you schedule anyone.
  2. Ask every candidate the same core questions at the same stage. Structured follow-ups are fine; new core questions per candidate are not.
  3. Take evidence notes, not adjectives. "Quantified the outcome: cut close time from 9 days to 6" is a note. "Impressive" is not.
  4. Score right after the interview, before you talk to anyone about the candidate.
  5. If multiple people interview, everyone scores independently first, then debriefs: evidence first, scores second, recommendation last.

The 15-minute calibration, if more than one person interviews

Before your first real candidate: everyone silently scores the same sample answer, reveals scores simultaneously, cites two evidence bullets each, and agrees on what a 4 versus a 3 looks like for this role. Repeat quarterly or when ratings drift more than a point apart.

Rater errors to name in the debrief

  • Halo and horns: one great or awful moment drove every score.
  • Contrast effect: scoring against the previous candidate instead of the anchors.
  • Similarity bias: "they remind me of me."
  • Style overweighting: confident speaker read as competent operator. The rubric's whole job is to catch this one.

Fairness and documentation

Not legal advice, but these operational guardrails support consistent, defensible interviews: map every question to a role competency and drop the curiosity questions. Ask the same core questions for the same role. Offer accommodations (written responses, extra time, camera-off) when needed, and don't penalize tech disruptions unless they genuinely prevented evidence gathering in a remote interview. Store your scorecards and evidence notes, and never let a subjective descriptor be the only justification for a decision. Periodically review pass rates across groups where lawful, and investigate large disparities.

Run this scorecard automatically with Truffle

This template's weakness is you: after six back-to-back interviews, nobody scores consistently, and half the evidence notes never get written. Truffle is a candidate screening platform that combines one-way video interviews with resume screening and talent assessments, and it runs this exact structure for you. Every candidate answers the same questions on their own time, AI transcribes and scores each response against the criteria you set with the reasoning shown, and 30-second Candidate Shorts surface the moments worth watching. Your rubric, applied identically to candidate one and candidate forty. You still make every call.

Plans start at $49 a month. The 7-day free trial includes 30 credits and no credit card. Start free trial.

Built for anyone hiring

Run any of these assessments inside Truffle

Pair a skills test with one-way video interviews and resume screening. One Position Link. One ranked shortlist. Public pricing, 7-day free trial, no credit card.

Truffle is candidate screening software built for the AI age

Start free trial

7 days · 30 credits · no card required

Start typing to search 300+ pages on hiretruffle.com.