Conducting Interviews (Structured, Behavioral)
Scope
Covers
- Preparing and running structured interviews (screen + loop) with consistent criteria
- Behavioral interviewing mapped to competencies/values
- Getting to “substance over polish” (avoiding “confident but shallow” signal)
- Capturing evidence, scoring consistently, and writing a debrief-ready summary
When to use
- “Help me conduct interviews for a .”
- “Create an interview script / interview question set / scorecard for .”
- “Design an interview loop and structured rubric for .”
- “Improve interviewer consistency and reduce bias.”
When NOT to use
- You need to define the role outcomes or write the job description (use
writing-job-descriptions first)
- You need to make a final hiring decision, design work samples, or run reference checks (use
evaluating-candidates)
- You need to conduct user/customer research interviews (use
conducting-user-interviews — completely different skill)
- You need to onboard the person after hiring (use
onboarding-new-hires)
- You need legal/HR compliance guidance or to adjudicate complex employment risk (this skill is not legal advice)
- You need compensation/offer strategy or negotiation coaching
Inputs
Minimum required
- Role + level + function (e.g., “Senior PM”, “Engineering Manager”)
- Interview stage(s) to design/run (screen, hiring manager, panel, etc.) + duration(s)
- Evaluation criteria: 4–8 competencies/values to measure (or your existing rubric)
- Company/team context candidates should know (mission, what’s hard, why now)
- Candidate materials (resume/portfolio) + any areas to probe
Missing-info strategy
- Ask up to 5 questions from references/INTAKE.md.
- If criteria aren’t provided, propose a default criteria set and clearly label it as an assumption.
Outputs (deliverables)
Produce an Interview Execution Pack in Markdown (in-chat; or as files if requested):
- Interview plan (stage purpose, criteria, agenda, timeboxes)
- Question map (questions → competency/value → what good looks like → follow-up probes)
- Interviewer script (opening, transitions, probes, close)
- Notes + scorecard (rating anchors + evidence capture)
- Debrief summary template (evidence-based strengths/concerns + hire/no-hire signal + follow-ups)
- Risks / Open questions / Next steps (always included)
Templates: references/TEMPLATES.md
Expanded guidance: references/WORKFLOW.md
Workflow (7 steps)
1) Intake + define the stage
- Inputs: user request; references/INTAKE.md.
- Actions: Confirm role, stage(s), duration, and who else interviews. Identify must-measure criteria and any “must not” red flags.
- Outputs: Interview brief + assumptions/unknowns list.
- Checks: You can state the stage goal in one sentence (e.g., “screen for X; sell Y; decide Z”).
2) Lock evaluation criteria (don’t improvise later)
- Inputs: competencies/values; role context.
- Actions: Choose 4–8 criteria; define 1–2 “strong” and “weak” anchors per criterion. Ensure each criterion is observable via evidence.
- Outputs: Criteria table with anchors.
- Checks: Every criterion has a definition + evidence hints; no criterion is “vibe”.
3) Build the question map (behavioral first)
- Inputs: criteria table.
- Actions: Write 1–2 primary questions per criterion (behavioral: “tell me about a time…”). Add probes that force specifics (role, constraints, trade-offs, results, what you’d do differently). Add two global questions: “How did you prepare?” and “Why here?”
- Outputs: Question map table.
- Checks: Each question maps to exactly one primary criterion; no double-barreled questions.
4) Write the interviewer script (runbook)
- Inputs: question map; timeboxes.
- Actions: Assemble an interview flow: opening (set context + structure), question sequence, note-taking reminders, and a consistent close: “Is there anything else you want to make sure we covered?”
- Outputs: Interviewer script with timestamps.
- Checks: Script fits in time; includes “sell” moments appropriate to stage; includes candidate questions time.
5) Prepare for “substance over polish”
- Inputs: question map; candidate materials.
- Actions: Add “substance checks” for polished communicators (ask for concrete examples, counterfactuals, and specific decisions). Add “structure help” for less polished candidates (rephrase, clarify what’s being asked) without leading.
- Outputs: Substance-vs-delivery guardrails embedded in the script.
- Checks: The plan reduces false positives from confident delivery and false negatives from imperfect structure.
6) Score using evidence (immediately after)
- Inputs: notes; scorecard template.
- Actions: Fill the scorecard with evidence snippets before discussing with others. Rate each criterion with anchors. Write a 5–8 sentence evidence-based summary and list follow-up questions.
- Outputs: Completed notes + scorecard + summary.
- Checks: Every rating has supporting evidence; the overall recommendation is consistent with criterion ratings.
7) Debrief + quality gate + finalize pack
- Inputs: completed scorecard; debrief template.
- Actions: Produce the debrief-ready packet; run references/CHECKLISTS.md and score with references/RUBRIC.md. Include Risks/Open questions/Next steps.
- Outputs: Final Interview Execution Pack.
- Checks: Clear recommendation + uncertainty; fair process; next steps defined (additional interview, reference check, work sample, etc.).
Quality gate (required)
- Use references/CHECKLISTS.md and references/RUBRIC.md.
- Always include: Risks, Open questions, Next steps.
Examples
Example 1 (Screen): “Create a 30-minute phone screen for a Senior Product Manager. I want to evaluate product sense, execution, and collaboration. Output the Interview Execution Pack with a question map and scorecard.”
Expected: timeboxed script, behavioral questions, clear anchors, and a scorecard that captures evidence.
Example 2 (Loop): “Design a structured interview loop for a Staff Engineer, including a hiring manager interview and a cross-functional panel. Map questions to our values and include a debrief template.”
Expected: stage goals, consistent criteria across interviewers, and artifacts that make debriefs evidence-based.
Boundary example (redirect): “Help me decide which of these 3 candidates to hire based on their interview notes and references.”
Response: redirect to evaluating-candidates — this skill designs and runs interviews, it does not synthesize cross-candidate hiring decisions.
Boundary example (wrong domain): “I need to interview 10 users about their onboarding experience with our product.”
Response: redirect to conducting-user-interviews — this skill is for hiring interviews, not user/customer research.
Boundary example (missing criteria): “Just tell me if this candidate is good; I don’t have criteria or notes.”
Response: require criteria + evidence; propose default criteria and ask the user to paste notes or run a structured interview first.
Anti-patterns (common failure modes)
- ”Wing it” interviewing — Skipping structured criteria and question maps, then relying on gut feel. This produces inconsistent signals and legal risk. Always lock criteria before writing questions.
- Brainteaser / hypothetical-only questions — Asking “How many golf balls fit in a school bus?” or purely hypothetical scenarios instead of behavioral evidence. These predict interview prep, not job performance.
- Halo/horns from first 5 minutes — Letting initial rapport (or lack thereof) color all subsequent scoring. Mitigate by scoring each criterion independently with evidence before writing an overall summary.
- Identical questions for every role — Reusing the same generic question bank regardless of role, level, or competency. Every question should map to a specific criterion for this specific role.
- Skipping the “substance over polish” check — Rewarding confident, articulate delivery while penalizing candidates who need a moment to organize their thoughts. Always include specificity probes and structural support for less polished communicators.
1---2name: conducting-interviews3description: Run structured behavioral HIRING interviews: plan, questions, scorecard, debrief. See also: conducting-user-interviews (product research).4---56# Conducting Interviews (Structured, Behavioral)78## Scope910**Covers**11- Preparing and running structured interviews (screen + loop) with consistent criteria12- Behavioral interviewing mapped to competencies/values13- Getting to “substance over polish” (avoiding “confident but shallow” signal)14- Capturing evidence, scoring consistently, and writing a debrief-ready summary1516**When to use**17- “Help me conduct interviews for a <role>.”18- “Create an interview script / interview question set / scorecard for <role>.”19- “Design an interview loop and structured rubric for <role>.”20- “Improve interviewer consistency and reduce bias.”2122**When NOT to use**23- You need to define the role outcomes or write the job description (use `writing-job-descriptions` first)24- You need to make a final hiring decision, design work samples, or run reference checks (use `evaluating-candidates`)25- You need to conduct user/customer research interviews (use `conducting-user-interviews` — completely different skill)26- You need to onboard the person after hiring (use `onboarding-new-hires`)27- You need legal/HR compliance guidance or to adjudicate complex employment risk (this skill is not legal advice)28- You need compensation/offer strategy or negotiation coaching2930## Inputs3132**Minimum required**33- Role + level + function (e.g., “Senior PM”, “Engineering Manager”)34- Interview stage(s) to design/run (screen, hiring manager, panel, etc.) + duration(s)35- Evaluation criteria: 4–8 competencies/values to measure (or your existing rubric)36- Company/team context candidates should know (mission, what’s hard, why now)37- Candidate materials (resume/portfolio) + any areas to probe3839**Missing-info strategy**40- Ask up to 5 questions from [references/INTAKE.md](references/INTAKE.md).41- If criteria aren’t provided, propose a default criteria set and clearly label it as an assumption.4243## Outputs (deliverables)4445Produce an **Interview Execution Pack** in Markdown (in-chat; or as files if requested):46471) **Interview plan** (stage purpose, criteria, agenda, timeboxes)482) **Question map** (questions → competency/value → what good looks like → follow-up probes)493) **Interviewer script** (opening, transitions, probes, close)504) **Notes + scorecard** (rating anchors + evidence capture)515) **Debrief summary template** (evidence-based strengths/concerns + hire/no-hire signal + follow-ups)526) **Risks / Open questions / Next steps** (always included)5354Templates: [references/TEMPLATES.md](references/TEMPLATES.md) 55Expanded guidance: [references/WORKFLOW.md](references/WORKFLOW.md)5657## Workflow (7 steps)5859### 1) Intake + define the stage60- **Inputs:** user request; [references/INTAKE.md](references/INTAKE.md).61- **Actions:** Confirm role, stage(s), duration, and who else interviews. Identify must-measure criteria and any “must not” red flags.62- **Outputs:** Interview brief + assumptions/unknowns list.63- **Checks:** You can state the stage goal in one sentence (e.g., “screen for X; sell Y; decide Z”).6465### 2) Lock evaluation criteria (don’t improvise later)66- **Inputs:** competencies/values; role context.67- **Actions:** Choose 4–8 criteria; define 1–2 “strong” and “weak” anchors per criterion. Ensure each criterion is observable via evidence.68- **Outputs:** Criteria table with anchors.69- **Checks:** Every criterion has a definition + evidence hints; no criterion is “vibe”.7071### 3) Build the question map (behavioral first)72- **Inputs:** criteria table.73- **Actions:** Write 1–2 primary questions per criterion (behavioral: “tell me about a time…”). Add probes that force specifics (role, constraints, trade-offs, results, what you’d do differently). Add two global questions: “How did you prepare?” and “Why here?”74- **Outputs:** Question map table.75- **Checks:** Each question maps to exactly one primary criterion; no double-barreled questions.7677### 4) Write the interviewer script (runbook)78- **Inputs:** question map; timeboxes.79- **Actions:** Assemble an interview flow: opening (set context + structure), question sequence, note-taking reminders, and a consistent close: “Is there anything else you want to make sure we covered?”80- **Outputs:** Interviewer script with timestamps.81- **Checks:** Script fits in time; includes “sell” moments appropriate to stage; includes candidate questions time.8283### 5) Prepare for “substance over polish”84- **Inputs:** question map; candidate materials.85- **Actions:** Add “substance checks” for polished communicators (ask for concrete examples, counterfactuals, and specific decisions). Add “structure help” for less polished candidates (rephrase, clarify what’s being asked) without leading.86- **Outputs:** Substance-vs-delivery guardrails embedded in the script.87- **Checks:** The plan reduces false positives from confident delivery and false negatives from imperfect structure.8889### 6) Score using evidence (immediately after)90- **Inputs:** notes; scorecard template.91- **Actions:** Fill the scorecard with evidence snippets before discussing with others. Rate each criterion with anchors. Write a 5–8 sentence evidence-based summary and list follow-up questions.92- **Outputs:** Completed notes + scorecard + summary.93- **Checks:** Every rating has supporting evidence; the overall recommendation is consistent with criterion ratings.9495### 7) Debrief + quality gate + finalize pack96- **Inputs:** completed scorecard; debrief template.97- **Actions:** Produce the debrief-ready packet; run [references/CHECKLISTS.md](references/CHECKLISTS.md) and score with [references/RUBRIC.md](references/RUBRIC.md). Include Risks/Open questions/Next steps.98- **Outputs:** Final Interview Execution Pack.99- **Checks:** Clear recommendation + uncertainty; fair process; next steps defined (additional interview, reference check, work sample, etc.).100101## Quality gate (required)102- Use [references/CHECKLISTS.md](references/CHECKLISTS.md) and [references/RUBRIC.md](references/RUBRIC.md).103- Always include: **Risks**, **Open questions**, **Next steps**.104105## Examples106107**Example 1 (Screen):** “Create a 30-minute phone screen for a Senior Product Manager. I want to evaluate product sense, execution, and collaboration. Output the Interview Execution Pack with a question map and scorecard.” 108Expected: timeboxed script, behavioral questions, clear anchors, and a scorecard that captures evidence.109110**Example 2 (Loop):** “Design a structured interview loop for a Staff Engineer, including a hiring manager interview and a cross-functional panel. Map questions to our values and include a debrief template.” 111Expected: stage goals, consistent criteria across interviewers, and artifacts that make debriefs evidence-based.112113**Boundary example (redirect):** “Help me decide which of these 3 candidates to hire based on their interview notes and references.”114Response: redirect to `evaluating-candidates` — this skill designs and runs interviews, it does not synthesize cross-candidate hiring decisions.115116**Boundary example (wrong domain):** “I need to interview 10 users about their onboarding experience with our product.”117Response: redirect to `conducting-user-interviews` — this skill is for hiring interviews, not user/customer research.118119**Boundary example (missing criteria):** “Just tell me if this candidate is good; I don’t have criteria or notes.”120Response: require criteria + evidence; propose default criteria and ask the user to paste notes or run a structured interview first.121122## Anti-patterns (common failure modes)1231241. **”Wing it” interviewing** — Skipping structured criteria and question maps, then relying on gut feel. This produces inconsistent signals and legal risk. Always lock criteria before writing questions.1252. **Brainteaser / hypothetical-only questions** — Asking “How many golf balls fit in a school bus?” or purely hypothetical scenarios instead of behavioral evidence. These predict interview prep, not job performance.1263. **Halo/horns from first 5 minutes** — Letting initial rapport (or lack thereof) color all subsequent scoring. Mitigate by scoring each criterion independently with evidence before writing an overall summary.1274. **Identical questions for every role** — Reusing the same generic question bank regardless of role, level, or competency. Every question should map to a specific criterion for this specific role.1285. **Skipping the “substance over polish” check** — Rewarding confident, articulate delivery while penalizing candidates who need a moment to organize their thoughts. Always include specificity probes and structural support for less polished communicators.