Hiring on evidence rather than impression?

Share a role

Behavioural interview

A behavioural interview asks what somebody did rather than what they would do, on the theory that past behaviour is the best evidence available to you. The theory holds. The execution usually does not, because every candidate has been coached on the STAR format and arrives with four polished stories, and an interviewer who accepts the story as given has tested preparation. The design below is built to get past the prepared version.

The format, and what it settles

One interviewer, one hour, three or four competencies, and a written scale agreed before anybody meets the candidate. This is not a panel format. It works better one to one, because probing hard is easier when the candidate is not being watched by four people while they revise an answer.

Who is in the room

The interviewer
Three or four competencies, no more. Each one needs a primary question, at least three probes, and enough silence for the candidate to fill.
The scribe, if you have one
Verbatim notes of what the candidate said, not a summary of what the interviewer concluded. Quotes survive a debrief. Adjectives do not.
Nobody else
Additional observers change the format. If a second person needs a read, give them their own competencies in their own conversation.

The competencies come from the success profile agreed before sourcing started. Choosing them in the meeting is how a company ends up hiring for whichever trait the interviewer happens to value.

The five moves, in order

  1. 01

    The prepared story

    Run by Interviewer

    Settles: Nothing yet. Ask the question, let them run, do not interrupt. The rehearsed answer is the starting position and its job is to give you something to probe.

    Ends it: Cannot produce an example at all, after a rephrase and a pause. Either it did not happen or it cannot be retrieved under mild pressure, and both matter.

  2. 02

    The particulars

    Run by Interviewer

    Settles: Whether this is a memory or a construction. Dates, names, sequence, who else was in it.

    Ends it: The story moves from "I" to "we" and never comes back, and they cannot say which part was theirs when asked directly.

  3. 03

    The cost

    Run by Interviewer

    Settles: Judgement. Every real decision traded something away. Ask what it cost, who paid it, and what they would trade differently now.

    Ends it: A decision with no downside, told twice. Real work has costs and the people who did it remember them.

  4. 04

    The aftermath

    Run by Interviewer

    Settles: Durability. What happened to the thing once their involvement ended.

    Ends it: No idea, and no curiosity about it. It is the clearest tell that somebody was present at something rather than responsible for it.

  5. 05

    The counter-example

    Run by Interviewer

    Settles: Calibration. Ask for a time the same approach did not work.

    Ends it: Refuses the premise. A method that has never failed has either not been used much or will fail here without warning.

What ends it

  • Every story is a success story, including the one you explicitly asked to be a failure.
  • No name, no date and no number anywhere in an hour of examples.
  • The scale of the story grows as your interest visibly grows.
  • Blame sits entirely outside the candidate in every example, including the ones where you asked what they would do differently.
  • The answer to "what did it cost" turns out to be a benefit.

How to score it

Owned it
First person throughout, the disagreement named, the cost named, and an accurate account of what they would do differently and why they did not do it then.
Was in it
Real involvement and real detail, with the decisions belonging to somebody else. Perfectly hireable at that level. A problem only if you are hiring for the level above it.
Was near it
Correct vocabulary, no particulars, the story told at team scale. Usually not dishonesty, usually somebody describing the environment they worked in.
Evidence against
A specific observation that predicts a problem: the failure that was somebody else’s fault twice, the cost that turned out to be a benefit, the peer who cannot be named.

Score the competency, not the story. One example can be strong evidence of ownership and no evidence at all of collaboration, and collapsing it into a single overall impression throws away most of what you just heard.

The probes that break a rehearsed answer

STAR is a preparation format, and by now it is a very well prepared one. Somebody who has practised will hand you situation, task, action and result in a clean four-minute block. Everything useful comes after that block.

  • "When was this?" Prepared stories are undated. Real ones arrive with a quarter attached, and often with what else was going on at the time.
  • "Who disagreed with you?" Every real project has one. Somebody who was actually there names the person and characterises the objection fairly, occasionally better than they characterise their own position.
  • "What did you have to give up to do that?" It turns a result back into a decision.
  • "What happened to it after you left?" It separates the person who built something from the person who was standing next to it at launch.
  • "Whose idea was it originally?" Asked plainly and without suspicion. The answer is frequently generous and specific, and the generosity is itself evidence.
  • "What would your manager at the time say was the hardest part of working with you?" A better version of the weakness question, because it is attributed and therefore harder to invent.
  • Silence, for four seconds after an answer ends. The addendum people volunteer to fill it is regularly the most useful sentence in the hour.

Behavioural questions worth an hour, by competency

Ownership

"Tell me about something you owned that failed. When did you know, and who did you tell first?" The gap between knowing and telling is the competency, and it is the thing that will happen again here.

Judgement under pressure

"Describe a decision you had to make before you had the information you wanted." Listen for what they did to cheaply reduce the uncertainty, and for whether they set a condition that would reverse it.

Working with difficulty

"Who is the hardest person you have worked well with, and what did you change?" The pivot is what they changed. An answer that stays on the other person tells you how your team will hear their feedback.

Standards

"Tell me about work of yours that shipped below your own bar. What made you accept it?" People with real standards name the compromise precisely and are usually still slightly annoyed about it.

Learning

"What did you believe about this work two years ago that you no longer believe?" Anybody doing the work has been structurally wrong about something. A blank here means somebody stopped paying attention.

Behavioural and situational are different instruments

Behavioural: what you did

Produces evidence, because it happened and can be probed for particulars.

Produces reasoning, because it has not happened and cannot be checked.

Rewards experience, and quietly penalises a strong candidate stepping up who has not yet had the situation.

Levels that, which makes it the fairer instrument for somebody changing level or domain.

Breaks under coaching, which is why the probes above exist.

Breaks under a candidate who is articulate about work they have never done, which is why you follow it with "when have you come closest to this?"

Use it where a failure would be expensive and history is the best predictor available.

Use it where the person will genuinely be doing something new, which includes most first-time management hires.

Where behavioural interviews go wrong

  • Ten competencies in one hour, which produces ten prepared answers and no probing at all.
  • A scale invented during the debrief, so every interviewer has been applying a private definition of a good answer.
  • The same competency scored by four interviewers, each hearing a better-rehearsed version of the same story, with the improvement recorded as consistency.
  • Treating a fluent narrator as a strong candidate. Storytelling is a real skill and it is a job-relevant one for only some jobs.
  • Penalising a candidate whose examples are smaller than yours. Scale of company is not scale of ownership, and the reverse mistake is more common than people admit.

Behaviour you can watch instead of ask about

The strongest behavioural evidence is not always an answer. How somebody handles the parts of your process that are not the interview is behaviour too: whether they read the brief, how they respond to a reschedule, what they do when a task is deliberately underspecified, and whether they paste in an answer they were asked to write themselves.

Behavioural signals of that kind sit alongside the structured conversations in how Continuity1 screens, because the way somebody handles a test is evidence in a way that their account of handling a test is not.

What this looked like on a real role

27 days

to a Director of Engineering offer

The instrument carries into senior hiring, where a rehearsed answer is more polished and the cost of accepting it is higher. On one Director of Engineering search the decision makers ran two conversations and made the offer, because the assessment work had been done before they were in the room.

Read the SuperProcure engagement in full

The reason to probe this hard is what the alternative costs when it goes wrong: put a number on the wrong hire.

How Continuity1 runs this funnel

The screening above is the job. These are the numbers it produces when a function owns it end to end, set against the published benchmarks for the same market.

1 in 3
Shortlisted candidates you meet who become the hire
Aligned engagements run nearer 1 in 2, distant ones nearer 1 in 10. The market takes about 180 applicants to make one hire, and that sifting lands on your team rather than ours.
Continuity1 tracked engagements
~3
Interviews your team sits in, per hire
Ashby puts technical roles at 17.6 interviews per hire across the whole process, up 52% since 2021. The rest of that load sits with the function, not with you.
Ashby talent-trends report
1 in 9
Accepted offers that ghost before joining
Indian employers report nearly 4 in 10 offers dropped. We lose 1 in 9.
nasscom community
95%
Offers that close inside your stated band
20 of the last 21. A flat fee earns nothing from an inflated offer; a percentage of CTC earns more.
Continuity1 tracked engagements

Every brief becomes a success profile before sourcing starts, calibrated with the people who will manage the role. That calibration is the step most hiring skips, and it is why a shortlist either matches the job or matches the job advert.

You review a scored shortlist and make the calls. The filtering never lands on your calendar.

Questions teams ask

What is a behavioural interview?

An interview that asks for specific past examples rather than opinions or hypotheticals, on the theory that what somebody did is better evidence than what they say they would do. It is only as good as the probing, because the first version of any answer is the rehearsed one.

Is it spelled behavioural or behavioral?

Both. Behavioural is the British and Indian spelling, behavioral the American one, and they describe the same format. Nothing about the method changes with the spelling.

How many behavioural questions should I ask in an hour?

Three or four, each probed properly. Interviewers who plan ten get ten prepared answers and no evidence, then need another interview to find out what they missed.

How do I score a behavioural interview?

Against the competency, on a scale written down before you met anybody, with the observation attached to the score. Score each competency separately, because a single story can be strong evidence of ownership and no evidence at all of collaboration.

What if a candidate has no relevant example?

Ask for the closest thing they have done, then ask what would be different here. Missing evidence is missing evidence rather than a negative. Somebody stepping up is often accurate about the gap, and the accuracy is a real signal.

Do behavioural interviews work for junior hires?

Partly. Two years of work means fewer examples and smaller ones, so the same probes apply to smaller stories and situational questions carry more of the load. What does not work is scoring a junior candidate against examples the size of your own.

Related

Continuity1

Your talent acquisition function, delivered as a product. The system, the process, and the operators, in one unit you switch on.

Stay updated

Get hiring insights delivered monthly.

© 2026 Continuity1. All rights reserved.