Every candidate looks like a great hire now. AI made sure of it.

Candidates now arrive at interviews pre-coached by AI, with their resumes optimized to pass every checkpoint. Polish has stopped being a signal. The traditional hiring process was built to read exactly the cues that AI is now best at producing, and the signals hiring managers once relied on have weakened as a result. And for roles where the wrong hire carries real business consequences, losing the ability to tell who will actually perform is not a minor inconvenience. It is a material risk, and it exposes the business to unnecessary turnover, reduced performance, and heavier investment for talent growth and development.
So how do you observe the behaviors that matter most, before someone is in the role?
Not by asking better questions, but rather by putting candidates in situations designed to elicit that behavior.
The limits of predicting from paper
Credentials tell you what someone has done. Structured interviews tell you what someone says they would do. Neither lets you observe what they actually do in the moments that count.
This distinction matters most in client-facing, relationship-driven roles, where the performance gap between a strong hire and a weak one plays out in real business outcomes (revenue, retention, client growth). Organizations that hire at scale in these roles carry that gap across hundreds of decisions at a time.
The better approach is to watch candidates do the work before you hire them. Put them in simulated, role-relevant scenarios, and pair the simulation with a second, different kind of measure so no single method carries the whole decision. That combination is what lets you evaluate real performance before anyone is in the role. Organization-specific simulations provide a clear read on who is ready and capable of performing on day one. In a world of AI-supported candidate signals, the use of simulations makes the process harder to prep for. It is harder to fake. And, when designed well, it is substantially more predictive than other hiring methods.
What counts as evidence
Claims about predictive power are easy to make. Evidence for them is rarer than you would expect.
A predictive validity study, the kind that links pre-hire assessment scores to how someone actually performs once hired, is some of the hardest evidence to produce and the rarest to see. Many assessments are validated against proxies: another test, or a theoretical model of the role, rather than real results on the job. Connecting scores to concrete business outcomes and doing the statistical work to show the link holds, takes years of shared data and a level of commitment from both the assessment provider and the client that most partnerships never reach. That is precisely why it is worth asking for. A provider who can show how assessment scores track to training completion, retention, and first-year output is offering something categorically different from one who can only show a correlation with another test.
Why simulation holds up where other methods do not
When a candidate sits across from a trained assessor (someone playing the client or prospect on the other side of the conversation) and has to work through a real situation, they cannot rely on a rehearsed answer. The scenario is specific. The stakes feel real. What you see is close to what you would get on the job.
That is the value of simulation-based assessment: it does not test what candidates know about the role.
It shows how they use what they know when a real person is on the other side of the conversation, before the stakes are real.
For roles that carry significant business responsibility, this distinction is the whole game. The cost of the wrong hire in a high-stakes client-facing role is not just a missed quota for a quarter - It plays out in relationships that do not develop, clients who leave, and productivity losses that compound over time. Getting those hiring decisions right, at scale, with consistency, requires methods that are built for predictive accuracy, not just candidate experience or hiring speed.
What this means for how organizations think about hiring
Most organizations are still optimizing the wrong things in their hiring process. They invest heavily in employer branding, application flow, and interview structure, all of which matter, but less in the core question: does our hiring process actually predict who will succeed in this role?
AI has sharpened the stakes here. If every candidate can present as polished and prepared, screening based on presentation becomes less useful. What holds up is direct observation of the behaviors that the job requires.
A few principles worth building from:
- Measure what the job requires, not what is easy to measure. Cognitive tests and personality questionnaires have their place, but they do not look much like the job. The closer the assessment is to the actual work, the better it predicts performance in it.
- Ask what your assessment predicts. Training completion? Retention? First-year output? Most organizations cannot answer that question today, largely because providers have rarely been asked to prove it. It is a fair thing to ask for.
- Take the human element seriously. In a simulation, a candidate is having a real conversation, responding in real time, navigating a situation that requires judgment. Even with the help of AI, that is hard to game. And it remains one of the strongest predictors of on-the-job performance available.
The data exists to make hiring decisions more accurate, fairer, and more directly tied to business outcomes. For organizations operating in high-stakes roles at scale, there is too much on the line to rely on methods that cannot hold up to that standard.
You may be interested in BTS’ thought leadership in the five talent shifts AI is forcing now.
Related content




