Research20 September 2026

The interviewer is the bias

The objection is that an AI interviewer changes what people say. The research says the human one already does.

Noah McDonough4 minute read
The short version
  • People edit what they tell an interviewer to avoid embarrassment or consequences. That is the bias, and it has been measured since the 1990s.
  • Remove the human listener and reporting of sensitive information goes up. Validated against records, not self-report.
  • It is the belief that no person is judging that does the work. Whether the interviewer was actually automated made no difference.

The first question a research-literate buyer asks about an AI interviewer is whether it biases what people say.

It is the right question. It is also aimed at the wrong interviewer.

People edit what they say to a person

Tourangeau and Yan reviewed three decades of survey research on sensitive questions for Psychological Bulletin. Their conclusion was that misreporting is common, situational, and motivated: people edit what they report to avoid embarrassing themselves in front of an interviewer or to avoid consequences from third parties.

That is not a technology finding. It is a finding about what happens when another person is listening.

The fix has been known for thirty years

Richman, Kiesler, Weisband and Drasgow pooled 61 studies and 673 effect sizes in the Journal of Applied Psychology. Two results matter here. Computer questionnaires and paper questionnaires produced almost no difference in social desirability distortion. Computerised interviews and face-to-face interviews did: less distortion without the person, and least of all when respondents were alone, anonymous, and could go back and change an answer.

61 studies

Less social desirability distortion in computerised interviews than in face-to-face interviews. Almost no difference between computer and paper. The interviewer is the variable, not the screen.

Richman, Kiesler, Weisband and Drasgow, Journal of Applied Psychology, 1999

Kreuter, Presser and Tourangeau then did the study most of this field cannot: they checked the answers. Recent university graduates were randomly assigned to a human telephone interviewer, an automated voice system, or a web form, and asked about things like low grades and academic warnings. The answers were compared to university records. Reporting of sensitive information and accuracy were highest on the web, lowest with the human interviewer, and the automated voice sat in between.

Validated against records

Randomised to a human phone interviewer, an automated voice, or a web form. Reporting of sensitive information and accuracy rose as the human was removed.

Kreuter, Presser and Tourangeau, Public Opinion Quarterly, 2008

It is the belief that matters

Lucas, Gratch, King and Morency ran the cleanest test of the mechanism. Two hundred and thirty-nine people were interviewed by a virtual interviewer about their health. Half were told it was automated. Half were told a person was operating it. Separately, and without their knowledge, half actually were automated and half actually were operated by a person.

The people who believed no human was involved reported less fear of disclosing, managed their impression less, showed more sadness on their faces, and were rated by independent observers as more willing to disclose. Whether a person was actually behind the screen changed nothing except how usable they found the system.

239 participants

Believing the interviewer was automated increased disclosure. Whether it actually was made no difference.

Lucas, Gratch, King and Morency, Computers in Human Behavior, 2014

So the honest sentence is this: an AI interviewer does not remove bias by being clever. It removes bias by not being a person the participant has to face afterward.

What about the questions it asks?

That is the second half of the objection, and it is fair. A leading question biases an answer no matter who asks it.

Chopra and Haaland had an AI conduct 381 qualitative interviews and coded every question it asked. About 95 percent were open-ended, relevant, and non-leading, and the interview data predicted what participants actually did eight months later.

Jabarian and Henkel ran the largest test so far: 70,000 job applicants randomly assigned to a human recruiter or an AI voice agent, with humans making every hiring decision. Applicants interviewed by the AI were 12 percent more likely to receive an offer and more likely to stay in the job. The transcripts show why. The AI spoke less, asked more consistently, and collected more decision-relevant information. Given a choice, 78 percent of applicants picked the AI.

70,000 applicants

Randomised to an AI voice interviewer or a human recruiter. AI-interviewed applicants were 12 percent more likely to receive an offer and more likely to stay. Working paper, and the recruiting firm has a commercial interest.

Jabarian and Henkel, 2025

Where the evidence stops

Three boundaries, because this is the study a sceptic will read closest.

The effect is about the perceived absence of a human judge, not about screens. Dodou and de Winter found no difference in social desirability across paper, offline and online surveys. If a participant believes a person will read their answer with their name attached, the benefit shrinks.

Kreuter's automated voice was a set of recorded prompts, not a conversation, so it is a floor for what a voice interviewer can do rather than a description of one. What carries over to a conversational agent is Lucas's mechanism: the participant's belief that no person is judging. Anonymity on request strengthens that belief.

Candour is not the same as accuracy. A machine that people speak to more freely still cannot verify what they say. That is what the human review and the verbatim quotes are for.

What Candidly AI does with this

Participants are told plainly that they are speaking to an AI. Calls can be fully anonymous on request. The questions are written with you in the Design phase and approved before launch, so what gets asked is on the record before anyone answers. Every conversation is read by a person, and the Ground Truth Report puts the verbatim quotes beside the themes, every participant quoted at least once.

The bias a human interviewer introduces is the participant's own editing. The research says take the person out of the room and the editing drops. That is not a claim about what our software can do. It is a claim about what people do, and it has held for thirty years.

An AI interviewer does not remove bias by being clever. It removes bias by not being a person the participant has to face afterward.
Common questions

Do people disclose more to an AI interviewer than to a human?

The evidence says yes, when they believe no person is listening. Lucas, Gratch, King and Morency found that participants who believed a virtual interviewer was automated reported lower fear of disclosure and were rated by observers as more willing to disclose, and whether it was actually automated made no difference. Richman and colleagues' meta-analysis of 61 studies found less social desirability distortion in computerised interviews than face-to-face interviews.

What is social desirability bias?

The tendency to answer in the way that makes you look better to whoever is asking. Tourangeau and Yan describe it as a motivated process in which respondents edit what they report to avoid embarrassing themselves in front of an interviewer or to avoid consequences from third parties.

Does an automated voice interviewer reduce bias compared to a human on the phone?

Kreuter, Presser and Tourangeau randomly assigned people to a human telephone interviewer, an automated voice system, or a web form, and checked answers against university records. Reporting of sensitive information and accuracy were lowest with the human interviewer and highest on the web, with the automated voice in between.

Does an AI interviewer ask biased or leading questions?

Chopra and Haaland coded every question asked across 381 AI-led interviews and found about 95 percent were open-ended, relevant, and non-leading. Jabarian and Henkel's field experiment with 70,000 applicants found the AI interviewer more structured and consistent than human recruiters, and applicants interviewed by AI were 12 percent more likely to receive an offer.

What is Candidly AI?

Candidly AI is a Calgary-based Canadian platform that holds hundreds of one on one voice conversations with the people whose perspective a decision depends on. Participants call a dedicated number on their own schedule, the conversation adapts with follow-up questions, and the organisation receives the Ground Truth Report: themed findings backed by verbatim quotes, with every participant quoted at least once. Candidly AI is ISO 27001 Certified and SOC 2 Audited, runs on fully Canadian infrastructure with self-hosted models, and makes no third-party model API calls.

Design, Listen, Findings. Fully Canadian infrastructure, self-hosted models, no third-party model API calls. SOC 2 Audited. ISO 27001 Certified.

See how it works

Sources

We tier our sources so you can see which claims are load-bearing and which are supporting. Where a study is a working paper or has a commercial interest behind it, we say so.

Load-bearing

Lucas, G. M., Gratch, J., King, A., and Morency, L.-P. (2014). It's only a computer: Virtual humans increase willingness to disclose. Computers in Human Behavior, 37, 94-100. 239 participants, 2×2 design.
doi.org/10.1016/j.chb.2014.04.043Used for: belief that the interviewer was automated increased disclosure; actual automation changed only usability.

Tourangeau, R. and Yan, T. (2007). Sensitive questions in surveys. Psychological Bulletin, 133(5), 859-883.
doi.org/10.1037/0033-2909.133.5.859Used for: misreporting as motivated editing in the presence of an interviewer.

Richman, W. L., Kiesler, S., Weisband, S., and Drasgow, F. (1999). A meta-analytic study of social desirability distortion in computer-administered questionnaires, traditional questionnaires, and interviews. Journal of Applied Psychology, 84(5), 754-775. 61 studies, 673 effect sizes.
doi.org/10.1037/0021-9010.84.5.754Used for: less distortion in computerised vs face-to-face interviews; near-zero difference computer vs paper; least distortion when alone, anonymous, able to backtrack.

Kreuter, F., Presser, S., and Tourangeau, R. (2008). Social desirability bias in CATI, IVR, and Web surveys: The effects of mode and question sensitivity. Public Opinion Quarterly, 72(5), 847-865. Randomised, validated against university records.
doi.org/10.1093/poq/nfn063Used for: reporting and accuracy highest on web, lowest with human interviewer, automated voice intermediate.

Supporting

Chopra, F. and Haaland, I. (2023). Conducting Qualitative Interviews with AI. CESifo Working Paper No. 10666. 381 interviews.
ssrn.com/abstract=4583756Used for: about 95 percent of AI questions open-ended, relevant and non-leading; predictive validity at eight months.

Jabarian, B. and Henkel, L. (2025). Voice AI in Firms: A Natural Field Experiment on Automated Job Interviews. SSRN working paper. 70,000 applicants. Not yet peer-reviewed; conducted with a recruiting firm that sells the AI interviewer studied.
ssrn.com/abstract=5395709Used for: 12 percent more offers, higher retention, 78 percent choosing AI, AI interviews more structured and consistent.

Correctives we apply to our own sources

Dodou, D. and de Winter, J. C. F. (2014). Social desirability is the same in offline, online, and paper surveys: A meta-analysis. Computers in Human Behavior, 36, 487-495.
doi.org/10.1016/j.chb.2014.04.005The reason we do not claim the screen reduces bias. The effect is about the interviewer.

Weisband, S. and Kiesler, S. (1996). Self disclosure on computer forms: Meta-analysis and implications. CHI '96. 30 studies, d = 0.20, effect shrinking over time.Early evidence that the computer-vs-paper effect was small and fading, consistent with Richman and Dodou.

What would you ask your stakeholders?

Tell us the questions your programme is trying to answer and we will build a demo conversation around those, not a generic one.