Home OverviewVoice Interview Features Trust All productsEight products, one data layer
Get in touch See a transcript

Products › Voice Interview

No camera. No app.
Just a conversation.

A spoken round, transcribed as it happens

A voice round any candidate can take on the phone they already have. No camera, no download, barely any bandwidth — transcribed live and scored against the same rubric as every other round.

Works on a phone, on a weak connection, in a shared room.

Demo video goes here Drop the file at assets/voice-interviewer-demo.mp4 — 16:9, 1920×1080, H.264. Autoplays muted on loop; keep it under roughly 8 MB and 30 seconds.

A spoken round, transcribed live, scored against the rubric.

The camera is the
part that loses you people.

Some candidates have no quiet room, no good light and no confidence on camera — and none of that tells you whether they can do the job. A voice round removes the barrier and keeps the assessment.

  • The lowest barrier of the three rounds. No camera, no video file to upload, and far less bandwidth than a recorded round — so a candidate on mobile data can actually finish it.
  • For roles where speaking is the work. Sales, support, hospitality, collections, anything client-facing, and any role with a real language-fluency requirement. A voice round tests the skill rather than a proxy for it.
  • A speaker-attributed transcript with timestamps. Which is what makes a spoken answer reviewable at all — read it in two minutes instead of listening for twenty, and jump to any moment.
  • Tone and confidence, from the voice. Reported as one input among several, never as a verdict, and never as a reason to pass on somebody by itself.
  • The same rubric as video and chat. A voice score and a video score sit on one axis, so mixing modalities across a pipeline does not break the comparison.
Transcript — customer success, round 1
Q01:04 · Tell me about a time you had to say no to a customer.
NA01:11 · “They wanted the integration live before the contract was signed. I said no, but I got them a sandbox the same day so they weren’t blocked…”
Q02:20 · What did their team say?Adaptive
Objection handling4.3
Clarity of speech4.5
Composure3.8
Objection handling 4.3 links to 01:11 · the follow-up at 02:20 was not scripted

A phone, five minutes, a scored round.

Five stages, and only one of them needs a person. The invitation sends itself and the transcript is written as the candidate talks.

Invite sent
automatic
Mic check
30 sec
They speak
their time
Transcribed
as they talk
You decide

What the Voice Interview does

Why a voice round reaches more candidates+
Three barriers disappear at once. There is no camera, so nobody needs a presentable room or good light. There is no video upload, so a weak mobile connection is enough. And there is nothing to install — it runs in the browser they already have open. For high-volume, frontline and shift-based hiring, that difference shows up directly in how many people finish the round.
The transcript, and what it is actually for+
A spoken answer is hard to review, because you cannot skim audio. So every round produces a speaker-attributed transcript, timestamped to the second, which turns a twenty-minute recording into two minutes of reading with the ability to jump to any moment. That is what makes it practical for a panel rather than just for one recruiter.
Question sets, and follow-ups that were not scripted+
Reusable question sets grouped into sections, with the interview adapting as it goes — so a thin answer draws a follow-up and a complete one does not. Question generation runs in the background, so opening a new role does not wait on somebody writing a bank from scratch.
Scoring, with the excerpt attached+
Answers are scored against your rubric by a model tier reserved for scoring, validated against a fixed schema rather than accepted as free text. Each score links to the passage of transcript it came from, so a rating always points at something the candidate actually said. It is the same rubric the video and chat rounds use, so scores stay comparable across modalities.
Tone and confidence, handled with care+
Emotional tone and confidence are read from the voice and reported as structured signal. It is deliberately one input among several: an accent, a nervous first minute or a bad line are not competence, and this is the wrong page to pretend otherwise. A human makes every call.
Pipelines, leaderboards and evaluation+
Candidates move through pipelines with weighted scorecards and leaderboards, so evaluation lives in one place. Hiring managers assess the same recorded round rather than a summary of it, and transcripts and analytics are retained for reporting on the process.

A recording of somebody’s
career

Audio and transcripts are the most sensitive data you will hold about a person who does not yet work for you.

A human makes the call

It ranks and evidences. It does not reject anyone. The hire, the pass and the pipeline stage are decisions a named person takes, on the record.

One AI gateway, no loose ends

Every model call passes through a single choke point with schema validation, caching, per-tenant rate limits and spend caps. There is no second path to a provider.

Media behind signed links

Audio sits in private storage reached only through short-lived signed URLs, with lifecycle rules that expire recordings on a schedule rather than keeping them indefinitely.

And plainly. A voice round still needs a microphone and a connection, which is why the check runs first — and the Chat round exists for candidates who have neither, or who assess better in writing. Tone and confidence signal is an input and should never on its own be the reason somebody is passed over; a strong accent and a weak line are the two easiest things in this field to mistake for a weak candidate. We publish no accuracy figure, because the honest measure is how often a human changed a score in your own pipeline. And there is no telephone dial-out — the round runs in a browser, not over a phone call.

Frequently asked questions

Do candidates need a camera, an app or a strong internet connection to take the interview?+
No. There is no camera, so nobody needs a presentable room or good light; there is no video upload, so a weak mobile connection is enough; and there is nothing to install, because it runs in the browser they already have open.
How are spoken answers scored, and can I see why a score was given?+
Answers are scored against your own rubric by a model tier reserved for scoring, and the output is validated against a fixed schema rather than accepted as free text. Each score links to the passage of transcript it came from, so every rating points at something the candidate actually said.
Does the AI judge candidates on their tone of voice or accent?+
Emotional tone and confidence are read from the voice and reported as structured signal, but deliberately as one input among several: an accent, a nervous first minute or a bad line are not competence. A human makes every call.
Can I use voice interviews alongside video or chat interviews in the same hiring pipeline?+
Yes. The voice round uses the same rubric as the video and chat rounds, so a voice score and a video score sit on one axis and mixing formats across a pipeline does not break the comparison.
When is a voice interview the wrong format for a candidate?+
When they have no microphone or no connection, which is why a check runs before the round starts. For those candidates, or for anyone who assesses better in writing, the Chat Interview round is the alternative.
How accurate is the AI scoring in the voice interview?+
TalbotIQ publishes no accuracy figure. Its stated reason is that the honest measure is how often a human changed a score in your own pipeline.

Published 9 September 2026 · Updated 9 September 2026 · Written by the TalbotIQ team

Ready to see it?

Tell us the role and where your candidates actually are — on a laptop, on mobile data, on a shop floor — and we will run a round against it.

0 / 250
Request demo

Take the camera out of it.

Book a 30-minute demo. Bring a role you are struggling to fill and we will run a voice round against it live.

Malaysia-based team · Response within one business day · No obligation