Voice, beta
Shipping todayVoice screens nobody acts on until a person has read them
An AI voice can run a consistent first-stage screen at nine in the evening. What it cannot do is decide anything, and we have built it so it does not.
AI interview calls run a fixed set of screening questions by voice and produce a transcript, a summary and structured answers on the candidate record. The capability ships as a controlled beta on the Professional tier, behind a feature flag, enabled only in one-party-consent jurisdictions, with a spoken recording disclosure before recording begins.
By Surhires Editorial · Published · Reviewed
In the product: AiInterviewCallsSection.tsx, InterviewAiSessionPanel.tsx and StartAiInterviewDialog.tsx
The release posture, stated plainly
This is the part most vendors bury, so it goes first. AI interview calls are a controlled beta at Wave 1. The capability is available on the Professional tier, it sits behind a feature flag rather than being switched on for every account, and it is enabled only in jurisdictions that operate one-party consent for call recording. Where a jurisdiction requires all parties to consent, we disable recording rather than assume consent has been given.
Every call opens with a spoken disclosure that the conversation is being recorded and transcribed, played before recording starts. Retention of the recording and the transcript is configurable per tenant. Every transcript is reviewed by a person before it influences anything. An administrator can disable the entire capability for the tenant with one switch, and when it is off, no call can be initiated by any user in the account.
Two conditions sit underneath all of that. Calls only run once the account's AI Knowledge Base is set up, meaning a company summary, a value proposition and at least one product or service; until it is, calls are skipped rather than placed with nothing to say about who is calling. And a candidate on the blacklist is always skipped, whatever their score, so automation never reopens contact that a person closed on purpose.
None of the above is a statement that using this feature makes you compliant with any statute. It is a description of what the product does and what controls it gives you. Whether a particular call in a particular jurisdiction is lawful depends on how you configure it, who you call and what you have captured from them, and that assessment is yours.
- Controlled beta at Wave 1, not a general-availability feature
- Calls run only once the AI Knowledge Base is set up; until then they are skipped
- Blacklisted candidates are never called
- Professional tier, behind a feature flag rather than on by default
- Enabled only in one-party-consent jurisdictions
- Spoken recording disclosure played before recording begins
- Configurable retention for recordings and transcripts
- Human review of every transcript before it affects a candidate
- A single administrator switch that disables the capability entirely
What the call covers, and why it is the same every time
The question set is authored by you, attached to the requisition stage, and fixed for the duration of the search. Every candidate at that stage hears the same questions in the same order, which is exactly what a first-stage screen should be and rarely is when four recruiters run it by phone across a fortnight.
Answers are captured as structured fields against the questions that produced them, not only as prose. That means an availability answer lands in the availability field and a notice period answer lands in the notice period field, so the record is directly comparable across a shortlist and can be filtered without anyone re-reading twenty transcripts.
The candidate can decline, and declining costs them nothing
A candidate who does not want to speak to an artificial voice can say so, and the call ends. The record shows that the candidate opted for a human conversation, and a task is raised for the recruiter to make that call. Declining is not a disposition, it does not affect a score, and it does not move the candidate anywhere in the pipeline.
This is a substantive design position rather than a courtesy. A screening step that penalises the people who refuse it stops being optional in any meaningful sense, and it collects a signal about a candidate that has nothing to do with whether they can do the job.
A person reads every transcript before anything moves
When a call completes, the session panel presents the recording, the transcript, the structured answers and a summary side by side. Nothing has advanced. The candidate is at the stage they were at before the call, and the transcript is a document awaiting review.
A recruiter reads it, corrects anything the transcription got wrong, and then makes the stage decision themselves with a coded disposition. The review is recorded against the session, so the record shows both the machine output and the human who acted on it. There is no configuration that lets a call advance or dispose of a candidate on its own.
Retention is a setting, and deletion removes the audio too
Recordings, transcripts and derived summaries are three artefacts, and a retention policy that deletes one and keeps the others is the failure most systems ship with. Retention here is configured per tenant, applies to all three together, and can be set shorter for the audio than for the structured answers if that is the balance you want.
When a candidate exercises an erasure request, the erasure workflow removes the audio file, the transcript, the summary and the structured answers as one operation, and records that the request was fulfilled and when. Where a placement has already been made and financial records must legally be retained, the workflow says so rather than quietly keeping the interview material as well.
What we do not let the model do
The system does not infer personality, emotion, confidence or cultural fit from a voice. It does not score an accent, a speech pattern or a hesitation. It does not produce an employability rating. Those capabilities are technically available and we have chosen not to expose them, because they are not measuring what they claim to measure and because they carry obvious discrimination risk.
What is produced is a transcript, structured answers to the questions you wrote, and a summary of what was said. That is the honest boundary of what a voice screen can contribute, and it is already worth having when the alternative is four hundred unreturned calls.
- No emotion, sentiment or confidence inference from audio
- No accent, fluency or speech-pattern scoring
- No automated advance, rejection or ranking from a call
- No use of the recording as training data outside your tenant
Where the beta is going
Wave 1 is a narrow, well-instrumented release with a small number of accounts running real volume, because a voice product needs to be tested against genuine phone lines, background noise and people who talk over the prompt rather than against a demo script.
The changes we expect ahead of general availability are broader language coverage, better handling of interruption and clarification, all-party-consent jurisdiction support once the consent capture flow around it is properly built, and a published operation weight for call minutes. Progress is posted on the shipping page rather than announced when it is finished.
What you get
Feature-flagged beta
Off by default, enabled per account on the Professional tier during Wave 1.
Jurisdiction gating
Callable only where one-party consent applies; elsewhere recording is disabled rather than assumed.
Spoken disclosure
A recording and transcription notice is played before recording starts on every call.
Authored question sets
You write the questions and attach them to a requisition stage; every candidate hears the same set.
Structured answers
Responses land in the fields they belong to, not only in a block of transcript text.
Session panel
Recording, transcript, structured answers and summary reviewable together on one screen.
Mandatory human review
No call can advance, reject or rank a candidate; a person makes the stage decision.
Decline path
A candidate can ask for a human instead, with no penalty and no disposition recorded.
Configurable retention
Per-tenant retention across audio, transcript and summary, settable independently.
Unified erasure
An erasure request removes audio, transcript, summary and answers as one operation.
Knowledge Base prerequisite
Calls are skipped until the company summary, value proposition and a product are set up.
Blacklist respected
A blacklisted candidate is always skipped by screening calls, whatever the score.
Admin kill switch
One tenant-level switch disables the capability for every user in the account.
No affect inference
Emotion, confidence, accent and fluency scoring are not built and will not be exposed.
Call audit trail
Who initiated the call, when, under which question set, and who reviewed the result.
Questions recruiters ask
Is this generally available?
No. It is a controlled beta at Wave 1, on the Professional tier, behind a feature flag, and enabled account by account in one-party-consent jurisdictions only. If you buy Professional today you should not assume access to it; ask during onboarding whether your account and your jurisdiction are in scope.
We switched it on and no calls went out. Why?
The most common reason is that the AI Knowledge Base is not set up yet. Calls need a company summary, a value proposition and at least one product or service so the voice can say who it is calling for; until those exist, calls are skipped rather than placed. Blacklisted candidates are also always skipped, whatever their score.
Does the candidate know they are speaking to an AI?
Yes. The call opens by identifying the voice as artificial and naming the company on whose behalf it is calling, and a spoken disclosure that the call is recorded and transcribed is played before recording begins. A candidate who would rather speak to a person says so and the call ends there.
Are you telling us this is legal where we operate?
No, and be wary of any vendor who does. We describe capabilities: jurisdiction gating, spoken disclosure, configurable retention, human review and a kill switch. Whether a specific call is lawful depends on your jurisdiction, your consent capture and how you configure the feature. That assessment belongs to you and your counsel.
Can a call reject a candidate?
No. A completed call produces a transcript and structured answers and leaves the candidate exactly where they were. A recruiter reads the transcript and makes the stage decision with a coded disposition under their own name. There is no setting that permits a call to advance or dispose of anyone automatically.
How long are recordings kept?
For as long as you configure, per tenant, with audio retention settable shorter than transcript retention if you want the substance without the voice. Deletion removes audio, transcript, summary and structured answers together, because a retention policy that clears one and keeps the rest is not a retention policy.
What happens if we switch it off after using it?
The administrator switch stops any further calls immediately for every user in the account. Existing sessions remain on the candidate records they belong to and follow your retention setting; they are not deleted by the switch. If you want them gone as well, run the erasure workflow against those records.
How is this different from the AI calling agent?
AI interview calls are the interview session itself: your question set, run by voice, reviewed by a person. The AI calling agent is the outbound qualification call that dials a candidate, works through availability and eligibility, and books the interview at the end. They share the same voice infrastructure and the same release posture.
Keep reading
- Outbound screening calls that end with an interview booked
- Book the interview in one message, not eleven
- Give every interviewer the same brief
- Structured feedback instead of an opinion
- Know why you hold every candidate record, and for how long
- What the product supports, and what the sender still owns
- Every feature area, with its real build status
See it against your own reqs
Bring one live role and three resumes. In twenty minutes you will see the match scores, the shortlist and the placement invoice that comes out the other end.