Skip to content
RoleRasta
Back to results
Confirmed Pakistan eligible
RemoteFull-time

Senior Voice AI Engineer

Employer

Sign up to view employer

remotePosted 2 Oct 2026Last verified 8 Oct 2026

Why this is on RoleRasta

Verified 8 Oct 2026

The posting does not contain a country restriction that excludes Pakistan.

Confirm final eligibility and employment terms with the employer before applying.

Job description

ABOUT THE ROLE As a founding engineer on a small conversational AI team, you will own the real-time voice layer, from incoming speech through AI reasoning to spoken responses. You will help make natural, responsive voice interactions work reliably in production, with a focus on end-to-end latency. WHAT YOU'LL DO - Build and own streaming speech-to-text, LLM turn-taking, text-to-speech, and telephony or WebRTC transport. - Measure and reduce latency, targeting first audio under 800 milliseconds on real calls. - Address interruptions, barge-in, silence detection, overlapping speech, poor audio, accents, and mid-sentence changes. - Build an evaluation harness from recorded calls, transcripts, and scored turns to detect regressions and guide product decisions. - Compare voice providers and models through evidence-based testing, and make changes based on results. - Instrument production systems for turn latency, transcription confidence, drop-offs, and cost per minute. - Work directly with founders and make technical decisions in a fast-moving team. WHAT WE'RE LOOKING FOR - At least 5 years building production software, including 2 or more years shipping voice, speech, or real-time audio systems. - Experience building and shipping end-to-end real-time voice pipelines, including streaming speech recognition, LLM turn-taking, speech synthesis, and telephony or…

Sign up to view full details

Create your account to continue to the complete posting and employer details.

Sign up to view full details