DATA SCIENCE
Voice & Speech AI Engineer
Build the voice layer of enterprise AI: real-time transcription, speech-to-speech agents and contact-centre systems where latency, accuracy and tone all decide whether customers stay on the line.
- Dallas
- Data Science
- Full-time
What you’ll do
- Design speech-to-speech agent pipelines that hold conversation-grade latency budgets
- Tune ASR and TTS for client domains: jargon, accents, acoustic conditions
- Build analytics on voice interactions — intent, sentiment, compliance phrases
- Integrate voice AI with client contact-centre and CRM systems
- Evaluate the fast-moving speech-model market and keep recommendations current
What we’re looking for
- MS or PhD in Computer Science, Linguistics, or equivalent experience
- Hands-on with modern speech stacks — Whisper, Deepgram, AssemblyAI — and TTS (ElevenLabs, cloud-native voices)
- Experience building real-time voice agents with streaming LLM pipelines and barge-in handling
- Strong Python and audio processing fundamentals (VAD, diarisation, noise robustness)
- Understanding of telephony integration (SIP, WebRTC, contact-centre platforms)
What we offer
- Competitive compensation package
- Health, dental, and vision insurance
- Professional development budget
- Flexible work arrangements
- Creative work environment