Skip to content
DATA SCIENCE

Voice & Speech AI Engineer

Build the voice layer of enterprise AI: real-time transcription, speech-to-speech agents and contact-centre systems where latency, accuracy and tone all decide whether customers stay on the line.

  • Dallas
  • Data Science
  • Full-time

What you’ll do

  • Design speech-to-speech agent pipelines that hold conversation-grade latency budgets
  • Tune ASR and TTS for client domains: jargon, accents, acoustic conditions
  • Build analytics on voice interactions — intent, sentiment, compliance phrases
  • Integrate voice AI with client contact-centre and CRM systems
  • Evaluate the fast-moving speech-model market and keep recommendations current

What we’re looking for

  • MS or PhD in Computer Science, Linguistics, or equivalent experience
  • Hands-on with modern speech stacks — Whisper, Deepgram, AssemblyAI — and TTS (ElevenLabs, cloud-native voices)
  • Experience building real-time voice agents with streaming LLM pipelines and barge-in handling
  • Strong Python and audio processing fundamentals (VAD, diarisation, noise robustness)
  • Understanding of telephony integration (SIP, WebRTC, contact-centre platforms)

What we offer

  • Competitive compensation package
  • Health, dental, and vision insurance
  • Professional development budget
  • Flexible work arrangements
  • Creative work environment
Back to job search