Audio AI Engineer Job at Propio, Overland Park, KS

eHlwb2d3QUlVZmo3b1dkU1V4eE5uUEt2QVE9PQ==
  • Propio
  • Overland Park, KS

Job Description



Full-time

Description

Propio is on a mission to make communication accessible to everyone. As a leader in real-time interpretation and multilingual language services, we connect people with the information they need across language, culture, and modality. We’re committed to building AI-powered tools to enhance interpreter workflows, automate multilingual insights, and scale communication quality across industries.

We are hiring an Audio AI Engineer that will develop and optimize end-to-end systems that enable real-time, high-fidelity speech-to-speech interpretation at Propio. This role focuses on seamlessly connecting speech recognition, translation, and synthesis technologies to create natural, low-latency interpretation experiences. 

Key Responsibilities:

  • Design and optimize end-to-end Speech-to-Speech pipelines that integrate ASR, translation, and TTS with minimal latency
  • Build bidirectional interpretation systems that handle turn-taking, speaker identification, and context preservation across language boundaries 
  • Collaborate with the Audio/Speech Engineer to optimize latency, quality, and robustness of speech components in the full pipeline
  • Work with the Staff ML Engineer to design efficient inference architectures and deployment strategies for real-time streaming systems
  • Develop streaming ASR and TTS systems capable of handling continuous, overlapping speech in interpretation scenarios
  • Benchmark and optimize latency across all pipeline stages (speech capture, recognition, translation, synthesis)
  • Integrate speaker diarization, acoustic environment adaptation, and speech enhancement into interpretation workflows
  • Partner with linguists and product teams to validate interpretation quality and gather domain-specific feedback

Requirements

Qualifications:

  • Bachelor's or Master’s Degree in Electrical Engineering, Computer Science, or related field
  • 3+ years of experience in speech processing, audio engineering, or conversational AI systems
  • Deep expertise in ASR, TTS, and streaming audio architectures
  • Proficiency in Python, ML frameworks, and experience with real-time signal processing 
  • Experience building low-latency production systems and optimizing for inference performance
  • Strong understanding of interpretation workflows, multilingual challenges, and speech quality metrics

Preferred Qualifications:

  • Experience building speech-to-text pipelines or hybrid ASR + LLM systems
  • Familiarity with real-time audio processing or latency-sensitive applications

#LI-JS1

Job Tags

Full time,

Similar Jobs

Medical On Demand

Primary Care Physicians, A Better Work/Life Position is in San Antonio Tx Job at Medical On Demand

Are you looking for a No Weekends-No Holidays Managed Care Opportunity We have a Manage Care Group providing high quality...  ...care in Beautiful San Antonio, Texas. They"re looking for a Primary Care Physician to see Medicare and Medicare Advantage patients. They offer... 

Sitter.com

Sitter Wanted - Looking For Children To Watch In My Home In Dauphin Pa Job at Sitter.com

Hi my name is Mary. I am a stay at home mom and have 3 children of my own. I amlooking to watch children in my home. My hours are Monday through Friday from7am-5pm. I have reasonable rates. I am also CPR and First Aid certified and I also have my CDA in child care. If...