Research Engineer, Audio and Speech

Added
8 days ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

python deep learning pytorch machine learning speech recognition
JobCopilot logo
Meet JobCopilot: Your Personal Al Job Hunter
Automatically Apply to Your Dream Jobs While You Sleep
Try it now β†’

πŸ“‹ Description

  • Design and build next-generation agent harnesses optimized for streaming speech, turn-taking
  • Research and train multimodal and full-duplex models that jointly understand audio, reason, and
  • Improve speech recognition, voice activity detection, endpointing, and speech generation across
  • Build evaluations and use production calls to ship measurable improvements in accuracy, latency
  • Optimize end-to-end inference for responsiveness, throughput, stability, and cost, partnering with

🎯 Requirements

  • 2+ years of experience in speech, audio ML, multimodal ML, or production machine learning
  • Experience developing or adapting autoregressive, diffusion, flow-matching, or codec-based speech
  • Hands-on experience with streaming agent systems, low-latency inference, production model serving
  • Fluency in Python and a modern deep-learning framework such as PyTorch, with strong foundations in
  • A track record of taking research ideas from prototype to reliable, measurable production impact
  • Familiarity with speech-to-speech or full-duplex models

🎁 Benefits

  • Medical, Dental, and Vision benefits for you and your family
  • Life Insurance and Disability Benefits
  • Retirement Plan (e.g., 401K, pension)
  • Parental Leave
  • Fertility and family building benefits through Carrot
  • Monthly stipend to support your wellness, lifestyle, and work-life balance
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’