Staff Machine Learning Engineer, Voice AI

Added
18 minutes ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

python pytorch cuda gpu vllm
JobCopilot logo
Meet JobCopilot: Your Personal Al Job Hunter
Automatically Apply to Your Dream Jobs While You Sleep
Try it now β†’

πŸ“‹ Description

  • Own the voice inference roadmap end-to-end for STT, TTS, and speech-to-speech.
  • Drive best-in-class inference performance for voice workloads (latency and throughput).
  • Lead productionization of voice models at scale with serverless and dedicated endpoints.
  • Build the voice evaluation platform for WER, latency, and naturalness across accents and noise.
  • Shape architecture for next-gen model support (audio-native LLMs, SNAC/Encodec).
  • Serve as technical DRI for model partner integrations (Cartesia, Deepgram, Rime).

🎯 Requirements

  • 8+ years of ML engineering experience focused on model serving at production scale.
  • Deep expertise in LLM serving engines (vLLM, SGLang, TensorRT-LLM) or equivalent.
  • Expert-level Python and PyTorch; strong GPU optimization (CUDA, profiling).
  • Proven system design judgment with scalable architectural decisions.
  • Strong technical leadership; high autonomy and engineering quality.
  • Familiarity with audio codecs/tokenization (SNAC/Encodec/DAC) is a plus.

🎁 Benefits

  • Startup equity
  • Health insurance
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’