Senior Machine Learning Engineer, Runtime and Serving

Added
1 day ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

python pytorch jax cuda tensorrt

πŸ“‹ Description

  • Architect and develop an efficient, high-performance ML runtime and serving system tailored for
  • Lead integration and feature development for ML inference runtimes across domains, balancing
  • Drive migration of ML workloads toward a JAX-native runtime architecture (OpenXLA/PjRT, TensorRT
  • Collaborate with Waymo ML teams to analyze system-level ML workloads and apply hardware-aware
  • Design robust tooling for profiling, benchmarking, and bottleneck identification across the ML

🎯 Requirements

  • B.S. or M.S. in CS, EE, Deep Learning or related field
  • 5+ years software engineering experience building, scaling, or maintaining ML systems
  • 5+ years production programming in C++
  • 3+ years production experience in Python and major DL frameworks (e.g., PyTorch, JAX)
  • Experience optimizing ML software for hardware accelerators (GPUs, TPUs, custom silicon)
  • Experience building low-latency, highly concurrent distributed backend systems

🎁 Benefits

  • The salary range is $213,000 - $263,000 USD, with bonus/equity benefits as described by Waymo
  • Waymo employees are eligible for discretionary annual bonus, equity incentive plan, and benefits
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’