Member of Technical Staff - RL Inference

Added
1 day ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

rust python pytorch jax cuda

πŸ“‹ Description

  • Design and optimize our inference stack for RL workloads at SpaceXAI, from small scale ablations to
  • Analyze, profile and address performance bottlenecks in large scale RL systems.
  • Work closely with the modelling team to efficiently implement novel RL techniques and algorithms.

🎯 Requirements

  • Experience in building, debugging, and optimizing efficiency of large-scale distributed systems.
  • Experience in LLM inference.
  • Proficiency in Python, C++ and/or Rust; frameworks PyTorch, Jax, CUDA.
  • Willingness to dive deep and solve hardcore problems at all levels of the stack.
  • Strong knowledge in quantization and numerics in LLM inference and training (preferred).
  • Experience in developing inference engines, e.g. SGLang, vLLM (preferred).

🎁 Benefits

  • Equity, comprehensive medical, vision, and dental coverage.
  • 401(k) retirement plan access.
  • Short and long-term disability insurance.
  • Life insurance and other discounts and perks.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’