Research Engineer (Reinforcement Learning)

Added
5 days ago
Type
Full time
Salary
Salary not provided

Related skills

python reinforcement learning gpus lora vllm

๐Ÿ“‹ Description

  • Build training environments, verifiers, and supporting infrastructure for post-training models.
  • Own the synthetic data pipeline from data generation through quality assurance and validation.
  • Run end-to-end training experiments, analyze results, and identify factors driving model
  • Design and maintain evaluations that models must pass before production releases.
  • Select and adapt open-weight foundation models for agent/product requirements.
  • Develop trained behaviors for voice and text-based agents.

๐ŸŽฏ Requirements

  • Strong Python engineering skills and ability to build reliable, production-quality systems.
  • Experience taking an ML model from raw data through experimentation to production.
  • Data-centric mindset with attention to coverage, quality, and data leakage.
  • Ability to anticipate reward exploitation and design robust rewards, verifiers, and evaluation
  • Practical experience with GPUs and understanding of their capabilities/limitations.
  • Good judgment on when training is the right solution versus simpler approaches.

๐ŸŽ Benefits

  • Opportunity to make a significant impact on a fast-growing developer platform.
  • Collaboration with a small, experienced team valuing excellence, creativity, and ownership.
  • Competitive salary and equity package.
  • Health, dental, and vision benefits.
  • Flexible vacation policy.
  • Remote-friendly working environment with autonomy.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest โ€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs โ†’