Research Engineer (Reinforcement Learning)

Added
5 days ago
Type
Full time
Salary
Salary not provided

Related skills

python reinforcement learning rl vllm post-training

๐Ÿ“‹ Description

  • Build environments, verifiers, and infra for post-training models.
  • Own synthetic data pipeline from generation to QA.
  • Run end-to-end training experiments and analyze results.
  • Design evaluations for production releases.
  • Adapt open-weight foundation models for agents/products.
  • Develop robust multi-turn agent behaviors (voice/text).

๐ŸŽฏ Requirements

  • Strong Python engineering for production systems.
  • Experience taking ML models from data to production.
  • Data-centric mindset: coverage, quality, leakage awareness.
  • Design robust rewards, verifiers, and eval mechanisms.
  • GPU experience; understanding of their limits.
  • Judgment on when to train vs simpler approaches.

๐ŸŽ Benefits

  • Impact on a fast-growing developer platform.
  • Small, experienced team with ownership.
  • Competitive salary and equity.
  • Health, dental, vision benefits.
  • Flexible vacation policy.
  • Remote-friendly environment with autonomy.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest โ€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs โ†’