Researcher, Agent Safety, Oversight and System Mitigations

Added
16 days ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

ai permissions evaluations sandboxing oversight

πŸ“‹ Description

  • Design, build, and evaluate system-level controls for agent actions like agent-based review
  • Plan how controls fit in a broader system including sandboxing with process isolation and
  • Work closely with a Codex harness engineering team to productionize the AI controls
  • Red-team end-to-end agentic systems to measure whether controls prevent data exfiltration, unsafe
  • Improve the safety-productivity tradeoff by measuring and reducing missed harmful actions

🎯 Requirements

  • Have strong systems or security instincts and can reason concretely about isolation boundaries
  • Enjoy turning ambiguous safety questions into concrete threat models, reproducible experiments, and
  • Can build robust experimental infrastructure and design evaluations that distinguish promising
  • Are deeply interested in frontier AI alignment, safety and control
  • A background in AI control or security is welcome but not required

🎁 Benefits

  • Offers Equity
  • $380K – $500K salary
  • Relocation assistance

🚚 Relocation support

Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’