OpenAI Jobs logo
Posted 4w ago•San Francisco

Researcher, Agent Safety, Oversight and System Mitigations

MiddleOn-site (San Francisco)Salary undisclosed
Job Description

About the Team

The Agent Safety team works to ensure that increasingly capable AI agents act safely, exercise sound judgment, and remain aligned with user intent. Our mission is to reduce the probability of severe unintended outcomes from increasingly capable AI agents while preserving their ability to act effectively and autonomously.

Our work spans three areas:

  • Training: Create training methods, environments and data that teach agents to make better decisions in consequential situations. We turn real-world failures into training signals that prevent similar incidents, and identify precursor behaviors and mitigations to address emerging risks.

  • Measurements: Build evaluations and production metrics that identify emerging risks and measure whether our interventions work.

  • Oversight: Develop oversight and system mitigation mechanisms that reduce harmful actions while preserving useful agent autonomy (for example future versions of auto-review).

About the Role

This role focuses on oversight and system-level mi

Similar Openings in Other

View all in category➔