OpenAI

Researcher, Frontier Risk Mitigations

OpenAI

London, England, United Kingdom · Full Time

Be the first to apply

Experience
2+ yrs
Salary
USD 295,000 – USD 445,000 / year
Openings
1
Posted
1 week ago
Work mode
In office
Education
Ph.D. in Computer Science or related discipline
Resume
Required to apply

Where you'll work

Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.

Job description

About The Team

The Preparedness team focuses on addressing significant risks associated with highly capable AI models that transition from assisting humans to independently planning and acting in real-world environments. Their efforts involve monitoring AI capabilities, ensuring safeguards against misuse and alignment failures, and collaborating across teams to maintain OpenAI’s preparedness framework.

Role Overview

This position is aimed at outstanding researchers dedicated to innovating safety mitigation strategies for frontier AI systems. Responsibilities include developing new safety techniques inspired by interpretability, control theory, and alignment research to ensure safe deployment of OpenAI’s models. The role involves defining the future shape of safe AI systems and critically contributing to OpenAI’s mission of building safe Artificial General Intelligence (AGI).

Responsibilities

  • Identify and analyze emerging AI safety risks, developing innovative approaches to evaluate and reduce their impact.
  • Create and continuously improve evaluation methods to assess risk levels, collaborating with domain experts as necessary.
  • Guide the research agenda and strategic approach towards enhancing AI safety, alignment, and robustness.
  • Participate in crafting industry and organizational best practices for AI safety standards.
  • Design and review red-teaming processes to rigorously test safety systems and highlight potential improvements.

Ideal Candidate Attributes

  • Strong alignment with OpenAI's mission to develop safe and beneficial AGI.
  • Passion for long-term AI safety and deep understanding of technical pathways toward safe AGI.
  • Practical experience applying interpretability, robustness, alignment, and control techniques to AI systems.
  • Minimum of 2 years applied experience in AI safety domains including reinforcement learning from human feedback (RLHF), human-AI collaboration, or control.
  • Ph.D. or an advanced degree in computer science, machine learning, or related disciplines.
  • Ability to work with large-scale AI architectures effectively.
  • At least 4 years of research engineering experience with proficiency in Python or comparable programming languages.

Equal Opportunity and Additional Information

OpenAI is dedicated to fostering diversity and inclusion, ensuring equal opportunity regardless of background or protected characteristics. Background checks are conducted in compliance with applicable laws. Reasonable accommodations are available for applicants with disabilities upon request.

Compensation

The salary range for this role is between $295,000 and $445,000 per year.

Minimum education

Doctorate

How they work

Teamwork & Collaboration Problem Solving Attention to Detail Adaptability
🤖
Online · instant AI help
Broxer