OpenAI

Researcher, Frontier Risk Mitigations

OpenAI$150K — $180K *
Enterprise Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • 2+ years of experience in AI safety, particularly in RLHF, human-AI collaboration, interpretability, or control
  • Ph.D. or degree in computer science, machine learning, or a related field
  • Over 4 years of research engineering experience
  • Proficiency in Python or similar languages
  • Ability to thrive in large-scale AI environments

Responsibilities

  • Identify emerging AI safety risks and develop new methodologies for mitigation
  • Build and refine evaluations to assess safety risks, collaborating with domain experts
  • Set research directions to enhance AI system safety and alignment
  • Contribute to AI safety best practices for OpenAI and the broader industry
  • Evaluate and design red-teaming pipelines for robustness of safety systems

Benefits

  • Opportunity to work on cutting-edge AI safety initiatives
  • Collaborative environment with cross-functional teams
  • Direct impact on the future of safe AGI development
  • Access to a network of domain experts across various fields
  • Involvement in shaping industry standards for AI safety
Full Job Description
About the Role

We are seeking exceptional researchers who can push the frontier of safety mitigations. You will help derisk frontier models by developing novel safety mitigations, developing and applying new techniques from domains like interpretability, control, and alignment to ensure the safety of OpenAI's deployed models. You will play a critical role in defining how a safe AI system should look in the future at OpenAI, making a significant impact on our mission to build and deploy safe AGI.

This role requires strong technical depth and close cross-functional collaboration to ensure our safety mitigations are enforceable, scalable, and effective. We seek researchers who can partner with experts across domains such as misalignment, cybersecurity, and biology in order to develop the best possible end-to-end safety stack.

In this role, you will:
  • Work on identifying emerging AI safety risks and new methodologies for exploring and mitigating the impact of such risks
  • Build (and then continuously refine) the evaluations that enable us to assess the extent of these risks; this might include working with domain experts (be it internal or external)
  • Set research directions and strategies to make our AI systems safer, more aligned, and more robust
  • Contribute to the development of "best practices" guidelines for AI safety for OpenAI and across the industry
  • Evaluate and design effective red-teaming pipelines to examine the end-to-end robustness of our safety systems, and identify areas for future improvement.
You might thrive in this role if you:
  • Are excited about OpenAI's mission of building safe, universally beneficial AGI and are aligned with OpenAI's charter.
  • Show enthusiasm for long-term AI safety, and have thought deeply about technical paths to safe AGI.
  • Are willing to "get your hands dirty" and apply methods from domains such as interpretability, robustness, alignment, and control to make OpenAI's models safe.
  • Bring 2+ years of experience in the field of AI safety, especially in areas like RLHF, human-AI collaboration, interpretability, or control.
  • Hold a Ph.D. or other degree in computer science, machine learning, or a related field.
  • Thrive in environments involving large-scale AI systems.
  • Possess 4+ years of research engineering experience and proficiency in Python or similar languages.


About OpenAI

OpenAI is an artificial intelligence research laboratory consisting of the for-profit corporation OpenAI LP and its parent company, the non-profit OpenAI Inc. The company was founded in 2015 by a group of technology leaders, including Elon Musk, Sam Altman, Greg Brockman, Ilya Sutskever, and John Schulman. OpenAI's mission is to develop and promote friendly AI for the betterment of humanity. The company has developed a number of cutting-edge AI technologies, including GPT-3, a language processing system that can generate human-like text. OpenAI has received funding from a number of high-profile investors, including LinkedIn co-founder Reid Hoffman and venture capitalist Peter Thiel.
Learn more about OpenAI
Size
100 employees
Industry
Founded
2015

Similar Jobs

More Jobs at OpenAI

More Enterprise Technology Jobs

Find similar Researcher, Frontier Risk Mitigations jobs: