The Effective Altruism
Opportunities Board
Work on the world's most pressing problems. Browse jobs, fellowships, internships, courses, and more at high-impact organisations.
Researcher, Recursive Self-Improvement Safety
OpenAISan Francisco, CA
San Francisco, CA
Today
Routes to impact
Direct high impact on an important cause
Description
Conduct technical research and build safety measures for advanced AI systems, focusing on risks from recursive self-improvement and loss of control.
- Develop scalable oversight, automated auditing, and model behavior evaluations
- Stress-test monitoring systems for scheming and other misalignment risks
- Build prototypes and integrate successful approaches into safety pipelines
- Coordinate technical safety work across teams and external communications
This text was generated by AI. If you notice any inconsistencies, please let us know using this form.
Related opportunities
Researcher, Alignment Chain of Thought Monitorability
OpenAISan Francisco, CA
San Francisco, CA
5 days ago
Founding Member, Technical Staff
NeolithicSan Francisco, CA
San Francisco, CA
Today
Research Scientist
GoodfireLondon, UK
London, UK
3 days ago
AI Security Researcher
MicrosoftRedmond, WA
Redmond, WA
3 days ago
Automation Lead
Alignment Research CenterBerkeley, CA
Berkeley, CA
5 days ago
AI Security Researcher
Carnegie Mellon UniversityPittsburgh, PA
Pittsburgh, PA
6 days ago
Assistant AI Security Software Engineer
Carnegie Mellon UniversityPittsburgh, PA
Pittsburgh, PA
6 days ago
Computer Scientist, AI Test and Evaluation
Defense Threat Reduction Agency Fort Belvoir, VA
Fort Belvoir, VA
6 days ago
Join 60k subscribers and sign up for the EA Newsletter, a monthly email with the latest ideas and opportunities