The Effective Altruism
Opportunities Board
Work on the world's most pressing problems. Browse jobs, fellowships, internships, courses, and more at high-impact organisations.
Researcher, Recursive Self-Improvement Safety
OpenAISan Francisco, CA
San Francisco, CA
1 month ago
Routes to impact
Direct high impact on an important cause
Description
Conduct technical research and build safety measures for advanced AI systems, focusing on risks from recursive self-improvement and loss of control.
- Develop scalable oversight, automated auditing, and model behavior evaluations
- Stress-test monitoring systems for scheming and other misalignment risks
- Build prototypes and integrate successful approaches into safety pipelines
- Coordinate technical safety work across teams and external communications
This text was generated by AI. If you notice any inconsistencies, please let us know using this form.
Related opportunities
Researcher, Alignment Interpretability
OpenAISan Francisco, CA
San Francisco, CA
3 weeks ago
Researcher, Alignment Chain of Thought Monitorability
OpenAISan Francisco, CA
San Francisco, CA
1 month ago
Research Engineer / Research Scientist, Zero-Knowledge Verification
Singapore AI Safety Hub (SASH)Remote (UK / Singapore)
Remote (UK / Singapore)
1 day ago
Research Engineer, Responsible Frontier AI Research
Google DeepMindLondon, UK
London, UK
1 day ago
Member of Technical Staff, Network Engineer, Software
Lucid ComputingSan Francisco, US / London, UK
San Francisco, US / London, UK
3 days ago
Software Engineer, Cyber and Autonomous Systems Team
AI Security Institute (AISI)London, UK
London, UK
1 week ago
Cyber Security Engineer, Cyber and Autonomous Systems Team
AI Security Institute (AISI)London, UK
London, UK
1 week ago
Head of Evals
Trajectory LabsBerkeley, CA
Berkeley, CA
1 week ago
Join 60k subscribers and sign up for the EA Newsletter, a monthly email with the latest ideas and opportunities