The Effective Altruism
Opportunities Board
Work on the world's most pressing problems. Browse jobs, fellowships, internships, courses, and more at high-impact organisations.
Researcher, Recursive Self-Improvement Safety
OpenAISan Francisco, CA
San Francisco, CA
1 month ago
Routes to impact
Direct high impact on an important cause
Description
Conduct technical research and build safety measures for advanced AI systems, focusing on risks from recursive self-improvement and loss of control.
- Develop scalable oversight, automated auditing, and model behavior evaluations
- Stress-test monitoring systems for scheming and other misalignment risks
- Build prototypes and integrate successful approaches into safety pipelines
- Coordinate technical safety work across teams and external communications
This text was generated by AI. If you notice any inconsistencies, please let us know using this form.
Related opportunities
Researcher, Alignment Interpretability
OpenAISan Francisco, CA
San Francisco, CA
1 week ago
Researcher, Alignment Chain of Thought Monitorability
OpenAISan Francisco, CA
San Francisco, CA
1 month ago
Research Engineer, Takeoff Intel
AnthropicSan Francisco, CA
San Francisco, CA
5 days ago
Expression of Interest
Trajectory LabsBerkeley, CA
Berkeley, CA
5 days ago
Tech Lead, AI Safety
Trajectory LabsBerkeley, CA
Berkeley, CA
5 days ago
Engineering Lead, Chem Bio
AI Security Institute (AISI)London, UK
London, UK
5 days ago
Founding Member, Technical Staff
ParallaxLondon, UK / Remote (Europe)
London, UK / Remote (Europe)
1 week ago
AI Cyber Red Teamer
2 weeks ago
Join 60k subscribers and sign up for the EA Newsletter, a monthly email with the latest ideas and opportunities