The Effective Altruism
Opportunities Board
Work on the world's most pressing problems. Browse jobs, fellowships, internships, courses, and more at high-impact organisations.
Research Engineer, AGI Safety and Alignment
Google DeepMindSan Francisco, CA / Mountain View, CA / New York, NY
San Francisco, CA / Mountain View, CA / New York, NY
2 weeks ago
Salary
$174,000 – $252,000
Routes to impact
Direct high impact on an important cause
Description
Research and engineer methods to reduce catastrophic risks from advanced AI systems through alignment, control, and interpretability work.
- Study alignment failures and develop scalable alignment techniques for frontier models
- Build adversarially robust control systems for production use
- Research interpretability methods and support adoption by product teams
This text was generated by AI. If you notice any inconsistencies, please let us know using this form.
Related opportunities
Research Scientist, Safety Oversight
Google DeepMindMountain View, CA
Mountain View, CA
2 weeks ago
PhD Position, Responsible Machine Learning
ELLIS Institute TübingenVienna, Austria
Vienna, Austria
Yesterday
Research Engineer, Value Persistence Through Reinforcement Learning
Yesterday
Research Engineer / Research Scientist, Control Red Team
AI Security Institute (AISI)London, UK
London, UK
5 days ago
Research Engineer / Research Scientist, Misuse Red Team
AI Security Institute (AISI)London, UK
London, UK
5 days ago
Research Scientist, Control
Apollo ResearchSan Francisco, CA / London, UK
San Francisco, CA / London, UK
6 days ago
AI Red Team Engineer
Apollo ResearchSan Francisco, CA / London, UK
San Francisco, CA / London, UK
6 days ago
AI Security and Control Researcher
Apollo ResearchLondon, UK / San Francisco, CA
London, UK / San Francisco, CA
6 days ago
Join 60k subscribers and sign up for the EA Newsletter, a monthly email with the latest ideas and opportunities