The Effective Altruism
Opportunities Board
Work on the world's most pressing problems. Browse jobs, fellowships, internships, courses, and more at high-impact organisations.
Machine Learning Research Scientist
AI X-risk Institute (AIXI Labs)London, UK / Berkeley, CA
London, UK / Berkeley, CA
Today
Deadline
2026-10-01
Routes to impact
Direct high impact on an important cause
Description
Research scientist designing and running experiments on LLM-based agents to test theoretically-predicted risk factors for loss of control.
- Build and evaluate safety mitigations grounded in algorithmic information theory
- Test for deception, power-seeking, and specification gaming under realistic conditions
- Help shape the research agenda within a three-person core team
- Publish papers and posts for the AI safety and ML communities
This text was generated by AI. If you notice any inconsistencies, please let us know using this form.
Related opportunities
Member of Technical Staff, Engineering
AI DigestRemote
Remote
1 week ago
Red Team Specialist, Cyber
OpenAISan Francisco, CA / Seattle, WA / Washington, DC
San Francisco, CA / Seattle, WA / Washington, DC
1 week ago
Research Assistant on AI Safety
University of OxfordOxford, UK
Oxford, UK
2 weeks ago
Research Engineer / Research Scientist
Center for AI Safety (CAIS)San Francisco, CA
San Francisco, CA
1 month ago
Research Scientist, Manipulation Evaluations
Apart ResearchRemote (Europe preferred)
Remote (Europe preferred)
2 months ago
Research Engineer, Scalable Interpretability
TransluceSan Francisco, CA
San Francisco, CA
2 months ago
Technical Associate, Robotics and Physical AI
Cambridge, MA
1 week ago
Data Scientist
Innovations for Poverty ActionKenya / Ghana / Colombia / Peru
Kenya / Ghana / Colombia / Peru
2 months ago
Join 60k subscribers and sign up for the EA Newsletter, a monthly email with the latest ideas and opportunities