The Effective Altruism
Opportunities Board
Work on the world's most pressing problems. Browse jobs, fellowships, internships, courses, and more at high-impact organisations.
Research Scientist, Agent Robustness
ScaleSan Francisco, CA / New York, NY
San Francisco, CA / New York, NY
Today
Salary
$216,000 – $270,000
Routes to impact
Direct high impact on an important cause
Description
Conduct technical research on AI agent robustness, safety evaluations, and mitigations for emerging failure modes.
- Build evaluation harnesses and prototypes for agent safety research
- Study harmful actions, adversarial behavior, and multi-agent risks
- Develop exploits and mitigations for agent failure modes
- Collaborate across research, industry, academia, and public-sector teams
This text was generated by AI. If you notice any inconsistencies, please let us know using this form.
Related opportunities
Member of Technical Staff, Embedded Assessments
Model Evaluation & Threat Research (METR)Berkeley, CA
Berkeley, CA
Today
Research Scientist
GoodfireLondon, UK
London, UK
2 weeks ago
Research Engineer, Scalable Interpretability
TransluceSan Francisco, CA
San Francisco, CA
2 months ago
Machine Learning Manager
LawZeroMontreal, Canada
Montreal, Canada
2 months ago
Member of Technical Staff, Design Engineer
ValthosNew York, NY / San Francisco, CA
New York, NY / San Francisco, CA
4 weeks ago
Member of Technical Staff, Applied Computational Biology
ValthosNew York, NY / San Francisco, CA
New York, NY / San Francisco, CA
4 weeks ago
Data Scientist
Innovations for Poverty ActionKenya / Ghana / Colombia / Peru
Kenya / Ghana / Colombia / Peru
1 month ago
Member of Technical Staff, Design Engineering
AI DigestRemote
Remote
Today
Join 60k subscribers and sign up for the EA Newsletter, a monthly email with the latest ideas and opportunities