The Effective Altruism
Opportunities Board
Work on the world's most pressing problems. Browse jobs, fellowships, internships, courses, and more at high-impact organisations.
Safeguards Enforcement Analyst, Violence and Extremism
AnthropicSan Francisco, CA / New York City, NY / Washington, DC
San Francisco, CA / New York City, NY / Washington, DC
1 month ago
Salary
$285,000 – $330,000
Routes to impact
Direct high impact on an important cause
Skill-building & building career capital
Description
Help build and improve AI safeguards against violence and extremist misuse through policy enforcement, evaluations, and threat analysis.
- Design scalable enforcement workflows and automated detection systems.
- Review high-risk content and identify emerging misuse patterns.
- Partner with engineering to improve policy enforcement and model evaluations.
- Inform policy updates with structured feedback from enforcement decisions.
This text was generated by AI. If you notice any inconsistencies, please let us know using this form.
Related opportunities
Regional Research Economist, Economic Research, London
AnthropicLondon, UK
London, UK
2 months ago
Research Resident, Emerging Technology and Security, Open Source
RANDRemote (U.S.)
Remote (U.S.)
1 month ago
Research Resident, AIxBio Threats and Mitigations
RANDRemote (U.S.)
Remote (U.S.)
1 month ago
Senior Research Scientist, AI Safety Evaluations
FacultyLondon, UK
London, UK
1 month ago
Risk Modeling Lead
SaferAIParis, France / London, UK / San Francisco, CA / Remote
Paris, France / London, UK / San Francisco, CA / Remote
2 months ago
Research Scientist/Engineer (Science of Scheming)
Apollo ResearchLondon, United Kingdom
London, United Kingdom
6 months ago
Research Scientist/Engineer (Evaluations)
Apollo ResearchLondon, United Kingdom
London, United Kingdom
6 months ago
Model Policy Manager
OpenAISan Francisco, CA
San Francisco, CA
2 weeks ago
Join 60k subscribers and sign up for the EA Newsletter, a monthly email with the latest ideas and opportunities