The Effective Altruism
Opportunities Board
Work on the world's most pressing problems. Browse jobs, fellowships, internships, courses, and more at high-impact organisations.
Safeguards Enforcement Analyst, Violence and Extremism
AnthropicSan Francisco, CA / New York City, NY / Washington, DC
San Francisco, CA / New York City, NY / Washington, DC
Today
Salary
$285,000 – $330,000
Routes to impact
Direct high impact on an important cause
Skill-building & building career capital
Description
Help build and improve AI safeguards against violence and extremist misuse through policy enforcement, evaluations, and threat analysis.
- Design scalable enforcement workflows and automated detection systems.
- Review high-risk content and identify emerging misuse patterns.
- Partner with engineering to improve policy enforcement and model evaluations.
- Inform policy updates with structured feedback from enforcement decisions.
This text was generated by AI. If you notice any inconsistencies, please let us know using this form.
Related opportunities
Regional Research Economist, Economic Research, London
AnthropicLondon, UK
London, UK
1 month ago
Specialist, Technical AI Governance
Simon Institute for Longterm Governance (SI)Geneva, Switzerland
Geneva, Switzerland
2 weeks ago
Senior Research Scientist, AI Safety Evaluations
FacultyLondon, UK
London, UK
Today
Frontier AI Risks Lead
OpenAISan Francisco, CA
San Francisco, CA
Today
Research Strategist, Emerging Impacts Team
Google DeepMindMountain View, CA / New York, NY
Mountain View, CA / New York, NY
1 week ago
Risk Modeling Lead
SaferAIParis, France / London, UK / San Francisco, CA / Remote
Paris, France / London, UK / San Francisco, CA / Remote
3 weeks ago
National Security Analyst
Aiken, SC
2 months ago
Team Member, Model Policy
OpenAISan Francisco, CA
San Francisco, CA
2 months ago
Join 60k subscribers and sign up for the EA Newsletter, a monthly email with the latest ideas and opportunities