The Effective Altruism
Opportunities Board
Work on the world's most pressing problems. Browse jobs, fellowships, internships, courses, and more at high-impact organisations.
Head of Evaluations, AI Red Teaming
Trajectory LabsBerkeley, CA
Berkeley, CA
3 days ago
Salary
$200,000 – $400,000
Routes to impact
Direct high impact on an important cause
Description
Lead evaluation quality and design for prompt injection red-teaming, shaping safety work used by frontier AI labs.
- Review evaluation tasks, transcripts, grades, and red-team submissions
- Design methodologies and environments targeting model vulnerabilities
- Build agent tooling, checkers, and pipeline automation
- Improve internal evaluation workflows and scalable quality controls
This text was generated by AI. If you notice any inconsistencies, please let us know using this form.
Related opportunities
Research Scientist, Control
Apollo ResearchSan Francisco, CA / London, UK
San Francisco, CA / London, UK
1 day ago
Red Team Specialist, Cyber
OpenAISan Francisco, CA / Seattle, WA / Washington, DC
San Francisco, CA / Seattle, WA / Washington, DC
1 week ago
Machine Learning Engineer
10a LabsRemote (US)
Remote (US)
2 months ago
Machine Learning Researcher
Gray SwanRemote
Remote
3 months ago
AI Security Research Engineer
0LabsRemote
Remote
4 months ago
Research Engineer / Research Scientist, Control Red Team
AI Security Institute (AISI)London, UK
London, UK
Yesterday
Research Engineer / Research Scientist, Misuse Red Team
AI Security Institute (AISI)London, UK
London, UK
Yesterday
AI Cyber Red Teamer
6 days ago
Join 60k subscribers and sign up for the EA Newsletter, a monthly email with the latest ideas and opportunities