The Effective Altruism
Opportunities Board
Work on the world's most pressing problems. Browse jobs, fellowships, internships, courses, and more at high-impact organisations.
Head of Evaluations, AI Red Teaming
Trajectory LabsBerkeley, CA
Berkeley, CA
Today
Routes to impact
Direct high impact on an important cause
Description
Lead evaluation quality and design for prompt injection red-teaming, shaping safety work used by frontier AI labs.
- Review evaluation tasks, transcripts, grades, and red-team submissions
- Design methodologies and environments targeting model vulnerabilities
- Build agent tooling, checkers, and pipeline automation
- Improve internal evaluation workflows and scalable quality controls
This text was generated by AI. If you notice any inconsistencies, please let us know using this form.
Related opportunities
Red Team Specialist, Cyber
OpenAISan Francisco, CA / Seattle, WA / Washington, DC
San Francisco, CA / Seattle, WA / Washington, DC
4 days ago
Machine Learning Engineer
10a LabsRemote (US)
Remote (US)
2 months ago
Machine Learning Researcher
Gray SwanRemote
Remote
3 months ago
AI Security Research Engineer
0LabsRemote
Remote
4 months ago
AI Cyber Red Teamer
3 days ago
Software Engineer, Safeguards Evaluations
AnthropicSan Francisco, CA | New York City, NY
San Francisco, CA | New York City, NY
2 months ago
Applied Control Researcher
Apollo ResearchLondon, United Kingdom / San Francisco, USA
London, United Kingdom / San Francisco, USA
3 months ago
Applied Researcher (Product)
Apollo ResearchLondon, United Kingdom
London, United Kingdom
5 months ago
Join 60k subscribers and sign up for the EA Newsletter, a monthly email with the latest ideas and opportunities