The Effective Altruism
Opportunities Board
Work on the world's most pressing problems. Browse jobs, fellowships, internships, courses, and more at high-impact organisations.
Research Engineer, Evals
White CircleParis, France / London, UK
Paris, France / London, UK
2 months ago
Routes to impact
Direct high impact on an important cause
Skill-building & building career capital
Description
Help develop and evaluate benchmarks for AI model safety and reliability research.
- Build and maintain benchmarks for content and agentic safety.
- Study real-world AI agent failure modes and model behavior.
- Develop evaluation tooling supporting research and production systems.
- Collaborate across research and product teams on new safety features.
This text was generated by AI. If you notice any inconsistencies, please let us know using this form.
Related opportunities
Researcher, Alignment Chain of Thought Monitorability
OpenAISan Francisco, CA
San Francisco, CA
1 month ago
Machine Learning Engineer
Gray SwanPittsburgh, PA / remote
Pittsburgh, PA / remote
1 month ago
Senior Research Scientist, AI Safety Evaluations
FacultyLondon, UK
London, UK
1 month ago
Applied AI Data Scientist
AE StudioRemote / US
Remote / US
1 month ago
Applied AI Data Scientist
AE StudioFlorianopolis, Brazil / Remote
Florianopolis, Brazil / Remote
1 month ago
Senior Machine Learning Data Processing Developer
LawZeroMontreal, Canada
Montreal, Canada
2 months ago
Data Engineer, Safeguards
AnthropicSan Francisco, CA / New York City, NY
San Francisco, CA / New York City, NY
2 months ago
ML Engineer
Tilde ResearchSan Francisco, CA
San Francisco, CA
2 months ago
Join 60k subscribers and sign up for the EA Newsletter, a monthly email with the latest ideas and opportunities