The Effective Altruism
Opportunities Board
Work on the world's most pressing problems. Browse jobs, fellowships, internships, courses, and more at high-impact organisations.
Research Engineer, Scalable Interpretability
TransluceSan Francisco, CA
San Francisco, CA
2 months ago
Routes to impact
Direct high impact on an important cause
Skill-building & building career capital
Description
Develop scalable interpretability systems to improve oversight of advanced AI models.
- Build evaluations for undesirable model behaviors
- Design architectures and training objectives for interpretability assistants
- Scale training and inference for frontier models
- Conduct research on model activations and behavior prediction
This text was generated by AI. If you notice any inconsistencies, please let us know using this form.
Related opportunities
Data Scientist
Innovations for Poverty ActionKenya / Ghana / Colombia / Peru
Kenya / Ghana / Colombia / Peru
2 months ago
Member of Technical Staff, Engineering
AI DigestRemote
Remote
1 week ago
Associate, AI and Advanced Computing
Schmidt SciencesNew York, NY
New York, NY
2 weeks ago
Research Assistant on AI Safety
University of OxfordOxford, UK
Oxford, UK
3 weeks ago
Research Scientist
GoodfireLondon, UK
London, UK
4 weeks ago
Researcher, Alignment Chain of Thought Monitorability
OpenAISan Francisco, CA
San Francisco, CA
1 month ago
Machine Learning Engineer
Gray SwanPittsburgh, PA / remote
Pittsburgh, PA / remote
1 month ago
Research Resident, Emerging Technology and Security, Open Source
RANDRemote (U.S.)
Remote (U.S.)
1 month ago
Join 60k subscribers and sign up for the EA Newsletter, a monthly email with the latest ideas and opportunities