The Effective Altruism
Opportunities Board
Work on the world's most pressing problems. Browse jobs, fellowships, internships, courses, and more at high-impact organisations.
Senior Software Engineer, GPU Cluster Infrastructure
FAR.AIRemote
Remote
Today
Routes to impact
Direct high impact on an important cause
Description
Operate and scale GPU cluster infrastructure that enables large-scale AI safety research and experiments.
- Manage Kubernetes GPU fleets, batch scheduling, storage, networking, and capacity
- Improve fault tolerance for distributed training and multi-node workloads
- Harden cluster security and sandbox autonomous AI agents
- Work directly with researchers to turn infrastructure problems into platform improvements
This text was generated by AI. If you notice any inconsistencies, please let us know using this form.
Related opportunities
System Administrator
Model Evaluation & Threat Research (METR)Berkeley, CA
Berkeley, CA
2 weeks ago
Software Engineer, Infrastructure, Interpretability
AnthropicSan Francisco, CA / New York City, NY
San Francisco, CA / New York City, NY
4 weeks ago
Senior DevOps Engineer
IrregularTel Aviv, Israel
Tel Aviv, Israel
1 month ago
Software Engineer
Gray SwanRemote (US)
Remote (US)
1 month ago
Senior Software Engineer
Gray SwanPittsburgh, PA
Pittsburgh, PA
1 month ago
Security Engineer
Model Evaluation & Threat Research (METR)Berkeley, CA
Berkeley, CA
1 month ago
Founding Engineer, Infrastructure
SaferAIParis, France / London, UK / San Francisco, CA
Paris, France / London, UK / San Francisco, CA
2 months ago
Member of Technical Staff, Secure Infrastructure and Platform
Security Level 5San Francisco, CA
San Francisco, CA
2 months ago
Join 60k subscribers and sign up for the EA Newsletter, a monthly email with the latest ideas and opportunities