Skills
About the Role
Ai2 is a non-profit research institute advancing open-source AI. We share our findings, data, code, and models with the global scientific community. We’re looking for a Director of AI Infrastructure to lead the platforms that power our research—owning the end-to-end lifecycle of our high-performance computing (HPC) environment.
Responsibilities
- Oversee on-prem GPU cluster operations and reliability for research workloads
- Own the software orchestration layer that schedules and manages workloads across a hybrid cloud setup
- Drive architecture, deployment, and continuous improvement of infrastructure supporting AI training and experimentation
- Partner with research and engineering teams to translate compute needs into scalable infrastructure
- Establish operational best practices, monitoring, and incident response processes
- Manage infrastructure roadmaps, capacity planning, and vendor/partner coordination when needed
Requirements
- Proven leadership experience delivering and operating HPC and/or GPU compute environments
- Strong understanding of hybrid cloud concepts and workload scheduling/orchestration
- Experience with performance, reliability, and observability practices for large-scale systems
- Ability to collaborate across technical and research stakeholders
- Onsite work in Seattle is required; on-site expectations can vary by team and role
Compensation & Benefits
- Base salary range: $176,400–$264,600, plus generous bonus plans