Responsibilities
- Define and manage the product vision and long-term plan for foundational infrastructure, including computing resources, network systems, multi-region deployment, and cross-cloud reliability.
- Enhance programs that ensure system dependability and recovery, setting platform-wide service level objectives, managing capacity forecasts, enabling failover mechanisms, and reducing incident frequency to support continuous global service availability.
- Evaluate and integrate emerging infrastructure technologies such as container orchestration, service mesh, distributed data storage, monitoring tools, and automated provisioning systems, making strategic decisions on internal development versus third-party solutions to balance efficiency, cost, and growth potential.
- Ensure infrastructure spending aligns with organizational goals, optimize usage of cloud resources, and maintain adherence to legal and compliance standards across all regions of operation.
- Improve how internal teams access and use infrastructure by creating automated, user-friendly tools, streamlining setup, deployment, and monitoring processes, and enhancing overall developer productivity.
- Develop and execute a cost management strategy for infrastructure by analyzing usage data, prioritizing high-impact optimization initiatives, and presenting data-driven justifications for investment to leadership.
Work Arrangement
Hybrid
Other
- The company operates with a remote-first policy, allowing distributed teams while maintaining select in-person roles.
- Team members are expected to participate in quarterly in-person collaboration intensives known as 'surges'.