Responsibilities
- Serve as the main contact for development teams on infrastructure, deployment, and access-related issues, handling inquiries through Jira and Slack with an emphasis on prompt and clear communication.
- Manage and scale multi-cloud environments on AWS and GCP using Infrastructure as Code with Terraform.
- Enhance developer productivity by maintaining and refining command-line tools and internal APIs that streamline complex processes for engineering teams.
- Design, implement, and optimize CI/CD pipelines using GitHub Actions and adopt modern GitOps practices with ArgoCD.
- Ensure system reliability through proactive monitoring, performance tracking, and automated recovery mechanisms, including participation in on-call rotations.
- Explore and integrate AI-driven solutions, such as LLMs or autonomous agents, to automate repetitive support and documentation tasks.
Work Arrangement
Hybrid
Team
Engineering
Responsibilities (6)
- Tiered Troubleshooting: Act as the primary point of contact for developers on infrastructure, deployment, and access issues. You will manage incoming requests via Jira and Slack with a focus on high responsiveness and clarity.
- Automated Infrastructure: Assist in managing and scaling multi-cloud environments (AWS/GCP) using Infrastructure as Code (Terraform).
- Developer Experience (DevEx): Maintain and improve CLI tools and internal APIs that simplify complex workflows for our application teams.
- CI/CD Pipeline Management: Build and optimize deployment pipelines in GitHub Actions, and modern GitOps (ArgoCD) patterns.
- Observability & Reliability: Monitor system health and performance, implementing automated "self-healing" triggers to reduce manual intervention. This includes participating in an on-call rotation.
- AI Integration: Experiment with and implement LLM-based or agentic workflows to automate routine troubleshooting and documentation tasks.
Available