Responsibilities
- Evaluate system and database performance through benchmarking, analysis, and capacity planning to drive optimization.
- Diagnose and resolve application and server-level issues by analyzing logs and routing problems appropriately.
- Propose configuration adjustments and tuning strategies to alleviate performance constraints.
- Collaborate with core development, cloud infrastructure, and security teams to enhance cloud platform efficiency.
- Design and lead chaos engineering initiatives aligned with internal engineering priorities.
- Build, implement, and maintain tools for executing controlled chaos experiments and assessing system impact.
- Demonstrate strong interest in mastering complex, large-scale distributed systems.
- Analyze challenges in software resilience, operations, and deployment workflows.
- Enhance backend systems to support chaos engineering practices across environments.
- Monitor live systems and identify high-impact methods to test system robustness.
Work Arrangement
Remote