Responsibilities
- Collaborate with business teams to collect requirements and develop solutions that meet engineering standards and business objectives.
- Design, deploy, and operate Red Hat OpenShift clusters in on-premises environments, including physical and virtual infrastructure.
- Lead systems engineering initiatives focused on OpenShift platform development and maintenance.
- Work closely with development teams to containerize applications and streamline deployment and release processes.
- Develop and sustain CI/CD pipelines using OpenShift and complementary DevOps tooling.
- Automate deployment and operational tasks using Ansible and other scripting technologies.
- Perform regular patching and upgrades of OpenShift clusters, ArgoCD, GitLab CI, and associated components in line with change management policies.
- Create, implement, and manage backup, disaster recovery, and redundancy strategies for critical systems.
- Participate in on-call rotations to ensure continuous support for essential platform and infrastructure services.
- Support day-to-day operations through trend analysis, root cause investigation, monitoring, issue troubleshooting, and resolution to maintain system reliability.
- Enforce security standards such as role-based access control, vulnerability remediation, and compliance measures.
- Enhance alerting, monitoring, and observability across OpenShift and connected systems.
- Perform general system administration tasks such as Linux server management, infrastructure upkeep, and data center support when OpenShift workload is low.
- Create and keep current technical documentation, including architecture diagrams, operational guides, and runbooks.
- Deliver training, mentoring, and knowledge sharing on OpenShift, GitLab, and related infrastructure technologies.