Responsibilities
- Manage daily operations of VMware platforms including vSphere/ESX/ESXi, covering lifecycle, capacity, performance, patching, upgrades, backup coordination, and incident resolution.
- Collaborate on virtualization strategy, including planning and executing migration from VMware to Nutanix or similar platforms with minimal disruption and risk.
- Operate and enhance storage and compute systems such as Cisco UCS, NetApp FAS/AFF, and FlexPod, focusing on performance, troubleshooting, and fault tolerance.
- Administer and optimize Microsoft Azure deployments, including networking, compute resources, storage, security policies, cost efficiency, and monitoring.
- Maintain and troubleshoot Kubernetes clusters, including upgrades, networking, ingress, policy enforcement, and workload stability.
- Develop, deploy, and support containerized applications using Docker, including infrastructure for registries, image management, runtime security, and scaling.
- Implement and support Layer 4 and Layer 7 load balancing solutions, including TLS setup and monitoring.
- Design and manage enterprise-wide data, voice, and video networks across LAN, WAN, and Wi-Fi environments.
- Configure and maintain routing and switching infrastructure, including TCP/IP, BGP, OSPF, VLANs, QoS, and core services like DHCP and DNS.
- Administer and secure firewall and VPN systems from vendors such as SonicWall, Cisco ASA, and Palo Alto, supporting secure remote access.
- Ensure secure internal and external file transfers using protocols like SFTP and managed file transfer, including certificate and key management, access controls, and audit logging.
- Support remediation of vulnerabilities and implement security hardening measures in alignment with compliance standards such as PCI and SOC.
- Manage the full lifecycle of SSL/TLS certificates across all infrastructure components.
- Develop automated, repeatable operational processes using Infrastructure-as-Code and configuration management tools to reduce manual effort and improve system reliability.
- Enhance system observability through monitoring, alerting, packet capture, and traffic analysis to prevent recurring issues.
- Participate in on-call rotations to support critical incidents and scheduled maintenance outside normal business hours.
- Work with internal teams and external vendors or clients to coordinate changes, perform upgrades, resolve incidents, and document results.
- Create and maintain comprehensive technical documentation, including runbooks, network diagrams, and operational standards, to ensure consistency and knowledge transfer.
Work Arrangement
On-site — Charlotte, NC