Operate production and non-production RKE2/Rancher Kubernetes clusters, including upgrades, node lifecycle, networking, ingress, DNS, certificates, and capacity; troubleshoot control-plane, scheduling, CoreDNS, and memory (OOM) issues, and improve observability, alerting, and runbooks. Run Kubernetes on physical HPE servers and virtual machines in company data centers, including hardware, firmware, RAID, and out-of-band management; partner with infrastructure teams on Cisco networking, SAN/NVMe/object storage, failure domains, and capacity planning.