Bachelors Degree in Computer Science, an engineering-related field, or equivalent related experience 8+ years in a Site Reliability Engineering, DevOps, or Infrastructure focused role Deep experience operating large-scale, multi-tenant Kubernetes environments in production Strong systems background - comfortable troubleshooting across the full stack (network, OS, container runtime, application) Expert-level Go, with a track record of shipping and owning controllers or operators that other teams depend on. Experience with configuration management at scale (Puppet, Ansible, or equivalent) Demonstrated ability to drive cross-functional initiatives to completion Strong written and verbal communication skillsExperience with third-party cloud platforms (AWS, GCP, or Azure) Familiarity with bare-metal provisioning and lifecycle management at datacenter scale Understanding of cloud-native observability (Prometheus, Thanos, Splunk, or similar) Experience running infrastructure as an internal managed service with defined SLAs.