Apple Inc logo

Sr. Linux Engineer - DNS and Global Server Load Balancing, Infrastructure Services

Apple Inc
  • Sunnyvale, CA
    1 day ago

    Job Description

    Infrastructure Services is part of IS&T and the foundation of Apples global network operations - managing data center equipment and systems to deliver compute, storage, and networking services for teams across Apple, including its internal developer community. From individual facilities to a worldwide network, Infrastructure Services ensures the technology underneath everything works without question

    Apples services reach hundreds of millions of people every day, and Edge Services builds and runs the authoritative DNS and global server load balancing infrastructure that sits at the front door of all of them. Every request to an Apple service begins here, in the first few milliseconds where our systems decide how and where to answer. The reliability, correctness, and performance of that front door directly shape the experience people have with Apple, and keeping it fast and dependable at global scale is the work of our team.

    Were looking for a thoughtful, experienced engineer to bring technical leadership to this infrastructure. Youll operate and extend bare-metal and cloud systems distributed across dozens of data centers and network points of presence worldwide, build deep traffic and resolution visibility, drive performance and capacity planning, and create the automation that keeps our 24x7 operations calm and reliable. Youll also mentor engineers across the team and partner closely with our software, network, and product engineering groups. We believe great systems are built by great teams and care as much about how we work together as about what we ship. If youre energized by owning the reliability of critical infrastructure at global scale, wed like to hear from you.

    This is a senior individual-contributor role and a technical anchor for the Edge Services DNS and GSLB systems. The essential functions of the job are to operate and extend distributed Linux infrastructure as code across geographically dispersed sites; to define and maintain the observability, alerting, and service-level objectives that keep the systems healthy; to build automation that removes operational toil; to plan fleet capacity and hardware lifecycle across globally distributed sites; and to mentor engineers while partnering across software, network, and product teams. The role includes incident and process management and participation in a shared 24x7 production on-call rotation.Operates and extends bare-metal and cloud Linux infrastructure across dozens of edge sites and points of presence, treating infrastructure as code so that standing up or rebuilding a site is predictable and repeatable. Owns the reliability, correctness, and performance of the authoritative DNS and global server load balancing systems, including incident response and participation in a 24x7 on-call rotation. Builds deep visibility into the systems by defining meaningful service-level objectives, alerting, and dashboards that span every point of presence and surface problems early. Develops automation and tooling in Python, Go, Rust, and/or Swift - supported by generative AI tooling - to remove operational toil and turn manual tasks into reliable, self-sustaining systems. Guides fleet capacity, growth, and hardware lifecycle across geographically distributed sites through smart tooling, thoughtful planning, and collaboration with leadership and program management. Mentors and grows engineers at every career stage and partners with other senior engineers to raise the bar on architecture, tooling, and design. Partners with software, network, and product engineering teams to deliver large-scale rollouts and keep 24x7 operations running smoothly. Represents the teams work, needs, and ideas to partners and leadership through clear, collaborative communication. Demonstrated experience operating Linux systems in production, including software deployment and CI/CD workflows. Working knowledge of networking fundamentals, with hands-on experience troubleshooting TCP/UDP and common layer 2-3 issues. Experience with configuration management or infrastructure-as-code tooling (for example, Salt, Ansible, Puppet, or Terraform) and with observability tooling (for example, Prometheus and Grafana, or equivalents). Proficiency in at least one programming language used for automation and tooling (for example, Python, Go, Rust, or Swift). Willingness to participate in a shared 24x7 on-call rotation. Bachelors degree in Computer Science or a related field, or equivalent practical experience. Extensive experience operating large-scale infrastructure across multiple data centers or network points of presence. Expertise with anycast routing and BGP, and a strong understanding of how DNS, load balancing, and the network layer interact. Depth in Linux internals, including kernel networking, and in package management and software deployment at fleet scale. Experience defining service-level objectives and indicators (SLOs/SLIs) and driving reliability programs across distributed systems. Familiarity with secure-by-default operations, such as DNSSEC key management, change safety, and progressive configuration rollout. Experience with scale and performance testing, disaster recovery, and capacity planning. Experience managing hardware lifecycle across edge and regional network environments. Track record of leading end-to-end projects, defining technical roadmaps, and driving cross-functional alignment on architecture and best practices. Experience mentoring engineers and leading code reviews.

    Numbers & Facts

    LocationSunnyvale, CA
    IndustryComputer/IT Services
    Company Size10,000 employees or more
    Year Founded1976
    Websitehttps://www.apple.com/jobs

    About Company

    We bring amazing people together to make amazing things happen.

    We’re a diverse collection of thinkers and doers, continually reimagining what’s possible to help us all do what we love in new ways. The people who work here have reinvented entire industries with the Mac, iPhone, iPad, and Apple Watch, as well as with services, including iTunes, the App Store, Apple Music, and Apple Pay. And the same passion for innovation that goes into our products also applies to our practices — strengthening our commitment to leave the world better than we found it.

    About Apple

    There’s a place here for every kind of brilliant. Everyone here is an innovator, or an innovator-to-be, no matter what your team or your role. So bring your passion, courage, and original thinking and get ready to share it, because every new product, service, or feature we invent is the result of people working together to make each others’ ideas stronger. Innovation at this level depends on people who represent the variety of the human experience and inspire us with their own fresh perspectives. Together, we’ll do amazing work that can make a difference in people’s lives. Including your own. Learn more about working at Apple.

    Skills

    • Ansibleunmatched
    • Appleunmatched
    • Artificial Intelligence (AI)unmatched
    • Automationunmatched
    • BGPunmatched
    • Best Practicesunmatched
    • Capacity Managementunmatched
    • Capacity and Performance Managementunmatched
    • Cloud Computingunmatched
    • Code Reviewsunmatched
    • Communication Skillsunmatched
    • Computer Scienceunmatched
    • Configuration Managementunmatched
    • Continuous Deployment/Deliveryunmatched
    • Continuous Integrationunmatched
    • Cross-Functionalunmatched
    • DNS (Domain Name System)unmatched
    • Data Managementunmatched
    • Disaster Recoveryunmatched
    • Distributed Computingunmatched
    • Go Programming Language (Golang)unmatched
    • Incident Managementunmatched
    • Incident Responseunmatched
    • Kernel Programmingunmatched
    • Leadershipunmatched
    • Linux Operating Systemunmatched
    • Load Balancingunmatched
    • Machine Toolunmatched
    • Mentoringunmatched
    • Network Administration/Managementunmatched
    • Network Monitoringunmatched
    • Network Operations Centerunmatched
    • Network Routingunmatched
    • On Callunmatched
    • Performance Testingunmatched
    • Process Managementunmatched
    • Product Engineeringunmatched
    • Production Systemsunmatched
    • Programming Languagesunmatched
    • Project/Program Managementunmatched
    • Puppet (Configuration Management)unmatched
    • Python Programming/Scripting Languageunmatched
    • Reporting Dashboardsunmatched
    • Rust Programming Languageunmatched
    • Software Distributionunmatched
    • Software Engineeringunmatched
    • Systems Maintenanceunmatched
    • TCP (Transmission Control Protocol)unmatched
    • Team Lead/Managerunmatched
    • Team Playerunmatched
    • Technical Leadershipunmatched
    • Testingunmatched
    • UDP (User Datagram Protocol)unmatched
    • Vehicle Fleetsunmatched

    Be found by employers

    5,500+ employers search our resume database daily. Add yours to get found by recruiters looking for candidates like you.

    Level up your application

    Professional resume templates

    Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.

    Free resume templates

    Free resume builder

    Improve your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.

    Free resume builder