Amazon.com Inc logo

Software Development Manager, Elastic Kubernetes Service (EKS)

Amazon.com Inc
  • Seattle, WA
    6 days ago

    Job Description

    Job Summary

    Kubernetes only gets useful once you add things to it: networking, storage, DNS, observability, cert management, AWS resource control. Our team is how Amazon EKS customers get all of that without operating any of it: EKS add-ons and EKS capabilities, installed, upgraded, healed, and security-scanned by us, on live clusters, in every AWS region and partition.

    We"re hiring a Manager III, Software Development in Seattle to lead that team. You"ll own three connected problems: the operator that reconciles add-ons onto customer clusters and repairs them when they drift; the release supply chain that takes an add-on from a partner team"s commit to a globally available, CVE-scanned version; and the newest layer: Managed AWS Controllers for Kubernetes (ACK) and kro, where AWS runs the controllers so customers can manage AWS resources from Kubernetes without babysitting the controllers themselves.

    • High blast radius, high leverage: an add-on change lands on running production clusters, so correctness and rollout safety are the design constraints, not an afterthought
    • You"re upstream and downstream of everyone: AWS service teams ship add-ons through your pipeline, and your operator is what actually puts them on customer clusters
    • Real open source: ACK and kro are public projects; the team works in the upstream repos, not just around them

    Key job responsibilities

    • Own the team"s charter, roadmap, and delivery for add-on lifecycle, add-on release, and managed ACK/kro capabilities, including the trade-off calls between new capability and paying down operational load
    • Hire, grow, and retain SDEs across levels; set the bar for design review, code review, and on-call quality
    • Raise release velocity for add-on partner teams while tightening the safety envelope: staged rollouts, bake time, automated validation, fast rollback
    • Own operational excellence for a fleet-wide, multi-region, multi-partition service: alarms that mean something, dashboards leadership can read, COEs that produce fixes rather than paragraphs
    • Drive the security supply chain for add-on images: scan coverage, time-to-patch, and holding the line with provider teams on CRITICAL/HIGH findings
    • Partner with PM and adjacent EKS teams (control plane, connectivity, the internal cluster platform the capability controllers run on) on cross-team designs and launches
    • Represent the team in upstream ACK/kro and EKS add-on community work, and in customer and service-team escalations

    A day in the life

    You start with the health of a system that is always mid-flight: overnight reconciliation and canary signal, an alarm from one region, a partner team"s release stuck in a rollout wave. You spend the morning making those either fixed or owned, and asking why the system needed a human at all.

    Then the design work. An engineer wants to change how the operator detects drift; another is proposing how a managed controller gets scoped credentials to touch a customer"s AWS resources. You review the doc, push on failure modes, and cut scope so it ships in a quarter rather than a year.

    Afternoons are the seams between teams. An AWS service team wants their add-on GA"d in a new partition. A container security review needs an answer on patch latency. Your PM wants to know what the roadmap costs if you also take on the next capability type. You have 1:1s, you write, and you fight for the two or three things that actually matter this quarter.

    Your customers are both external and internal: EKS customers running these add-ons in production, and the AWS service teams and open-source partners who ship through your pipeline.

    About the team

    We"re the add-ons and capabilities team inside Amazon EKS, based in Seattle. Our mission is simple to state and hard to do: everything a customer adds to their cluster should be installed correctly, upgraded safely, patched fast, and boring to operate.

    Culturally we"re a code-is-truth team. Designs are written down and argued about; assumptions get checked against the source and against production rather than asserted; the person who finds the sharp edge is expected to file it down, not just report it. On-call is shared and taken seriously, and we treat operational pain as a design bug. We work in public repos alongside upstream maintainers, which keeps the standard honest.

    Numbers & Facts

    LocationSeattle, WA
    IndustryRetail
    Company Size10,000 employees or more
    Year Founded1994
    Websitehttp://Amazon.com/militaryroles

    About Company

    At Amazon, we don’t wait for the next big idea to present itself. We envision the shape of impossible things and then we boldly make them reality. So far, this mindset has helped us achieve some incredible things. Let’s build new systems, challenge the status quo, and design the world we want to live in. We believe the work you do here will be the best work of your life.

    Wherever you are in your career exploration, Amazon likely has an opportunity for you. Our research scientists and engineers shape the future of natural language understanding with Alexa. Fulfillment center associates around the globe send customer orders from our warehouses to doorsteps. Product managers set feature requirements, strategy, and marketing messages for brand new customer experiences. And as we grow, we’ll add jobs that haven’t been invented yet.

    It’s Always Day 1
    At Amazon, it’s always “Day 1.” Now, what does this mean and why does it matter? It means that our approach remains the same as it was on Amazon’s very first day – to make smart, fast decisions, stay nimble, invent, and stay focused on delighting our customers. In our 2016 shareholder letter, Amazon CEO Jeff Bezos shared his thoughts on how to keep up a Day 1 company mindset. “Staying in Day 1 requires you to experiment patiently, accept failures, plant seeds, protect saplings, and double down when you see customer delight,” he wrote. “A customer-obsessed culture best creates the conditions where all of that can happen.” You can read the full letter here

    Our Leadership Principles
    Our Leadership Principles help us keep a Day 1 mentality. They aren’t just a pretty inspirational wall hanging. Amazonians use them, every day, whether they’re discussing ideas for new projects, deciding on the best solution for a customer’s problem, or interviewing candidates. To read through our Leadership Principles from Customer Obsession to Bias for Action, visit https://www.amazon.jobs/principles

    Skills

    • Amazon Web Services (AWS)unmatched
    • Boringunmatched
    • Code Reviewsunmatched
    • DNS (Domain Name System)unmatched
    • ElastiCunmatched
    • Financial Controlunmatched
    • Leadershipunmatched
    • On Callunmatched
    • Open Sourceunmatched
    • Reporting Dashboardsunmatched
    • Resource Managementunmatched
    • Software Developmentunmatched
    • Software Patchesunmatched
    • Supply Chainunmatched
    • Team Lead/Managerunmatched
    • Vehicle Fleetsunmatched

    Be found by employers

    5,500+ employers search our resume database daily. Add yours to get found by recruiters looking for candidates like you.

    Level up your application

    Professional resume templates

    Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.

    Free resume templates

    Free resume builder

    Improve your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.

    Free resume builder