Our team builds the ML-inference stack that powers generative AI for Apple Intelligences Private Cloud Compute - running on Apple Silicon in the datacenter, distributing work across on-SoC acceleration hardware and multi-node clusters. Built on Private Cloud Computes privacy guarantees, were growing the team to scale across more platforms and support a widening set of features. As part of the team you will help engineer continuous improvements in stability and performance for Private Cloud Compute, help implement entirely new functionality as it emerges from the research community, and help bring our inference stack up on new generations of SoCs and hardware acceleration IP as we extend to more platforms - in collaboration with hardware, product and research teams throughout Apple.
We write performant and scalable frameworks (primarily in Swift, with C++ where we bridge to the hardware) to distribute and coordinate ML inference across the acceleration IP blocks of different SoCs, and to move data and coordinate work reliably across multi-node inference clusters. You will integrate inference code into a full service stack so that user traffic is served reliably and performantly, with a strong focus on code that is easy and safe to develop, update, and monitor in production.
Were a collection of highly skilled and friendly engineers who value each others opinions and experience. We strive for excellence and believe strongly in the quality of our output. We are a team of domain experts, each specializing in specific core subject areas, with broad collective experience across cloud software services and platforms.2 Years practical experience plus Bachelors degree in Computer Science, Computer Engineering, or a related field - or equivalent practical experience. Experience building large-scale distributed systems that serve ML inference, reasoning across processes, hosts, and service tiers as well as model behavior under load. A ML performance-centric mindset - able to reason about latency/throughput trade-offs and to distinguish what can be solved at the system level from what requires model-system co-design. Strong in a systems or server language - Python, Go, Rust, Java, C++, or similar. Swift is what we write, but we dont expect it going in.Low-level or close-to-the-metal work - systems programming, performance, or hardware/SoC bring-up. Server-side Swift, or RPC/networking stacks (gRPC, Protocol Buffers). Strong debugging and observability instincts - fluent in logs, metrics, and traces, with observability-stack (OpenTelemetry, Splunk), SLO/error-budget, and on-call experience, and able to drive a production incident to root cause. Depth in ML inference serving optimizations - quantization, sparsity, batching, KV cache, tokenization, GPU acceleration - enough to optimize the system and reason about the trade-offs and requirements it places on models (you wont be training them). Solid grasp of concurrency, async/streaming, resource lifecycle, and error/cancellation handling; bonus for Apple platform experience (XPC, Instruments, Swift Concurrency).
| Location | Seattle, WA |
| Industry | Computer/IT Services |
| Company Size | 10,000 employees or more |
| Year Founded | 1976 |
| Website | https://www.apple.com/jobs |
We’re a diverse collection of thinkers and doers, continually reimagining what’s possible to help us all do what we love in new ways. The people who work here have reinvented entire industries with the Mac, iPhone, iPad, and Apple Watch, as well as with services, including iTunes, the App Store, Apple Music, and Apple Pay. And the same passion for innovation that goes into our products also applies to our practices — strengthening our commitment to leave the world better than we found it.
There’s a place here for every kind of brilliant. Everyone here is an innovator, or an innovator-to-be, no matter what your team or your role. So bring your passion, courage, and original thinking and get ready to share it, because every new product, service, or feature we invent is the result of people working together to make each others’ ideas stronger. Innovation at this level depends on people who represent the variety of the human experience and inspire us with their own fresh perspectives. Together, we’ll do amazing work that can make a difference in people’s lives. Including your own. Learn more about working at Apple.
Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.
Free resume templatesImprove your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.
Free resume builder