iPhone System Reliability Engineer AppleiPhone System Reliability EngineerCupertino, CAFamiliarity with tools such as Tableau for building dashboards + Ability to travel internationally without restriction + Demonstrate an unwavering passion for engineering products with exceptional reliability **Preferred Qualifications** + Master’s degree or PhD in engineering (mechanical, electrical, materials, etc.) or science (new college graduate) + Statistical experience such as Weibull, JMP, or familiarity with accelerated test models + Excel at building and maintaining reliability data dashboards on Tableau that drive informed decision making + Proficient in using Python, R or other languages to for extracting data and performing reliability analysis + Experience in adopting emerging AI/ML tools to streamline workflow and identify opportunities for automation + Excellent written and verbal communication skills for audiences ranging from technicians to senior executives + Familiarity with Failure Analysis techniques (Optical Microscopy, X-ray/CT, Scanning Electron Microscopy/Energy Dispersive Spectroscopy, etc.), and the ability to use failure analysis methodology to derive a root cause of failure + Experience in spirited collaboration between multidisciplinary teams of engineers and scientists to solve complex electromechanical / mechanical design challenges **Minimum Qualifications** + Bachelor’s degree in engineering (mechanical, electrical, materials, etc.) or science with 3+ of industry experience + Collaborative attitude and ability to cross-functionally work effectively + Familiarity with failure analysis methodology to identify root cause + Ability to make clear and concise slides/ presentations through Keynote, Excel, etc.
Senior Software Engineer, Site Reliability Engineer Harvey, Inc.Senior Software Engineer, Site Reliability EngineerSan Francisco, CA$200,000–$260,000 / yearAutomate operational tasks and workflows, building tools and processes for capacity planning, graceful rollouts, and safe data access to maintain high reliability and reduce manual intervention. By combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise, we're reshaping how critical knowledge work gets done for decades to come.
Staff Software Engineer, Site Reliability Engineer Harvey, Inc.Staff Software Engineer, Site Reliability EngineerSan Francisco, CA$238,000–$290,000 / yearAutomate operational tasks and workflows, building tools and processes for capacity planning, graceful rollouts, and safe data access to maintain high reliability and reduce manual intervention. 10+ years of experience in Site Reliability Engineering or similar roles supporting production environments, with proven ability to mentor and guide technical teams.
NewReliability Engineer Macpower Digital Assets Edge Private LimitedReliability EngineerSan Leandro, CA$65–$70 / hourEssential Functions and Responsibilities: Conduct root cause analysis (RCA) on high-cost failures, high-downtime, and repetitive issues to identify engineering solutions. Job Summary: Enhance equipment reliability and drive plant-wide continuous improvement across maintenance, operations, safety, quality, and training.
Senior Site Reliability Engineer - Managed Kubernetes LambdaSenior Site Reliability Engineer - Managed KubernetesSan Francisco, CaliforniaOur investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove. *Note: This position requires presence in our San Francisco, San Jose, or Bellevue office location 4 days per week; Lambda’s designated work from home day is currently Tuesday.
Reliability Engineer - Principal / SPE MACOM Technology Solutions Holdings IncReliability Engineer - Principal / SPEMilpitas, CA$98,000–$200,000 / yearJob Description: We are seeking a highly experienced and technically recognized Principal Optoelectronic Device Reliability Engineer to lead reliability engineering and product qualification activities for our portfolio of III-V InP based laser diodes and photodetectors. MACOM sells and distributes products globally via a sales channel comprised of a direct field sales force, authorized sales representatives and leading industry distributors.
Sr. Reliability Engineer - SEM Systems KLA CorporationSr. Reliability Engineer - SEM SystemsMilpitas, CA$167,300–$284,400 / yearEnabling the movement towards advanced chip design, KLA's Global Products Group (GPG), which is responsible for creating all of KLA's metrology and inspection products, is looking for the best and the brightest research scientist, software engineers, application development engineers, and senior product technology process engineers. Our expert teams of physicists, engineers, data scientists and problem-solvers work together with the world's leading technology providers to accelerate the delivery of tomorrow's electronic devices.
Staff Sustaining Reliability Engineer Rivian and Volkswagen Group TechnologiesStaff Sustaining Reliability EngineerPalo Alto, CaliforniaRivian and Volkswagen Group Technologies may use your Candidate Personal Data for the purposes of (i) tracking interactions with our recruiting system; (ii) carrying out, analyzing and improving our application and recruitment process, including assessing you and your application and conducting employment, background and reference checks; (iii) establishing an employment relationship or entering into an employment contract with you; (iv) complying with our legal, regulatory and corporate governance obligations; (v) record keeping; (vi) ensuring network and information security and preventing fraud; and (vii) as otherwise required or permitted by applicable law. Rivian and Volkswagen Group Technologies may share your Candidate Personal Data with (i) internal personnel who have a need to know such information in order to perform their duties, including individuals on our People Team, Finance, Legal, and the team(s) with the position(s) for which you are applying; (ii) Rivian and Volkswagen Group Technologies affiliates; and (iii) Rivian and Volkswagen Group Technologies’ service providers, including providers of background checks, staffing services, and cloud services.
Senior Platform Reliability Engineer Grow TherapySenior Platform Reliability EngineerSan Francisco, California$182,000–$250,000 / yearThis is a hybrid role with the expectation to work onsite from our San Francisco, NYC, or Seattle hub location three days per week (Tuesday, Wednesday, and Thursday) and travel 2–3 times per year (e.g., company and department offsites). Evolving Incident Response Developing and improving incident response practices, from detection to post-incident learning, and helping teams build sustainable on-call and escalation patterns.
Software Reliability Engineer Waymo LLCSoftware Reliability EngineerSan Francisco, CASince its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo DriverThe World''s Most Experienced Driverto improve access to mobility while saving thousands of lives now lost to traffic crashes. The Waymo Driver has provided over ten million rider-only trips, enabled by its experience autonomously driving over 100 million miles on public roads and tens of billions in simulation across 15+ U.S. states.
Robotics Reliability Engineer - Motors, Mechanisms and Electronics MaticRobotics Reliability Engineer - Motors, Mechanisms and ElectronicsMenlo Park, CaliforniaAfter launch, once robots are in real homes, the real data starts coming from every direction: support tickets, customers posting when something breaks, robots sent back for repair, and the motor and sensor telemetry customers consent to share. Your job is to see every way a robot is failing or could fail, keep track of all of it, and turn that signal into action: new tests and test setups where we need them, and clear input on which failure modes are under-prioritized and need engineering bandwidth.
Staff Site Reliability Engineer AssuredStaff Site Reliability EngineerPalo Alto, CaliforniaThe challenges we face are deep and diverse, from creating digital experiences that provide comfort and clarity to claimants at their most stressed and vulnerable to orchestrating large-scale ML-driven decision-making on billions of dollars of claims payments, life at Assured is dynamic, collaborative, and rewarding. Move us beyond component-level availability targets to end-to-end SLOs for the claims journeys insurers and claimants depend on, spanning many services and teams, and extend that into how we measure and report SLA compliance.
Senior Site Reliability Engineer, Compute RobloxSenior Site Reliability Engineer, ComputeSan Mateo, CA$243,290–$295,250 / yearWe’re looking for a skilled Senior Site Reliability Engineer with strong programming skills to help us build Roblox's private cloud, productionize our growing Kubernetes-based infrastructure, and institute reliability best practices across the Roblox Compute team. Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators.
Sr. Quality Assurance & Reliability Engineer Reliable RoboticsSr. Quality Assurance & Reliability EngineerMountain View, CAPartner with design, manufacturing, and test teams through product development, acceptance testing, and quality checkouts to create reliable hardware verification for flight components and systems, including identification of key operating characteristics, prototype development testing for functionality and manufacturability, flight and prototype planning development, and collation of test anomalies and lessons learned. You will be responsible for the strategy and execution of the quality management system, risk mitigation, risk prevention, design reliability, hardware reliability, flight reliability, test validation, and production optimization.
Reliability Engineer, Control Systems, NA Vantage Data CentersReliability Engineer, Control Systems, NASanta Clara, CAFor each of the major systems Electrical, Mechanical, and Controls, the Reliability Engineering team is responsible for ensuring success in the commissioning stages of new construction, evaluating and improving the reliability and performance of existing critical infrastructure, sustaining equipment operational availability through maintenance program design, providing ongoing technical support to the Site Operations Teams, as well as providing systems reliability and maintainability feedback to the Design Engineering teams for future design considerations. Developing and operating across six markets in North America and five markets in Europe, Vantage has evolved data center design in innovative ways to deliver dramatic gains in reliability, efficiency and sustainability in flexible environments that can scale as quickly as the market demands.
Senior/Staff Design Reliability Engineer - Sensor, Compute And EE Systems ZooxSenior/Staff Design Reliability Engineer - Sensor, Compute And EE SystemsFoster City, CA$187,000–$246,000 / yearUsing reliability targets, DFMEA outputs, and physics-of-failure principles, partner with validation engineers on virtual and physical test plans combining multi-variable stressors (thermal, vibration, electrical loading, firmware states). Use current-fleet telemetry and PHM insights to find improvement opportunities and drive corrective action across design, validation, and operations, feeding lessons learned back into next-gen design.
Software Reliability Engineer Nuro IncSoftware Reliability EngineerMountain View, CA$145,830–$219,000 / yearWith years of real-world deployment experience and a flexible, partner-led business model, Nuro is working toward a future where millions of autonomous vehicles powered by our technology help make everyday life safer, easier, and more connected. Powered by the Nuro Driver, our universal autonomy platform enables the global mobility ecosystem to deploy autonomy at scale, from robotaxis and logistics fleets to personal vehicles.
Site Reliability Engineer SRE Thinking Machines Lab IncSite Reliability Engineer SRESan Francisco, CA$350,000–$475,000 / yearWe are scientists, engineers, and builders who've created some of the most widely used AI products, including ChatGPT and Character.ai, open-weights models like Mistral, as well as popular open source projects like PyTorch, OpenAI Gym, Fairseq, and Segment Anything. Tinker is our fine-tuning API that empowers researchers and developers to customize frontier AI to their needs - opening access to capabilities that have previously been concentrated in a handful of labs.
Site Reliability Engineer Obsidian SecuritySite Reliability EngineerPalo Alto, CAWe work closely with Engineering, Quality Engineering, and Customer Support to deliver end-to-end services that bring code to life and maintain our world-class SaaS security platform. The DevOps/SRE team at Obsidian ensures that engineering excellence translates into stable, scalable, and high-performing production systems.
Staff Software Engineer, Developer Experience Crusoe EnergyStaff Software Engineer, Developer ExperienceSan Francisco, CA$208,725–$253,000 / yearWe're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved - people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services. Language & Version Control Mastery: Expertise in modern programming languages (specifically Go) and advanced proficiency in Git-based workflows (GitLab/GitHub).