Senior Software Engineer, Reliability Tin CanSenior Software Engineer, ReliabilitySeattle, WashingtonBonus: production FreeSWITCH, Kamailio, or carrier SIP trunking; load-testing a system through a seasonal peak; Postgres at scale; operating a consumer hardware fleet; or observability tooling like Grafana. Our calls run on FreeSWITCH and Kamailio in AWS, gated by a set of Lambda services, with an activation flow every family walks through the morning they open the box.
Safety and Reliability Sr. Systems Engineer Blue Origin Enterprises LPSafety and Reliability Sr. Systems EngineerSeattle, WA$145,188–$203,263.20 / yearand/or transports placardable amounts of hazardous materials by ground in any vehicle on a public road while in commerce, may be subject to additional Federal Motor Carrier Safety Regulations including: Driver Qualification Files, Medical Certification (obtained before onboarding), Road Test, Hours of Service, Drug and Alcohol Testing, vehicle inspection requirements, CDL requirements (if applicable) and hazardous materials transportation/shipping training. Additionally, this role requires understanding and supporting requirements analysis and traceability, functional analysis, systems analysis, safety analysis, technical performance monitoring, developing strategies and approaches for specialty engineering (operability, reliability, maintainability, supportability) and system product life cycle processes.
Sr. Staff Engineer Software, Infrastructure Reliability (Chronosphere) Palo Alto Networks IncSr. Staff Engineer Software, Infrastructure Reliability (Chronosphere)Seattle, WARemote$126,000–$204,500 / yearAnalytical Debugging & Quality: The ability to debug your own code efficiently using logs and tests, while proactively identifying edge cases (like nulls or limits) during the design phase. For candidates who receive an offer at the posted level, the starting base salary (for non-sales roles) or base salary + commission target (for sales/com-missioned roles) is expected to be the annual range listed below.
Salesforce Reliability & Migration Engineer: 26-01344 Akraya Inc.Salesforce Reliability & Migration Engineer: 26-01344Bellevue, WA$60–$64.75 / hourMost recently, we were recognized Stevie Employer of the Year 2025, SIA Best Staffing Firm to work for 2025, Inc 5000 Best Workspaces in US (2025 & 2024) and Glassdoor's Best Places to Work (2023 & 2022)! As Talent solutions provider for Fortune 100 Organizations, Akraya's industry recognitions solidify our leadership position in the IT staffing space.
Research Engineer Graduate (AI Training Systems Reliability & Performance - Seed Infra) - 2026 Start (PhD) Beijing ByteDance Technology Co LtdResearch Engineer Graduate (AI Training Systems Reliability & Performance - Seed Infra) - 2026 Start (PhD)Seattle, WAIndividuals who are completing or have recently completed a PhD degree in Computer Science, Electrical Engineering, Electrical and Computer Engineering, Physics, Mathematics, or a related discipline. Improve the reliability and performance of large-scale training systems across pre-training, fine-tuning, evaluation, and inference.
Senior Reliability Engineer, Amazon LEO Amazon.com IncSenior Reliability Engineer, Amazon LEORedmond, WAYou will play an important role across key dimensions of a product's lifecycle: design trades, product development and release, manufacturing process qualification, supply chain reliability, environmental stress screens, and on-orbit fault and failure investigations. The ideal candidate will have a strong background in electronics manufacturing processes, hardware reliability concepts, and reliability test practices, with a focus on ensuring the functional reliability and performance of satellite hardware systems in the harsh environment of space.
ASE Compute - Site Reliability Engineering (SRE) Manager Apple IncASE Compute - Site Reliability Engineering (SRE) ManagerSeattle, WALead, grow, and mentor a team of Site Reliability Engineers focused on large-scale Kubernetes and compute infrastructure Own the reliability, availability, and performance of mission-critical cloud platform services Drive incident response, post incident reviews, and systemic improvements that reduce operational toil Partner with software engineering and architecture teams to influence system design for reliability, scalability, and operability Establish and refine SRE practices including SLOs, error budgets, capacity planning, and change management Champion automation - eliminate manual processes through tooling and self-service capabilities Manage on-call rotations and ensure sustainable, well-supported operational coverage Communicate clearly across teams to build a culture of visibility, transparency, and shared ownership5+ years of engineering management experience leading infrastructure or SRE teams Deep experience operating large-scale, multi-tenant Kubernetes environments in production Strong systems background - comfortable troubleshooting across the full stack (network, OS, container runtime, application) Experience with configuration management at scale (Puppet, Ansible, or equivalent) Track record of building high-performing teams through coaching, clear expectations, and psychological safety Demonstrated ability to drive cross-functional initiatives to completion Strong written and verbal communication skillsExperience with third-party cloud platforms (AWS, GCP, or Azure) Familiarity with bare-metal provisioning and lifecycle management at datacenter scale Experience with Java, Go, or Python services in production Understanding of cloud-native observability (Prometheus, Thanos, Splunk, or similar) CNCF Certified Kubernetes Administrator (CKA) or equivalent hands-on certification Experience running infrastructure as an internal managed service with defined SLAs. You will set the technical direction for reliability and operational excellence while mentoring engineers, driving automation, and partnering closely with software and infrastructure teams to ship improvements that matter.
Reliability Test Specialist, Amazon LEO Amazon.com IncReliability Test Specialist, Amazon LEORedmond, WAAmazon Leo is an initiative to launch a constellation of Low Earth Orbit satellites that will provide low-latency, high-speed broadband connectivity to unserved and underserved communities around the world. Export Control Requirement: Due to applicable export control laws and regulations, candidates must be a U.S. citizen or national, U.S. permanent resident (i.e., current Green Card holder), or lawfully admitted into the U.S. as a refugee or granted asylum.
__Safety & Reliability-May 2022 Keltia Design Inc__Safety & Reliability-May 2022Kirkland, WAThis includes preparing the Maintenance and Reliability teams, developing, and implementing the Maintenance and Reliability Asset Management Process, the CMMS, the spare parts inventory, the mechanical integrity program and developing the comprehensive plant equipment strategies. This position requires a minimum of 15+ years' experience in roles that develop and implement maintenance and asset integrity programs & processes of a PSM covered process Must exhibit a strong track record of completion and success in delivery of commissioning and projects.
Software Engineering Manager, GPU Reliability Google LLCSoftware Engineering Manager, GPU ReliabilitySeattle, WATeams work all across the company, in areas such as information retrieval, artificial intelligence, natural language processing, distributed computing, large-scale system design, networking, security, data compression, user interface design; the list goes on and is growing every day. With technical and leadership expertise, you manage engineers across multiple teams and locations, a large product budget and oversee the deployment of large-scale projects across multiple sites internationally.
Sr. Hardware Reliability Eng, Prime Air Amazon.com IncSr. Hardware Reliability Eng, Prime AirSeattle, WAYou will develop reliability requirements, test methodologies, and environmental qualification programs while partnering with engineering teams to improve drone hardware reliability throughout the product lifecycle. Working across drone system development and multiple engineering disciplines, you will lead design-for-reliability activities, verification planning, environmental and accelerated-life testing, and complex failure investigations.
Engineering Manager, Cloud Network Reliability Apple IncEngineering Manager, Cloud Network ReliabilitySeattle, WAApple Cloud Networking team builds and operates large-scale, software-defined networking platforms that enable secure, resilient, and highly available multi-cloud connectivity with a global footprint. We are seeking an experienced and visionary Reliability Engineering Manager to lead and grow a team of engineers focused on ensuring the availability, performance, scalability, and resiliency of Apple's global network services.
AWS Site Reliability Engineering With Dynatrace Akt LLCAWS Site Reliability Engineering With DynatraceSeattle, WARemoteConducting Post-incident reviews (PIRs): Evaluating Incidents after resolution to build work book and automation to avoid/solve them when happen again. Monitoring tools: Should have working experience on Dynatrace, AppDynamics, DataDog, Pager Duty and Cloud Watch.
SRE Architect, Ai-Powered Reliability WEX Inc.SRE Architect, Ai-Powered ReliabilitySeattle, WA$200,600–$250,400 / yearDefine and lead WEX's AI-Powered Reliability Engineering strategy, driving adoption of SRE agents across the software lifecycle-from design and development through deployment and operations, to improve reliability, automation, and operational efficiency. Serve as a technical advisor to engineering leaders and architects across WEX, reviewing system designs for reliability risk, providing guidance on high-availability and low-latency architecture patterns, and advising on operational tradeoffs.
Digital - Senior Data Engineer / Data Engineer, Digital Technology AritziaDigital - Senior Data Engineer / Data Engineer, Digital TechnologySeattle, WA$100,000–$300,000 / yearWith a Global Support Campus in Vancouver and additional Support Office hubs in major cities across North America - including Toronto, New York, Los Angeles and Seattle - our workplace network is strategically designed to support our People and serve our Clients a frictionless experience. As the Senior Data Engineer / Data Engineer, you will bridge data modeling and analytics to transform raw data into high-quality, reliable data assets, while generating actionable insights that answer critical business questions.
Systems Engineer III (PCB): 26-00744 Akraya Inc.Systems Engineer III (PCB): 26-00744Redmond, WA$75–$79 / hourJob Summary:The Build Reliability Engineer will ensure the reliability and quality of complex Optical Inter-Satellite Link (OISL) products from the design phase through to manufacturing, focusing on Design for Manufacturing (DfM), Quality (DfQ), and Reliability (DfR). Most recently, we were recognized Stevie Employer of the Year 2025, SIA Best Staffing Firm to work for 2025, Inc 5000 Best Workspaces in US (2025 & 2024) and Glassdoor's Best Places to Work (2023 & 2022)!
NewSoftware Engineer I ChewySoftware Engineer IBellevue, WashingtonThis team builds and supports services that power Chewy’s post-purchase customer experience, helping customers understand the status of their orders and receive accurate, timely information throughout the fulfillment and delivery lifecycle. We offer parental leave, family services benefits, backup dependent care, flexible spending accounts, telemedicine, pet adoption reimbursement, employee assistance program, and many discounts including 10% off pet insurance and 20% off at Chewy.com.
PRINCIPAL RF/MICROWAVE ENGINEER Kapta Space CorpPRINCIPAL RF/MICROWAVE ENGINEERSeattle, WA$185,000–$230,000 / yearTo conform to U.S. Government space technology export regulations, including the International Traffic in Arms Regulations (ITAR), Kapta Employees must be U.S. citizens, lawful U.S. permanent residents (i.e., current Green Card holder), or lawfully admitted into the U.S. as a refugee or granted asylum, or be eligible to obtain the required authorizations from the U.S. Department of State or Commerce as applicable. As a critical contributor to payload development, this role operates within a highly multidisciplinary environment requiring engineering trade studies and design decisions across RF performance, DC power, thermal management, reliability, manufacturability, environmental robustness, and overall payload performance.
NewPrincipal Penetration Test Engineer ArcfieldPrincipal Penetration Test EngineerHome, VirginiaFull timeThere are differentiating factors that can impact a final salary/hourly rate, including, but not limited to, Contract Wage Determination, relevant work experience, skills and competencies that align to the specified role, geographic location (For Remote Opportunities), education and certifications as well as Federal Government Contract Labor categories. As an industry-leading solutions provider in digital engineering and model-based systems engineering (MBSE), the company delivers MBSE-as-a-Service, integrated digital engineering environment deployments, training and consulting to both commercial and public sector customers.
NewGNC Engineer Cowboy SpaceGNC EngineerSeattle, WashingtonOur modular, scalable satellites collect sunlight in Low Earth Orbit to enable multiple integrated applications: transmitting energy via infrared lasers (space-to-earth and space-to-space), powering on-orbit high-performance computing clusters (GPU/TPU), and providing secure, high bandwidth optical data transport. Export Control Requirement: To conform to U.S. Government space technology export regulations, including the International Traffic in Arms Regulations (ITAR), applicants must be a U.S. citizen, lawful permanent resident of the U.S., protected individual as defined by 8 U.S.C. 1324b(a)(3), or eligible to obtain the required authorizations from the U.S. Department of State.