Site Reliability Engineer, AI & Agentic Systems ServiceLink IP Holding Co LLCSite Reliability Engineer, AI & Agentic SystemsPlano, TXThe ideal candidate will leverage Azure-native AI services and agentic systems to reduce toil, improve incident response, and enable intelligent operations-while also driving performance testing practices to validate system resilience under load. Design, develop, and execute performance testing strategies for distributed systems and microservices, including load testing, stress testing, soak testing, and capacity planning.
NewSite Reliability Engineer Intern CopartSite Reliability Engineer InternDallas, TexasCopart, Inc. a technology leader and the premier online vehicle auction platform globally, with over 200 facilities located across the world, Copart links vehicle sellers to more than 750,000 buyers in over 190 countries. Copart, a global leader in online vehicle auctions, is seeking a highly skilled and proactive Site Reliability Engineer Intern to join our SRE team at our Dallas location.
IT Site Reliability Engineer - API Management Platforms Texas Instruments IncIT Site Reliability Engineer - API Management PlatformsDallas, TXAs an IT Site Reliability Engineer within the Enterprise Platforms team, you will serve as the primary technical platform owner for TI''s Apigee Edge private cloud environment - the backbone of TI''s API management and integration strategy - while also providing platform SRE support across a broader portfolio of automation and DevOps tooling used throughout the enterprise. Multi-Platform SRE: Provide site reliability engineering support for other automation and DevOps platforms within the ITS DevOps Platforms portfolio - which may include CI/CD tools, artifact management, source control, integration middleware, and automation orchestration platforms.
Site Reliability Engineer Cisco Systems IncSite Reliability EngineerRichardson, TX$152,500–$219,200 / yearPartner with other engineering teams, product management, and business partners across multiple teams and time zones to understand platform dependencies, seek opportunities for improvement, and deliver solutions that enhance reliability and reduce operational overhead. The team operates with a high degree of autonomy, giving engineers the opportunity to drive both critical initiatives and grassroots improvements that solve real operational challenges.
Staff Hardware Reliability Engineer Shield AI IncStaff Hardware Reliability EngineerDallas, TX$158,542–$237,812 / yearYoull participate in design reviews and FMEA activities, shape material selection and manufacturing requirements, analyze test and field data using reliability modeling tools, and help develop corrective actions and process improvements that elevate hardware reliability across the program. You will lead environmental and stress testing efforts, including temperature cycling, vibration, HALT, and HASS, conduct failure analysis and materials characterization, and analyze root cause investigations for manufacturing non-conformances and field returns.
Site Reliability Engineer II General Motors Financial Company, Inc.Site Reliability Engineer IIArlington, TXContinuous Improvement: Design and implement automated pipelines, security controls, and monitoring solutions to minimize manual intervention and maximize uptime. Experience working in Agile Scrum teams with demonstrated success leading improvements (getting better/faster/happier).
Technical Lead Site Reliability & Performance Engineer ECLAROTechnical Lead Site Reliability & Performance EngineerAddison, TX$80 / hourPlay a critical role in ensuring the reliability, performance, security, and operational health of Client's externally facing digital platforms through ownership of key platform services including observability, domain services, DNS, certificate management, CDN technologies, web application firewalls, cloud platform infrastructure, and Infrastructure-as-Code (IaC) solutions. This position provides the opportunity to work with globally distributed engineering teams operating in a follow-the-sun support model, influence platform strategy across multiple regions, partner with technical leaders worldwide, and develop toward a future Technical Manager role.
Sr. Site Reliability Engineer - Observability AppFolioSr. Site Reliability Engineer - ObservabilityDallas, TX$138,400–$173,000 / yearAlong with your team, you'll ensure all aspects of our shared product spaces have a plan to address any opportunities for exception reporting, capacity planning, monitoring and alerting, backups, runbooks, configuration management, DDoS protection, infrastructure as code, and disaster recovery. Proven ability to diagnose and monitor performance and reliability issues across the stack: relational databases, web servers, networking, OS, containers, load balancers, etc.
Product Reliability Engineer Motorola Solutions IncProduct Reliability EngineerAllen, TX$100,000–$120,000 / yearProduct Reliability bridges the gap between development engineering, manufacturing, and the customer by understanding fielded product or customer issues then bringing all aspects of the problem together to provide actionable direction for full root cause and resolution of issues. Performing reliability testing including Electrical & Mechanical Stress Testing, Shock/Drop/Vibration Testing, and Environmental testing to reproduce field failures.
Reliability Engineer HROC Heidelberg MaterialsReliability Engineer HROCIrving, Texas$107,480–$143,307 / yearDevelop and implement effective spare parts management strategies across the plants supported, ensuring identification and availability of critical spare parts, optimizing inventory levels and minimizing costs. Monitor equipment health using condition monitoring technologies (vibration, thermography, oil analysis, and other predictive tools), analyze reliability data and KPIs, and translate insights into actionable maintenance plans.
Reliability Engineer Hroc Heidelberg MaterialsReliability Engineer HrocIrving, TX$107,480–$143,307 / yearDevelop and implement effective spare parts management strategies across the plants supported, ensuring identification and availability of critical spare parts, optimizing inventory levels and minimizing costs. Monitor equipment health using condition monitoring technologies (vibration, thermography, oil analysis, and other predictive tools), analyze reliability data and KPIs, and translate insights into actionable maintenance plans.
Software Engineer Lead - Site Reliability Engineering Center The PNC Financial Services Group IncSoftware Engineer Lead - Site Reliability Engineering CenterFarmers Branch, TX$86,250–$158,125 / yearIn addition, PNC generally provides the following paid time off, depending on your eligibility: maternity and/or parental leave; up to 11 paid holidays each year; 9 occasional absence days each year, unless otherwise required by law; between 15 to 25 vacation days each year, depending on career level; and years of service. This position is subject to the requirements of Section 19 of the Federal Deposit Insurance Act (FDIA) and, for any registered role, the Secure and Fair Enforcement for Mortgage Licensing Act of 2008 (SAFE Act) and/or the Financial Industry Regulatory Authority (FINRA), which prohibit the hiring of individuals with certain criminal history.
Sr Principal Engineer, NPD Quality and Reliability ON Semiconductor CorpSr Principal Engineer, NPD Quality and ReliabilityAllen, TXWith a highly differentiated and innovative product portfolio, onsemi creates intelligent power and sensing technologies that solve the world's most complex challenges and leads the way in creating a safer, cleaner, and smarter world. In this highly visible role, the candidate will work with multiple business units, multidisciplinary teams and interface with Tier 1 customers to ensure we exceed customers' needs with high quality, reliable products.
Senior Reliability and Maintainability Engineer RTX CorpSenior Reliability and Maintainability EngineerRichardson, TXQualifications You Must Have: Typically Requires a bachelor's degree in science, Technology, Engineering or Math (STEM) and minimum 5 years of relevant experience conducting RCCA investigations to identify root cause, implement corrective action, and monitor to verify improvement. RTX considers several factors when extending an offer, including but not limited to, the role, function and associated responsibilities, a candidate's work experience, location, education/training, and key skills.
Sr. Systems Engineer, Site Reliability American Heart AssociationSr. Systems Engineer, Site ReliabilityDallas, TexasRequired Skills: Analytical Thinking; Application Development; Attention to Detail; Collaboration; Database Structures; Effective Communication; Initiative; IT Architecture; IT Environment; Network Architecture; Network Engineering; Network Security; Operating System; Problem Solving; Programming; Software Development Life Cycle; Software Engineering; Structured Query Language (SQL); System Testing; Time Management; DevOps; Infrastructure as Code; Multi-Cloud Ops Management. The American Heart Association’s 2028 Goal: Building on over 100 years of trusted leadership in cardiovascular and brain health, by 2028 the Association will drive breakthroughs and implement proven solutions in science, policy, and care for healthier people and communities.
Sr Software Engineer (Site Reliability) Austin or Dallas, TX H.E. Butt Grocery CoSr Software Engineer (Site Reliability) Austin or Dallas, TXDallas, TXQualifications & Key Requirements: Work Experience: 5+ years experience designing, analyzing, developing, or troubleshooting distributed systems 3+ years of SRE experience managing Google Kubernetes Engine (preferred), K8s, or AWS environments 2+ years of Java (Spring) programming experience preferred 3+ years of using Terraform to maintain cloud infrastructure 3+ years of CI Pipeline experience with either Gitlab Pipelines, or GitHub Actions Experience with tools such as Gitlab, JIRA, Slack, Confluence and Intellij is preferred Experience with microservices architecture patterns Experience working with PostgreSQL, Kubernetes, Docker, Linux, GCP, Terraform, and APIs using REST and GraphQL Experience working with monitoring and visualization tools such as Datadog, Grafana, or New Relic Strong proficiency with scripting languages such as Python, Ruby, Groovy, Bash Proven track record of researching, understanding, and effectively applying Scalability and High Availability principles Knowledge/Skills/Abilities: Advanced knowledge in system and data architecture, data modeling, and design and capable of architecting and designing at the application or service level using well-accepted design patterns - Able to review platform designs for strength of engineering solutions, namely performance, sustainability, and iterative development potential. Collaborates with development teams to design service architectures, software platforms and frameworks, capacity planning, release plans and launch reviews Monitors internal and vendor service level objectives (SLOs) and agreements (SLAs); identifies / resolves SLO / SLA gaps Serves as technical subject matter expert (SME) for cross-functional engineering Teams; assists with / troubleshoots systems-related issues and maintenance The responsibilities and essential functions outlined above describe the general nature and level of work assigned to this position.
Senior Engineer - Site Reliability Engineering London Stock Exchange Group PlcSenior Engineer - Site Reliability EngineeringAllen, TXStrong experience with Azure, including services such as AKS, Azure Container Apps, Virtual Machines, virtual networking (VNet), Azure Active Directory (Entra ID), and Azure managed services. Role Profile: We are evolving our Site Reliability Engineering capabilities to strengthen reliability, observability, security, and operational excellence across our Markets and Risk Intelligence division.
Senior Maintenance Reliability Engineer Teasdale FoodsSenior Maintenance Reliability EngineerCarrollton, TexasReporting directly to the Senior Director of Engineering, the Senior Maintenance Reliability Engineer will play a crucial role in ensuring the optimal functioning of our production facilities across multiple sites. Utilize reliability engineering principles, such as Failure Mode and Effects Analysis (FMEA) and Reliability Centered Maintenance (RCM), to optimize maintenance strategies and improve equipment performance.
Reliability Excellence (RX) Engineer Sherwin-Williams CoReliability Excellence (RX) EngineerGarland, TXThe wage range listed for this role takes into account the wide range of factors considered in making compensation decisions including skill sets; experience and training; licensure and certifications; and other business and organizational needs. Please be aware, Sherwin-Williams recruiting team members will never request a candidate to provide a payment, ask for financial information, or sensitive personal information like national identification numbers, date of birth, or bank account numbers during the application process.
Reliability and Maintainability Engineer II RTX CorpReliability and Maintainability Engineer IIRichardson, TXQualifications You Must Have: Typically Requires a Bachelor's degree in Science, Technology, Engineering or Math (STEM) and minimum 2 years of relevant experience supporting the program in the Failure Reporting And Corrective Action (FRACAS) process for product production and/or sustainment. RTX considers several factors when extending an offer, including but not limited to, the role, function and associated responsibilities, a candidate's work experience, location, education/training, and key skills.