Lead Software Engineer, Middleware Reliability Engineering Visa Technology and Operations LLCLead Software Engineer, Middleware Reliability EngineeringFoster City, CA$192,300–$307,600 / yearFamiliarity with AI/ML frameworks and chatbot integrations for operational automation Middleware Expertise • Understanding of middleware technologies (Message Queues, Service Bus, API • Gateways) • Experience with application servers (Tomcat, JBoss, WebSphere) • Knowledge of integration platforms and their deployment patterns • Proficiency in troubleshooting and performance optimization Education • Bachelor's or Master's degree in Computer Science or related field, or equivalent experience • We value hands-on experience and continuous learning over specific degrees. • Solid understanding of Linux/Unix systems, networking protocols, certificate management, secret management, system design, cloud platforms (AWS, • Azure, GCP), and containerization (Kubernetes, Docker • Proficiency with monitoring tools (Prometheus, Grafana, Datadog, etc.), logging systems (ELK stack, Splunk), and tracing tools (Jaeger, Zipkin).
Supplier Quality Engineer Raas Info Solutions Pvt LtdSupplier Quality EngineerFremont, CA$50–$54.99 / hourContractorThe role involves evaluating thermal/mechanical designs, supplier manufacturing capability, reliability test reports, and process quality to ensure products meet performance, manufacturability, serviceability, and compliance requirements. Job Summary: Responsible for supplier quality management of data center rack systems and advanced cooling solutions.
NewDevOps Platform Development Engineer (Bilingual Mandarin) ComriseDevOps Platform Development Engineer (Bilingual Mandarin)Palo Alto, CA$150,000–$240,000 / yearFull timeCollaborate closely with platform engineering teams in China as well as overseas infrastructure, networking, security, and business teams to manage implementation schedules, dependencies, and risks, while providing feedback to improve platform capabilities for international deployment scenarios. Participate in the day-to-day operations and incident response of overseas production environments, quickly troubleshoot deployment, infrastructure, and platform integration issues, and drive root cause analysis and long-term corrective actions.
NewSupplier Quality Engineer JobotSupplier Quality EngineerSunnyvale, CA$50–$85 / hourThe ideal candidate will have experience supporting suppliers in electro-mechanical and electronics manufacturing environments, including printed circuit board assemblies (PCBAs), manufactured components, and complex hardware assemblies. Information collected and processed as part of your Jobot candidate profile, and any job applications, resumes, or other information you choose to submit is subject to Jobot's Privacy Policy, as well as the Jobot California Worker Privacy Notice and Jobot Notice Regarding Automated Employment Decision Tools which are available at jobot.com/legal.
NewSenior AI Engineer JobotSenior AI EngineerSan Francisco, CA$225,000–$280,000 / yearInformation collected and processed as part of your Jobot candidate profile, and any job applications, resumes, or other information you choose to submit is subject to Jobot's Privacy Policy, as well as the Jobot California Worker Privacy Notice and Jobot Notice Regarding Automated Employment Decision Tools which are available at jobot.com/legal. The product sits at the intersection of AI, automation, and real world impact, using LLMs and agentic systems to replace hours or weeks of manual work with intelligent, scalable software.
NewSupport Engineer 3- Identity/Access Specific JobotSupport Engineer 3- Identity/Access SpecificSan Francisco, CA$70–$120 / hourOperating at the intersection of platform engineering, automation, and security, the team partners with high-growth enterprises to create repeatable, reliable, and compliant cloud environments across complex multi-cloud ecosystems. Information collected and processed as part of your Jobot candidate profile, and any job applications, resumes, or other information you choose to submit is subject to Jobot's Privacy Policy, as well as the Jobot California Worker Privacy Notice and Jobot Notice Regarding Automated Employment Decision Tools which are available at jobot.com/legal.
NewSenior Software Engineer JobotSenior Software EngineerSan Francisco, CA$225,000–$450,000 / yearInformation collected and processed as part of your Jobot candidate profile, and any job applications, resumes, or other information you choose to submit is subject to Jobot's Privacy Policy, as well as the Jobot California Worker Privacy Notice and Jobot Notice Regarding Automated Employment Decision Tools which are available at jobot.com/legal. We’re building the infrastructure layer for private capital markets — combining modern software systems, AI-native workflows, and deep market intelligence to transform how private market transactions happen.
NewData Engineer - Sr. Consultant level Visa Technology and Operations LLCData Engineer - Sr. Consultant levelFoster City, CA$169,100–$270,800 / yearVisa is a world leader in payments technology, facilitating transactions between consumers, merchants, financial institutions and government entities across more than 200 countries and territories, dedicated to uplifting everyone, everywhere by being the best way to pay and be paid. Masters, MBA, JD, MD) or 2 years of work experience with a PhD Preferred Qualifications • 9 or more years of relevant work experience with a Bachelor Degree or 7 or more relevant years of experience with an Advanced Degree (e.g.
Staff Data Engineer Visa Technology and Operations LLCStaff Data EngineerFoster City, CA$146,200–$233,700 / yearVisa is a world leader in payments technology, facilitating transactions between consumers, merchants, financial institutions and government entities across more than 200 countries and territories, dedicated to uplifting everyone, everywhere by being the best way to pay and be paid. At Visa, you'll have the opportunity to create impact at scale — tackling meaningful challenges, growing your skills and seeing your contributions impact lives around the world.
Lead Software Engineer Visa Technology and Operations LLCLead Software EngineerFoster City, CA$192,300–$307,600 / yearVisa is a world leader in payments technology, facilitating transactions between consumers, merchants, financial institutions and government entities across more than 200 countries and territories, dedicated to uplifting everyone, everywhere by being the best way to pay and be paid. Essential Functions Provide deep technical leadership within Shared Services Product Development, with a strong understanding of how shared platforms enable downstream product innovation.
Network DevOps Engineer Epitec StaffingNetwork DevOps EngineerPalo Alto, CASummary Our client is seeking a hands-on Network DevOps Engineer to support network automation, cloud infrastructure deployment, and multi-cloud connectivity initiatives across a large-scale enterprise environment. The ideal candidate enjoys solving complex infrastructure challenges, building automated solutions, and supporting highly available environments across multiple cloud platforms.
Principal Infrastructure Engineer Sidram techPrincipal Infrastructure EngineerSF, CAContractorHiring #InfrastructureEngineer #PrincipalEngineer #PlatformEngineering #CloudEngineering #Kubernetes #Docker #Terraform #CloudFormation #AWS #Azure #GCP #DevOps #SRE #CloudNative #CICD #DataEngineering #MachineLearning #Snowflake #Databricks #Istio #Linkerd #Prometheus #ELK #HybridJobs #SanFranciscoJobs #CaliforniaJobs #ContractJobs #C2C #W2 #USC #H1B #H4EAD #TechJobs #NowHiring #SidRamTech. We are looking for an experienced Principal Infrastructure Engineer to design and deploy scalable, secure infrastructure for Nextdata OS across multi-cloud environments.
Software Engineer, Deployment Infrastructure VercelSoftware Engineer, Deployment InfrastructureSan Francisco, CAWe’re looking for engineers who want to own high-scale systems end-to-end, thrive on improving the developer experience, and are driven to have a lasting impact on how modern software gets built and delivered. • Own and evolve the infrastructure that powers Vercel’s build and deployment lifecycle — from handling webhooks to designing resilient database schemas and building scalable APIs.
Data Engineer II Pinnacle Technical ResourcesData Engineer IISunnyvale,, CaliforniaContractorThe specific compensation for this position will be determined by several factors, including the scope, complexity, and location of the role, as well as the cost of labor in the market; the skills, education, training, credentials, and experience of the candidate; and other conditions of employment. The ideal candidate will be responsible for designing, building, and maintaining scalable data pipelines and systems to support our organization's data needs.
NewStaff Platform Engineer AmplitudeStaff Platform EngineerSan Francisco, CAAs a Staff Platform Engineer, youll set technical direction for the platform across teams, lead our highest-complexity and highest-leverage initiatives, and shape a platform where AI agents are first-class users alongside humans: kicking off deploys, opening pull requests against infrastructure, and triaging incidents, so a single engineer can get the throughput of a team. • Set technical direction — shape platform and domain-level technical strategy that improves developer experience, reliability, security, and cost, and lead the high-complexity, cross-cutting initiatives that deliver it with measurable impact for the organization.
NewSoftware Engineer, Distributed Systems FigmaSoftware Engineer, Distributed SystemsSan Francisco, CAJob level and actual compensation will be decided based on factors including, but not limited to, individual qualifications objectively assessed during the interview process (including skills and prior relevant experience, potential impact, and scope of role), market demands, and specific work location. As a Software Engineer on our Infrastructure team, you’ll help design, build, and operate the systems that power our real-time collaborative design tools used by millions of people worldwide.
NewSenior Hardware Manufacturing Engineer MicrosoftSenior Hardware Manufacturing EngineerMountain View, CA$119,800–$234,700 / yearThe Hardware Manufacturing Engineering Group is a team that collaborates with Microsoft’s suppliers, engineering design teams, business managers, and supply chain managers to ensure that the newest hardware that powers Microsoft’s cloud services are available to our customers on time and at the highest levels of quality. Masters Degree in Manufacturing, Material, Mechanical, Electrical, Thermal, or Industrial Engineering, or related field AND 4+ years experience in a manufacturing environment /repair OR Bachelors Degree in Manufacturing, Material, Mechanical, Electrical, Thermal, or Industrial Engineering, or related field AND 5+ years experience in a manufacturing environment/repair OR equivalent experience.
NewSr. Product Engineer – Web Services EsriSr. Product Engineer – Web ServicesSan Francisco, CARemote$93,600–$159,328 / yearExperience working with ArcGIS applications and technologies (such as ArcGIS Hub, ArcGIS Map Viewer, ArcGIS REST APIs, ArcGIS API for Python, ArcGIS Maps SDKs). You will act as a service owner and trusted decision‑maker, shaping priorities, balancing tradeoffs, and ensuring our web services remain reliable, scalable, and valuable to customers over time.
NewSenior Software Engineer, Data Management AmplitudeSenior Software Engineer, Data ManagementSan Francisco, CAYou will help design and build systems that govern the data lifecycle, including complex ingestion planning, semantic enrichment, and data observability — spanning backend infrastructure, APIs, and the product surfaces customers interact with directly. Our values of growth mindset, ownership, and humility are core to the way we work: we’re tenacious in the face of challenges, we take the initiative to solve problems that drive our shared success, and we operate from a place of empathy and openness, seeking to understand many points of view.
Staff Software Engineer - Payments RipplingStaff Software Engineer - PaymentsSan Francisco, CAPayroll Payments is the team within the Payments organization that owns converting payroll data into payment instructions - the what, who, and when of every debit and credit — and owning the admin, employee, contractor, Ops, and Support experience around those payments. • Raise the engineering bar for the entire organization by mentoring senior and junior engineers, leading in-depth design reviews, and championing best practices for code quality, testing, and observability.
Staff Software Engineer - Payroll Data RipplingStaff Software Engineer - Payroll DataSan Francisco, CADefine the Unified Query Layer: Partner with other Staff and Principal engineers to design and evolve a cohesive data query layer that spans batch and real-time processing, reducing fragmentation and enabling consistent access patterns. • Partner Cross-Functionally: Work closely with Product, Platform, AI, and other Payroll teams to translate product needs into robust platform capabilities while influencing roadmaps with data-informed tradeoffs.
Software Engineer, Next.js VercelSoftware Engineer, Next.jsSan Francisco, CAThis role is a strong fit for someone with a background in framework development, developer tooling, or product infrastructure: someone comfortable moving across the stack, working in public, and turning ambiguous problems into clear technical direction. In this role, you will work across framework internals, rendering, caching, routing, build systems, developer tooling, documentation, platform integration, and user-facing product behavior.
Senior Software Engineer II - AI Tooling and Interfaces DoubleVerifySenior Software Engineer II - AI Tooling and InterfacesSan Francisco, CA$107,000–$193,000 / yearSoftware Engineer II in AI Tooling and Interfaces, youll design and build systems purpose-built for AI agents, automations, and decisioning engines that drive real outcomes for our clients and internal teams. If you enjoy turning high-volume data into clean, powerful systems—and you want your work to drive real business outcomes rather than just dashboards—this role puts you right at the center of the action.
Senior Software Engineer, Backend - Financial Product RipplingSenior Software Engineer, Backend - Financial ProductSan Francisco, CAOur team builds and maintains the core financial products that enable businesses to operate seamlessly, covering areas such as: • Global Payroll: Developing the systems for calculating, filing, and processing payroll across multiple countries, including tax logic, compliance, and direct money movement infrastructure. This platform delivers secure, fault-tolerant, scalable primitives for money movement, billing, reconciliation, compliance, and financial intelligence, enabling rapid launch of financial products with minimal effort, and zero disbursement failures or manual intervention.
(Agile1)IT Solutions Engineer, Principal Axelon(Agile1)IT Solutions Engineer, PrincipalOakland, CA$100–$125 / hourThe Lead Software Engineer owns mobile and backend architecture, guides implementation decisions, coordinates integration across Client systems, and leads delivery through technical direction, mentorship, and increasing team leadership responsibility. Lead Software Engineer Mobile & Platforms to provide senior technical leadership for the design, development, and evolution of a new customerfacing mobile application platform.
Rotating Equipment Reliability Engineer PBF Energy IncRotating Equipment Reliability EngineerMartinez, CA$88,424.81–$156,910.67 / yearThe successful candidate will work with Machinery Inspectors on long-term objectives and day-to-day troubleshooting of refinery machinery which includes pumps, turbines, compressors, gearboxes, fans/blowers, etc. Rotating Equipment Reliability Engineer PBF Energy Inc. (NYSE:PBF) is one of the largest independent refiners in North America, operating, through its subsidiaries, oil refineries and related facilities.
NewRefinery Rotating Equipment Reliability Engineer Npa WorldwideRefinery Rotating Equipment Reliability EngineerMartinez, CA$88,000–$155,000 / yearThis full-time role focuses on maintaining the reliability of refinery machinery and working closely with Machinery Inspectors to troubleshoot issues. The ideal candidate will hold a Bachelor's degree in Mechanical Engineering and have over 3 years in refinery or petrochemical environments.
Senior Hardware Reliability Engineer SamsaraSenior Hardware Reliability EngineerSan Francisco, CA$204,000–$240,000 / yearYou should apply if: You want to impact the industries that run our world: Your efforts will result in real-world impact-helping to keep the lights on, get food into grocery stores, reduce emissions, and most importantly, ensure workers return home safely. Working at Samsara means you'll help define the future of physical operations and be on a team that's shaping an exciting array of product solutions, including Video-Based Safety, Vehicle Telematics, Apps and Driver Workflows, and Equipment Monitoring.
Senior Staff Site Reliability Engineer, AViD, YouTube Ads Google LLCSenior Staff Site Reliability Engineer, AViD, YouTube AdsMountain View, CAEnsure that Google Ads services within the AViD ecosystem have reliability and uptime appropriate to users" needs with a fast rate of improvement, while keeping an ever-watchful eye on capacity and performance. Lead and contribute to the cross-SRE AI Ops program, driving the strategic adoption of AI/ML tools to improve incident response, reduce toil, and enhance service reliability across Ads SRE.
Reliability Engineer, Mechanical Systems, NA Vantage Data Centers Management Co LLCReliability Engineer, Mechanical Systems, NASanta Clara, CAFor each of the major systems Electrical, Mechanical, and Controls, the Reliability Engineering team is responsible for ensuring success in the commissioning stages of new construction, evaluating and improving the reliability and performance of existing critical infrastructure, sustaining equipment operational availability through maintenance program design, providing ongoing technical support to the Site Operations Teams, as well as providing systems reliability and maintainability feedback to the Design Engineering teams for future design considerations. Developing and operating across North America, EMEA and Asia Pacific, Vantage has evolved data center design in innovative ways to deliver dramatic gains in reliability, efficiency and sustainability in flexible environments that can scale as quickly as the market demands.
Hardware Reliability Engineer Google LLCHardware Reliability EngineerMountain View, CADrive the failure analysis process for all failures discovered during reliability testing and lead the failure analysis process with internal and external cross-functional teams for all failures discovered during reliability testing. Proficiency in statistical data analysis including DOE, significance testing, sample sizes and confidence level using statistical tools such as JMP, Python.
NewSenior Site Reliability Engineer, Wikimedia Enterprise Social Impact GuideSenior Site Reliability Engineer, Wikimedia EnterpriseSan Francisco, CAWikimedia Enterprise aims to improve the user experience for Wikimedia/Wikipedia readers beyond our own websites; increase the reach and discoverability of Wikimedia/Wikipedia content; and improve awareness and ease of attribution and verifiability of Wikimedia/Wikipedia content by the organizations that reuse our content the most. We host Wikipedia and the Wikimedia projects, build software experiences for reading, contributing, and sharing Wikimedia content, support the volunteer communities and partners who make Wikimedia possible, and advocate for policies that enable Wikimedia and free knowledge to thrive.
Cell Reliability Engineer, Cell Quality & Reliability Tesla IncCell Reliability Engineer, Cell Quality & ReliabilityFremont, CA$109,600–$164,400 / yearAnalyze Field Reliability/Quality data for cell related failures and failure modes to predict expected failure rates, affected populations, verify effectiveness of the corrective actions at Tesla and at suppliers. Perform risk assessment with process engineers for proper documentation of control plan, pfmea/dfmea, ppap, IQC/OQC metrics along with supporting key production metrics to achieve yield, OEE.
Lead Database Reliability Engineer - 11606 Coupa Software IncLead Database Reliability Engineer - 11606San Francisco, CARemote$142,000–$198,667 / yearCollaborate effectively across cross-functional teams, mentor junior database engineers, stay current on emerging database technologies and best practices, and remain flexible to support global teams across multiple time zones. By submitting your application, you acknowledge that you have read Coupa's Privacy Policy and understand that Coupa receives/collects your application, including your personal data, for the purposes of managing Coupa's ongoing recruitment and placement activities, including for employment purposes in the event of a successful application and for notification of future job opportunities if you did not succeed the first time.
NewSr. Site Reliability Engineer Starlink Space Exploration Technologies CorpSr. Site Reliability Engineer StarlinkPalo Alto, CA$165,000–$280,000 / yearBASIC QUALIFICATIONS: Bachelor''s degree in computer science, engineering, math, or scientific discipline and 5 years of software development experience; OR 7+ years of professional experience building software with site reliability or DevOps in lieu of a degree. ITAR REQUIREMENTS: To conform to U.S. Government export regulations, applicant must be a (i) U.S. citizen or national, (ii) U.S. lawful, permanent resident (aka green card holder), (iii) Refugee under 8 U.S.C. § 1157, or (iv) Asylee under 8 U.S.C. § 1158, or be eligible to obtain the required authorizations from the U.S. Department of State.
Staff Site Reliability Engineer, Quota SRE Google LLCStaff Site Reliability Engineer, Quota SRESunnyvale, CAMany teams within Google use services such as Quotaserver and Bouncer to implement security policies, safety guardrails, regulations, internal protection, accounting mechanisms, etc or Slicer to enable dynamic autosharding, load balancing, and caching for stateful services. SRE ensures that Google Cloud"s services-both our internally critical and our externally-visible systems-have reliability, uptime appropriate to customer"s needs and a fast rate of improvement.
Lead Infrastructure and Reliability Engineer (Systems & Scale) Luma AI IncLead Infrastructure and Reliability Engineer (Systems & Scale)Palo Alto, CARequired: • Deep expertise in Linux and distributed systems • Experience operating GPU / accelerator clusters in real production environments • Strong fluency in Kubernetes and modern open-source infrastructure • Comfortable debugging across hardware kernel runtime orchestration • You understand how systems behave under contention and at scale • You write code and build automation • You think in bottlenecks, failure modes, and tradeoffs • Engineers trust your judgment, especially when things break. • Scaling Training & Inference Define how infrastructure and workloads evolve as cluster size and concurrency grow Design scheduling, placement, and resource management approaches for increasingly complex jobs Work directly with research to build the systems required for new model capabilities Ensure inference platforms scale rapidly without sacrificing reliability or latency Anticipate where today's abstractions will fail and redesign ahead of them.
NewSite Reliability Engineer - Senior Staff Vistance NetworksSite Reliability Engineer - Senior StaffSunnyvale, CA$118,000–$170,000 / yearHow You'll help us connect the world:Ruckus Networks is looking for a customer focused Senior Site Reliability Engineer (SRE) to help improve reliability, scalability, operational excellence, and customer experience across our cloud platform ecosystem. Overview Site Reliability Engineer - Senior StaffReq ID: 81736Location: Sunnyvale, California, United States, 94089In our ‘always on' world, we believe it's essential to have a genuine connection with the work you do.
Site Reliability Engineer, Cloud SQL SRE Google LLCSite Reliability Engineer, Cloud SQL SRESunnyvale, CAWe"re looking for engineers who bring fresh ideas from all areas, including information retrieval, distributed computing, large-scale system design, networking and data storage, security, artificial intelligence, natural language processing, UI design and mobile; the list goes on and is growing every day. Google Cloud offers a portfolio of managed database services, catering to a wide range of customer needs from enterprise applications to modern, cloud-native solutions, including those powering Gen AI/ML.
Site Reliability Engineer Akkodis Group AG.Site Reliability EngineerSunnyvale, CA$64–$68 / hourSite Reliability Engineer Job Responsibilities include: Design, deploy, and maintain highly available and scalable cloud infrastructure on AWS using services such as EC2, EKS, S3, RDS, Lambda, IAM, and VPC. The ideal candidate with experience maintaining highly available, scalable cloud infrastructure and driving operational excellence through automation, monitoring, and incident management.
Staff Infrastructure Reliability Engineer - Database & Storage Rocket Companies IncStaff Infrastructure Reliability Engineer - Database & StorageSan Francisco, CARemote$180,100–$278,700 / yearThis role requires depth in design, collaboration with internal teams, a proactive approach to problem solving, and the ability to share complex ideas with senior leadership and secure their support. You have a proven history in architecting, building, scaling, and supporting cloud infrastructure technologies, specializing in database and storage services and can communicate the direct business impact of this work.
Reliability Engineer Ecovyst IncReliability EngineerMartinez, CADrive the development of Key Performance Indicators (KPIs), such as Asset Utilization, Overall Equipment Effectiveness (OEE), Mean Time Between Failure (MTBF), On-Stream Time (OST), and benchmark performance against best-in-class, track progress, and measure improvements. This role will partner with and support Operations and Maintenance through identification and reduction and/or elimination of production losses and high maintenance cost assets to promote plant objectives in the areas of Health, Safety, and Environmental (HSE), asset capability, quality, and production.
Reliability Engineer, Energy Storage Redwood Materials IncReliability Engineer, Energy StorageSan Francisco, CA$122,500–$192,500 / yearRedwood is localizing a global battery supply chain that seamlessly integrates recovery, reuse, and recycling - keeping critical minerals in circulation and driving the energy transition. Deep knowledge of electronic and mechanical failure mechanisms (e.g., solder fatigue, corrosion, delamination, sealing failures, thermal expansion mismatch, material fatigue, etc.).
Wireless Module Reliability Engineer Apple IncWireless Module Reliability EngineerSunnyvale, CAAct as the primary technical liaison between internal cross-functional teams and reliability teams at vendors, ensuring alignment on reliability campaigns, determining stress test conditions and approving test hardware. As a Reliability Engineer for the Wireless Design group, you will be responsible for driving all reliability requirements across multiple Wireless vendors.
Sr. Site Reliability Engineer Backblaze IncSr. Site Reliability EngineerSan Mateo, CA$150,000–$200,000 / yearToday, Backblaze generates over $100m in revenue and is the leading specialized storage cloud - managing over three billion gigabytes of data storage for 500K+ customers in 175+ countries, including businesses, developers, IT professionals, and individuals. The SRE will collaborate with engineering, product, and operations teams to embed reliability practices into day-to-day development and operations while contributing to tools and processes that improve efficiency and reduce manual effort.
Senior Lead Site Reliability Engineer JPMorgan Chase Bank, N.A.Senior Lead Site Reliability EngineerPalo Alto, CAFull timeStrong experience building production-grade RESTful APIs and designing message queue architectures (Kafka, RabbitMQ, SQS) for event-driven systems; and expertise in graph databases (Neo4j, TigerGraph), vector databases (Pinecone, Weaviate, Chroma), and integrating multiple data stores for AI-powered systems. Hands-on experience building AI Agents and autonomous systems with proficiency in AI frameworks (LangChain, LangGraph, AutoGen, CrewAI) and leveraging AI development tools (GitHub Copilot, Claude, etc.) to accelerate development and innovation and Expertise in designing and implementing logging pipelines (Fluentd, Logstash, Vector) and systems for metrics collection, analysis, and distributed tracing.
Senior Site Reliability Engineer OutSystemsSenior Site Reliability EngineerSan Francisco, Northern Mariana IslandsAs an SRE at OutSystems here are your key responsibilities and duties: Lead and onboard services and teams to the reliability tenets; Establish and maintain Service Level Objectives (SLOs) and Service Level Agreements (SLAs); Design and implement scalable, reliable, and secure infrastructure, while ensuring cloud-native best practices; Collaborate with software development teams to ensure systems are resilient (observable, fault-tolerant, recoverable, scalable) and performant; Implement monitoring, alerting, logging, and tracing solutions to detect and respond to incidents; Lead incident response efforts, ensuring quick resolution and minimal downtime, and conduct RCA/post-mortems; Automate every operational task, with a special focus on fast incident detection & recovery; Programming in Python supported by Gen AI tooling to accelerate development of mission critical automation and tools. (CKA, CKAD, CKS certifications are valued); Experience with automation and Infrastructure as Code (IaC) tools, such as AWS CloudFormation, Terraform, Puppet, Chef, Spacelift, etc; Experience with Python, Go, Bash/Shell scripting, or other automation tools/languages; Familiarity with AWS services like EC2, RDS, ELB, CloudFront, Lambda, etc; Proficiency in monitoring and troubleshooting complex distributed systems; Experience with Grafana, ELK stack, Prometheus, or others; Strong understanding of designing resilient and fault-tolerant systems; Expertise in debugging complex distributed systems.
Site Reliability Engineer TELCOR IncSite Reliability EngineerSan Francisco, CA$125,000–$165,000 / yearCopy and paste the following link into your browser to learn more about TELCOR and what it means for TELCOR to be a certified Great Place To Work: https://www.greatplacetowork.com/certified-company/7054288 . This role will also design and operate resilient systems across cloud and containerized environments, as well as manage production infrastructure and deployment workflows across environments.
Sr. Reliability Engineer, Power Modules and AI Applications Monolithic Power Systems IncSr. Reliability Engineer, Power Modules and AI ApplicationsSan Jose, CA$135,000–$170,000 / yearClosely work with Design, Marketing/Customers, Application teams as integral part of Quality Assurance function to develop Reliability strategies and methodologies based on customer requirements and mission-profile analysis. Work as an integral function of Quality Assurance to analyze reliability data using statistical tools such as JMP and Reliasoft and present findings to internal and external teams.
NewSenior Site Reliability Engineer AbbottSenior Site Reliability EngineerSunnyvale, CaliforniaExcellent communication skills with the demonstrated ability to work effectively in cross-functional teams, translate technical complexity for non-technical stakeholders, and collaborate with development, quality, security, marketing, and regulatory teams. Work closely with software engineering, security, quality, and compliance teams to integrate SRE best practices into our operational processes and infrastructure, ensuring the integrity, availability, and confidentiality of our systems that carry sensitive patient health data.