Engineering Manager, Infrastructure SentryEngineering Manager, InfrastructureSan Francisco, CA$220,000–$450,000 / yearAs the Engineering Manager for Infrastructure Engineering, you'll lead a team of engineers building the tools that power Sentry's growth: internal admin and change management tools, configuration automation, and the routing layer that underlies Sentry's architecture. They build the internal control platforms, configuration systems, traffic routing, and automation that let product engineers operate services safely at scale without needing deep infrastructure expertise themselves.
Engineering Manager, Model Infrastructure Harvey, Inc.Engineering Manager, Model InfrastructureSan Francisco, CA$260,000–$340,000 / yearExperience working with multiple model providers such as OpenAI, Anthropic, Azure OpenAI, Fireworks, Baseten, or open-source model ecosystems. Drive the evolution of our Unified Model Controller (UMC) and Model Selector platform to automatically detect degraded models and intelligently route traffic based on health, latency, quality, compliance, and cost.
Data Center Infrastructure Electrical Engineer OpenAIData Center Infrastructure Electrical EngineerSan Francisco, CaliforniaFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. The ideal candidate has strong practical experience with critical electrical systems at data centers or comparable industrial scale, including medium-voltage and low-voltage distribution, utility interfaces, backup power, UPS and battery systems, rack power delivery, grounding, protection, controls, and monitoring systems.
Software Engineer, ChatGPT Infrastructure OpenAISoftware Engineer, ChatGPT InfrastructureSan Francisco, CaliforniaFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. We focus on high-leverage infrastructure: primitives and “golden paths” that incorporate operational lessons as defaults, so engineers don’t need to rediscover failure modes, latency pitfalls, or integration issues each time they build something new.
Software Engineer, Agentic AI Infrastructure AnrokSoftware Engineer, Agentic AI InfrastructureSan Francisco, CaliforniaRemoteWe're looking for an engineer who pairs a strong distributed-systems background with real, hands-on experience building AI agent systems—someone who understands both how to make a service fault-tolerant and what it takes to make an LLM-powered agent reliable in production. As the digital economy has grown 6x over the last decade, software businesses have gone from not worrying about sales tax to needing to monitor exposure, calculate rates, and file returns across 50 US jurisdictions and 100+ countries.
Software Engineer, Agent Infrastructure OpenAISoftware Engineer, Agent InfrastructureSan Francisco, CaliforniaFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. Some of the most challenging technical problems in scaling the capabilities and utility of agents and agentic models lie in the infrastructure layer – and our team is focused on building the research and production systems that enable OpenAI to train the most capable models in the world, and maximize the utility of our agentic products for users around the world.
Software Engineer, Infrastructure HebbiaSoftware Engineer, InfrastructureSan Francisco, CA$160,000–$300,000 / yearYou will spend your time designing multi-account architecture, writing the IaC and automation that makes environments reproducible, driving down cloud spend, hardening our infrastructure for enterprise and financial-services scrutiny, and making the path from a developers laptop to production fast and boring. Founded in 2020 by George Sivulka and backed by Peter Thiel and Andreessen Horowitz, Hebbia powers investment decisions for BlackRock, KKR, Carlyle, Centerview, and 40% of the worlds largest asset managers.
Member of Technical Staff (Software Engineer, Cloud Infrastructure) PerplexityMember of Technical Staff (Software Engineer, Cloud Infrastructure)San Francisco, CaliforniaDesign and operate Perplexity’s cloud networking fabric, including VPC architectures, private connectivity, and peering with hyperscalers and neocloud providers to support low-latency, high-throughput AI workloads. Architect and scale compute platforms (Kubernetes/EKS, autoscaling groups, and mixed CPU/GPU fleets) to efficiently serve online request traffic and background workloads across regions.
Infrastructure Supply Chain Manager, Autonomy & Robotics DoorDash IncInfrastructure Supply Chain Manager, Autonomy & RoboticsOakland, CA$130,600–$192,000 / yearWe value a diverse workforce - people who identify as women, non-binary or gender non-conforming, LGBTQIA+, American Indian or Native Alaskan, Black or African American, Hispanic or Latinx, Native Hawaiian or Other Pacific Islander, differently-abled, caretakers and parents, and veterans are strongly encouraged to apply. As such, we are looking for an Infrastructure Supply Chain Manager who has strong relevant experience to source and manage supply chain operations for colocation, equipment, and facilities in this field while having the analytical mindset, creativity, and vision to tackle the challenges in the present and future for DoorDash Labs.
Member of Technical Staff - GPU Infrastructure Engineer Liquid AIMember of Technical Staff - GPU Infrastructure EngineerSan Francisco, CaliforniaSpun out of MIT CSAIL, we build general-purpose AI systems that run efficiently across deployment targets, from data center accelerators to on-device hardware, ensuring low latency, minimal memory usage, privacy, and reliability. We are looking for a hands-on software engineer to keep our GPU clusters reliable, improve resource efficiency, and build the tooling that allows researchers to focus on model development rather than infrastructure.
Portal & Infrastructure Tech Lead Faraday FuturePortal & Infrastructure Tech LeadFremont, CA$155,000–$180,000 / yearThe ideal candidate is a seasoned platform engineer who has built systems used by large external developer communities, understands what makes an API a pleasure or a pain to work with, and has the leadership range to grow a high-performing engineering team while remaining deeply technically engaged. Lead the design and engineering of the platform's core developer-facing products: the robotics APIs (covering motion control, perception, sensor fusion, autonomous decision-making, and inter-device communication) and the multi-language SDKs (Python, Node.js, Go, Java, and others as needed).
NewSenior Software Engineer, Infrastructure, Platforms and Devices Google LLCSenior Software Engineer, Infrastructure, Platforms and DevicesSan Francisco, CAWere looking for engineers who bring fresh ideas from all areas, including information retrieval, distributed computing, large-scale system design, networking and data storage, security, artificial intelligence, natural language processing, UI design and mobile; the list goes on and is growing every day. The Platforms and Devices team encompasses Googles various computing software platforms across environments (desktop, mobile, applications), as well as our first party devices and services that combine the best of Google AI, software, and hardware.
NewSoftware Engineer, Anti-distillation Investigations, Enforcements and Infrastructure, DeepMind Google LLCSoftware Engineer, Anti-distillation Investigations, Enforcements and Infrastructure, DeepMindSan Francisco, CAAt Google DeepMind, we are a pioneering AI lab with exceptional interdisciplinary teams focused on advancing AI development to solve complex global challenges and accelerate high-quality product innovation for billions of users. Google is a global company and, in order to facilitate efficient collaboration and communication globally, English proficiency is a requirement for all roles unless stated otherwise in the job posting.
Engineering Manager, GPU Infrastructure CohereEngineering Manager, GPU InfrastructureSan Francisco, CaliforniaA background running large Kubernetes compute fleets in production, including in multi-cloud environments: multi-cluster operations, scheduling, node health at scale, and familiarity with IaC and infrastructure monitoring. You’ve gone deep in one of the layers that make a GPU training fleet work, whether that’s cluster-wide operations, GPU networking, or hardware, and you’re willing to get hands-on and learn the rest.
Global Supply Manager, Infrastructure Development Construction Tesla IncGlobal Supply Manager, Infrastructure Development ConstructionFremont, CA$84,000–$324,000 / yearThe lead role encompasses the selection of these through a competitive tender/bid process involving document preparation (scope, drawings, etc.) for projects while leading external negotiations, and collaborating closely with internal stakeholders. Synthesize supplier, industry and market research for assigned commodities (If needed) to develop sourcing strategy as well as issue RFx (RFQ, RFB, and RFP) for sourcing assigned commodities/services.
Staff Software Engineer, On-Device Machine Learning Infrastructure Google LLCStaff Software Engineer, On-Device Machine Learning InfrastructureSunnyvale, CAWere looking for engineers who bring fresh ideas from all areas, including information retrieval, distributed computing, large-scale system design, networking and data storage, security, artificial intelligence, natural language processing, UI design and mobile; the list goes on and is growing every day. Solve technically tests problems that exceed the scope of a generalist Software Engineers, specifically around optimizing Generative AI performance across heterogeneous hardware (CPUs, GPUs, and EdgeTPUs).
Product Manager, Platform - Cloud Infrastructure Commercial/Federal C3.ai IncProduct Manager, Platform - Cloud Infrastructure Commercial/FederalRedwood City, CA$150,000–$190,000 / yearC3 AI delivers a family of fully integrated products including the C3 Agentic AI Platform, an end-to-end platform for developing, deploying, and operating enterprise AI applications, C3 AI applications, a portfolio of industry-specific SaaS enterprise AI applications that enable the digital transformation of organizations globally, and C3 Generative AI, a suite of domain-specific generative AI offerings for the enterprise. Set and hold the bar on deployment KPIs - for example, time-to-stand-up a new tenant, deployment success rate, mean-time-to-recover for failed rollouts, and patch latency across the fleet - and partner with Forward Deployed Engineering on the feedback loop from real customer deployments back into the platform roadmap.
Staff Product Manager, Infrastructure Harvey, Inc.Staff Product Manager, InfrastructureSan Francisco, CA$213,600–$300,000 / yearFluency working with engineering and design teams on technical trade-offs, and comfort engaging with concepts like distributed systems, APIs, and cloud infrastructure at a level sufficient to make informed decisions. You'll also serve as the connective tissue across functions, ensuring that customer feedback, competitive dynamics, and technical realities all inform the product direction, and you'll raise the product bar across the organization through rigorous specs, reviews, and decision-making.
Member of Technical Staff — Reliability-CI Infrastructure RadixArkMember of Technical Staff — Reliability-CI InfrastructurePalo Alto, CaliforniaRadixArk is an infrastructure-first company built by engineers who've shipped production AI systems, created SGLang (30K+ GitHub stars, the fastest open LLM serving engine), and developed Miles (our large-scale RL framework). Our team has optimized kernels serving billions of tokens daily, designed distributed training systems coordinating 10,000+ GPUs, and contributed to infrastructure that powers leading AI companies and research labs.
ATE Test Infrastructure Engineer, Google Cloud Google LLCATE Test Infrastructure Engineer, Google CloudSunnyvale, CAFrom software to hardware our teams are shaping the future of world-leading hyperscale computing, with key teams working on the development of our TPUs, Vertex AI for Google Cloud, Google Global Networking, Data Center operations, systems research, and much more. Develop and validate test programs on Automated Test Equipment (ATE) platforms for new product integration (NPI) in preparation for high-volume manufacturing (HVM), working with ATE vendors and internal cross-functional teams.