Infrastructure Software Engineer, Enterprise Genai Scale AI, Inc.Infrastructure Software Engineer, Enterprise GenaiSan Francisco, CA$179,400–$224,250 / yearThe range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position and may be inclusive of several career levels at Scale; it will be determined during the interview process based on work location and additional factors, including job-related skills, experience, qualifications, interview performance, and relevant education or training. Our products provide the high-quality data and full-stack technologies that power the world's leading models, and help enterprises and governments build, deploy, and oversee AI applications that deliver real impact.
NewStaff Software Engineer, DC Infrastructure Crusoe EnergyStaff Software Engineer, DC InfrastructureSan Francisco, CA$215,000–$260,000 / yearWe're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved - people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services. Own the deployment, monitoring, and operational support of developed tooling, ensuring solutions maximize GPU fleet availability and performance to drive customer success.
Technical Program Manager, Safeguards (Infrastructure & Evals) AnthropicTechnical Program Manager, Safeguards (Infrastructure & Evals)San Francisco, CAYour primary responsibility is driving reliability - owning the incident-response and post-mortem process, ensuring SLOs are defined and met in partnership with various teams, and making sure that when things go wrong, the right people know, the right actions get taken, and those actions actually get closed out. What You'll Do: Own the Safeguards Engineering ops review- Drive the recurring cadence that keeps the team informed and coordinated: surfacing recent incidents and failures, bringing visibility to reliability trends, and making sure the right people are in the room when decisions need to be made.
Product Manager, AI Infrastructure Together AIProduct Manager, AI InfrastructureSan Francisco, CA$175,000–$220,000 / yearOur product surface is expanding fast- GPU clusters, managed storage, networking, and observability - and we're adding a Product Manager to the Together Cloud team to own the day-to-day product work that keeps these AI infrastructure products moving. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models.
Senior Manager, Infrastructure Engineering True AnomalySenior Manager, Infrastructure EngineeringSan Francisco, CA$165,000–$230,000 / yearBuild and lead the Infrastructure Engineering team: initially 3 Senior Engineers (Kubernetes/platform reliability, networking, CI/CD/cloud), growing into dedicated subteams as the organization scales. To conform to U.S. Government space technology export regulations, including the International Traffic in Arms Regulations (ITAR) you must be a U.S. citizen, lawful permanent resident of the U.S., protected individual as defined by 8 U.S.C. 1324b(a)(3), or eligible to obtain the required authorizations from the U.S. Department of State.
Engineering Manager, Infrastructure KikoffEngineering Manager, InfrastructureSan Francisco, California$307,000–$352,000 / yearWith record revenue growth in 2025 and a unicorn valuation, we've built a suite of products that help millions of people build credit, access liquidity, and save money. The team owns five connected areas: Observability, Developer Productivity, Compute Infrastructure, Networking and Storage, and Data Infrastructure.
Global Public Policy Manager, Compute, Infrastructure & Sovereign AI CohereGlobal Public Policy Manager, Compute, Infrastructure & Sovereign AISan Francisco, California$190,000–$230,000 / yearAll legitimate roles are listed on the Cohere careers page and LinkedIn only, with all communications from Cohere employees coming from an @cohere.com or @cw.cohere email alias. Governments increasingly view AI as critical national infrastructure and are investing heavily in compute capacity, energy resources, and domestic AI ecosystems.
Staff Engineer, Distributed Storage And HPC & AI Infrastructure Together AIStaff Engineer, Distributed Storage And HPC & AI InfrastructureSan Francisco, CA$250,000–$300,000 / yearYou'll manage and scale high-performance parallel filesystems and object stores, evaluate and integrate cutting-edge technologies such as Vast, Weka, Ceph, and Lustre, and solve the complex engineering challenges of operating at extreme throughput, low-latency data paths, and massive cluster-scale storage operations. Engineer end-to-end data paths to achieve 10+ GB/s per GPU node; architect multi-tier caching for model weights and datasets; tune parallel filesystems using advanced profiling; and scale storage infrastructure across thousands of nodes.
Staff Engineer, Distributed Storage and HPC & AI Infrastructure Together AIStaff Engineer, Distributed Storage and HPC & AI InfrastructureSan Francisco, California$250,000–$300,000 / yearYou’ll manage and scale high-performance parallel filesystems and object stores, evaluate and integrate cutting-edge technologies such as Vast, Weka, Ceph, and Lustre, and solve the complex engineering challenges of operating at extreme throughput, low-latency data paths, and massive cluster-scale storage operations. Engineer end-to-end data paths to achieve 10+ GB/s per GPU node; architect multi-tier caching for model weights and datasets; tune parallel filesystems using advanced profiling; and scale storage infrastructure across thousands of nodes.
Support Engineer, AI Infrastructure FIGMASupport Engineer, AI InfrastructureSan Francisco, CA$169,000–$245,000 / yearHands-on experience with LLM-powered workflows, AI automations, or AI-enabled customer/support experiences, including working with operational data to debug issues, improve workflows, and measure impact. This role is ideal for someone who can move from ambiguous support problems to working technical solutions: understanding the workflow, identifying the systems involved, building the integration or automation, validating the data flow, and measuring the impact on customer outcomes and Specialist efficiency.
Backend / Infrastructure Engineer Elite Talent ConsultingBackend / Infrastructure EngineerSan Francisco, CaliforniaIn this role, you will design and develop scalable infrastructure that powers our core product, enabling radiologists to work more efficiently and double their productivity. About the Role: We are hiring mid-level and senior backend engineers to help build an AI-powered platform that is revolutionizing workflows in radiology.
Staff AI Infrastructure Engineer Luma AI, Inc.Staff AI Infrastructure EngineerRedwood City, CAAs a Staff AI Infrastructure Engineer, you'll be a technical authority who turns deep systems knowledge into repeatable, company-wide reliability, and a leader other strong engineers want to work with. This is close-to-the-metal work - kernels, containers, schedulers, networking, storage, GPU behavior - under demand hard enough that yesterday's solutions break regularly.
Engineering Manager, Machine Learning Infrastructure, Ads RobloxEngineering Manager, Machine Learning Infrastructure, AdsSan Mateo, CA$295,250–$345,040 / yearEvery day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. With Roblox Ads business growing at a rapid rate, we are building large scale ads machine learning infrastructure to deliver effective performance ads to our users, and more business values to our advertisers.
Software Engineer - Infrastructure (Senior+) Cogent SecuritySoftware Engineer - Infrastructure (Senior+)San Francisco, CaliforniaYou'll join a small, fast-growing Core Platform pod (that includes infrastructure, security, and compliance) and work hands-on across our data platform, our security systems, and the platform layer our agents use to build and ship the product. Alongside our core product work, Cogent Research serves as our applied AI lab, providing the research horsepower needed to make truly agentic security workflows a reality.
Technical Program Manager, Data Center Infrastructure AnthropicTechnical Program Manager, Data Center InfrastructureSan Francisco, CAThis research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues.
Senior Security Software Engineer, Infrastructure Security RobloxSenior Security Software Engineer, Infrastructure SecuritySan Mateo, CA$233,570–$269,170 / yearWork closely with other InfoSec teams (AppSec, D&R, GRC, CorpSec, CloudSec, NetSec) and partner with engineering teams across Roblox, specifically the Infrastructure organization, to ensure the secure outcomes of security and product driven initiatives. Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators.
Manager, Insurance & Risk Management – Energy & Digital Infrastructure SB EnergyManager, Insurance & Risk Management – Energy & Digital InfrastructureRedwood City, CA$140,000–$165,000 / yearSupport the placement, renewal, and ongoing management of corporate insurance programs, including Property, General Liability, Excess, Auto, Workers' Compensation, Cyber, Management Liability, Builders Risk, and OCIP/CCIP programs. In 2026, SoftBank Group and OpenAI announced a $1 billion investment to support SB Energy as their leading development and execution partner for data center campuses, while Ares Infrastructure continues its long-standing support of the company's growth.
Staff Software Engineer - Data Infrastructure PlaidStaff Software Engineer - Data InfrastructureSan Francisco, CaliforniaLeading key data infrastructure projects such as improving ML development golden paths, implementing offline streaming solutions for data freshness, building net new ETL pipeline infrastructure, and evolving data warehouse or data lakehouse capabilities. We provide tooling and guidance to teams across engineering, product, and business and help them explore our data quickly and safely to get the data insights they need, which ultimately helps Plaid serve our customers more effectively.
Senior Infrastructure Endpoint Engineer EverOpsSenior Infrastructure Endpoint EngineerSan Francisco, CaliforniaRemoteThis role sits at the intersection of endpoint engineering, identity, and security, with deep expertise across leading EDR platforms (CrowdStrike Falcon, SentinelOne, and Microsoft Defender for Endpoint) and modern MDM platforms (Intune, Kandji/Iru, and Fleet) to drive automation, visibility, and user experience. Beyond hands-on execution, you will own the endpoint and security tooling roadmap, evaluate platform tradeoffs, and drive technical decisions using data - not opinion - while designing automated provisioning workflows tied to Autopilot or Apple Business Manager.
NewTech Lead, Software Infrastructure and Fleet Management Mind RoboticsTech Lead, Software Infrastructure and Fleet ManagementPalo Alto, CaliforniaHands-on experience with integration and systems testing for hardware-attached software — hardware-in-the-loop testing, simulation-based regression suites (Isaac Sim, Gazebo, MuJoCo, or similar), and CI/CD pipelines that gate on more than unit tests. Design and build testing infrastructure — hardware-in-the-loop (HIL) test reporting, simulation-based regression testing, and pre-deployment validation gates that catch failures before they reach a live site.