Senior Microsoft Cloud Infrastructure Engineer Crusoe EnergySenior Microsoft Cloud Infrastructure EngineerSan Francisco, CA$160,000–$195,000 / yearWe're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved - people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services. The role sits at the intersection of IT, Security, and site-specific operations to keep systems reliable, secure, and ready to scale with Crusoe's rapidly expanding footprint of offices, warehouses, and data centers.
Staff Software Engineer, Managed Orchestration (Managed Kubernetes) Crusoe EnergyStaff Software Engineer, Managed Orchestration (Managed Kubernetes)San Francisco, CA$220,000–$250,000 / yearWe're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved - people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services. Work collaboratively with tech leads and engineers to create a dynamic environment where creativity and technical excellence are encouraged, leading to the development of cutting-edge cloud solutions.
Senior Tools Development Engineer - Cosmos Platform Software Astera LabsSenior Tools Development Engineer - Cosmos Platform SoftwareSan Jose, CA$140,000–$175,000 / yearAstera Labs' Intelligent Connectivity Platform integrates CXL, Ethernet, NVLink, PCIe, and UALink semiconductor-based technologies with the company's COSMOS software suite to unify diverse components into cohesive, flexible systems that deliver end-to-end scale-up, and scale-out connectivity. We know that creativity and innovation happen more often when teams include diverse ideas, backgrounds, and experiences, and we actively encourage everyone with relevant experience to apply, including people of color, LGBTQ+ and non-binary people, veterans, parents, and individuals with disabilities.
NewSenior Staff Deployment Automation Engineer Crusoe EnergySenior Staff Deployment Automation EngineerSan Francisco, CA$250,000–$300,000 / yearWe're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved - people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services. Multi-Node Scaling Validation: Design and execute large-scale validation tests across multi-node virtualized clusters to ensure linear scaling and stability of GPU workloads.
Senior Staff Hardware Qualification Engineer D-MatrixSenior Staff Hardware Qualification EngineerSanta Clara, CASignal & Power Integrity Validation: Provide senior-level oversight for the validation of high-speed SerDes (112G/224G/PCIe), LPDDR, CoWOS designs, and complex PDN (Power Delivery Network) transients. This is a high-impact, technical leadership role responsible for defining the qualification strategies that ensure our hardware-from ASIC substrates to Rack scale systems - can withstand the rigorous demands of 24/7 data center environments.
Software Engineer, Runtime & C++ Middleware PlusAI Inc.Software Engineer, Runtime & C++ MiddlewareSanta Clara, CA$120,000–$200,000 / yearWe may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. In this role, you will contribute to existing frameworks, libraries, and tools while also designing and implementing new components across various mission-critical domains.
NewRegional Principal Product Application Engineer (Asia) Astera LabsRegional Principal Product Application Engineer (Asia)San Jose, CA$185,000–$240,000 / yearAs an Astera Labs Principal Product Application Engineer, you will need to provide technical guidance to customers to overcome design challenges, generate collateral for existing and new products, and drive innovation by providing insightful feedback to other internal teams to continuously improve products and processes. Astera Labs' Intelligent Connectivity Platform integrates CXL, Ethernet, NVLink, PCIe, and UALink semiconductor-based technologies with the company's COSMOS software suite to unify diverse components into cohesive, flexible systems that deliver end-to-end scale-up, and scale-out connectivity.
NewStaff Software Engineer, Storage Crusoe EnergyStaff Software Engineer, StorageSan Francisco, CA$240,000–$310,000 / yearWe're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved - people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services. Deep Performance Engineering: Lead "tiger teams" to solve the most ambiguous and difficult bottlenecks in the stack-from kernel-level IO context switching to global tail-latency in distributed clusters.
NewInfrastructure Design Engineer Crusoe EnergyInfrastructure Design EngineerSunnyvale, CA$155,000–$190,000 / yearWe're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved - people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services. The Infrastructure Design Engineer is responsible for planning, designing, and implementing data center whitespace environments - the operational computing areas where servers, storage, and network equipment are deployed.
Senior Robotics Test Engineer ATOMSSenior Robotics Test EngineerSan Francisco, CaliforniaSimulation-Based Testing — Build the simulation infrastructure that lets us run autonomy software against virtual scenarios, with deterministic playback and meaningful coverage of real-world conditions. CI/CD for Robotics — Build and maintain the testing pipeline that gates code changes — unit, integration, simulation, and replay-based regression — and integrate it cleanly into the developer workflow.
NewStaff+ Software Engineer, Product Sandboxing AnthropicStaff+ Software Engineer, Product SandboxingSan Francisco, CAYou might be a good fit if you: Have a minimum of 8 years of practical experience as a backend or platform engineer building scalable distributed systems, ideally operating at tech lead level or equivalent distributed systems, cloud-native products, developer tools, or external developer facing products. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences.
NewSenior Site Reliability Engineer Andromeda ClusterSenior Site Reliability EngineerSan Francisco, CaliforniaIncident Management: Proven track record leading incident response for complex distributed systems where the failure could be in hardware, firmware, networking, drivers, orchestration, or application code and you need to narrow it down fast. Observability & Monitoring: Hands-on experience building monitoring and alerting for GPU infrastructure, not just Prometheus/Grafana basics, but GPU-specific telemetry (DCGM, nvidia-smi, fabric manager metrics) integrated into actionable dashboards.
NewSenior mixed-signal electrical engineer (Board Level) ASMLSenior mixed-signal electrical engineer (Board Level)San Jose, CaliforniaLead the design, analysis, and implementation of high-performance analog and mixed-signal circuits on board level , including: High-speed (>50 MSPS), high-precision (16-bit) DAC/ADC signal chains. Individual pay is determined through interviews and an assessment of several factors that that are unique to each candidate, including but not limited to job-related skills, relevant education and experience, certifications, abilities of the candidate and pay relative to other team members.
Senior Robotics Systems Engineer ATOMSSenior Robotics Systems EngineerSan Francisco, CaliforniaYou’ll develop, test, and deploy features for our autonomous trucks, and you’ll serve as a technical lead during deployments, owning bring-up, triage, and cross-team coordination in mission-critical environments. Our systems are designed to understand, predict, and control the real world with precision, turning complex physical operations into something more reliable, more scalable, and more productive.
NewCo-Op, DevOps Engineer LiveRampCo-Op, DevOps EngineerSan Francisco, CaliforniaHundreds of global innovators, from iconic consumer brands and tech giants to banks, retailers, and healthcare leaders turn to LiveRamp to build enduring brand and business value by deepening customer engagement and loyalty, activating new partnerships, and maximizing the value of their first-party data while staying on the forefront of rapidly evolving compliance and privacy requirements. LiveRamp offers complete flexibility to collaborate wherever data lives to support the widest range of data collaboration use cases—within organizations, between brands, and across its premier global network of top-quality partners.
NewStaff Optical Design Engineer Lumentum Inc.Staff Optical Design EngineerSan Jose, CAWith our continual goal of making Lumentum a best place to work for our employees, we strive to offer employees competitive total compensation packages, which may include annual bonus, commission for certain sales roles, equity, and health and welfare benefits. Coordinate with internal teams, including Process, Testing, and Automation, to review and respond to their DFM items, optimize designs, and establish appropriate assembly and test workflows.
NewStaff Slurm Cluster & HPC Engineer BitdeerStaff Slurm Cluster & HPC EngineerSan Jose, CACluster health and reliability engineering- Build the passive and active health-check system expected of a top-tier GPU cloud: prolog/epilog checks, LBNL NHC or equivalent, DCGM diagnostics, and detection of XID/SXID errors, ECC faults, PCIe errors, GPUs falling off the bus, IB/RoCE link flaps, and NCCL stalls - with automatic drain and job requeue. Slurm cluster architecture and lifecycle- Design, deploy, and operate production Slurm clusters on bare metal and VMs: slurmctld/slurmdbd high availability, slurmrestd, configless slurmd, SACK/MUNGE and JWT authentication, and rolling version upgrades on live clusters without losing running jobs.
Software Engineer (SE / Sr Se), Data Infrastructure PlusAI Inc.Software Engineer (SE / Sr Se), Data InfrastructureSanta Clara, CA$120,000–$200,000 / yearIn this role, you will own critical parts of the path that makes this data reliable and useful-from high-throughput recording on the vehicle to transfer, validation, cataloging, and efficient access for replay, analytics, and machine learning. Responsibilities: Design and evolve high-throughput C++ systems that continuously record vehicle sensor and runtime data, selectively capture events, and collect system telemetry while operating safely under CPU, memory, disk, and network constraints.
NewStaff Maas Backend Engineer BitdeerStaff Maas Backend EngineerSan Jose, CAWe are seeking a Staff Backend Engineer to take our Model-as-a-Service (MaaS) working system and re-architect it into a commercial, globally distributed, multi-tenant token service - one that sustains at least millions of monthly active users, high sustained token throughput per GPU, and invoice-grade accounting, with scalability, reliability, and observability engineered in deliberately rather than absorbed under load. Global Topology and Reliability (SLOs): Scale the platform to a globally distributed architecture featuring regional inference pools, capacity-aware failovers, and an active-active control plane.
NewSenior GPU Systems & Fabric Engineer BitdeerSenior GPU Systems & Fabric EngineerSan Jose, CAThis role requires deep expertise in Linux kernel internals, GPU architectures, and high-speed interconnects, as you will be tasked with transforming raw, bare-metal compute resources into scalable, resilient, and multi-tenant cloud primitives. Build and manage automated hardware remediation pipelines using DCGM telemetry to proactively identify, isolate, and reset degraded GPU/NIC components before they impact production jobs.