NewStaff Software Engineer, Storage Crusoe EnergyStaff Software Engineer, StorageSan Francisco, CA$240,000–$310,000 / yearWe're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved - people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services. Deep Performance Engineering: Lead "tiger teams" to solve the most ambiguous and difficult bottlenecks in the stack-from kernel-level IO context switching to global tail-latency in distributed clusters.
NewInfrastructure Design Engineer Crusoe EnergyInfrastructure Design EngineerSunnyvale, CA$155,000–$190,000 / yearWe're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved - people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services. The Infrastructure Design Engineer is responsible for planning, designing, and implementing data center whitespace environments - the operational computing areas where servers, storage, and network equipment are deployed.
NewSenior Staff Deployment Automation Engineer Crusoe EnergySenior Staff Deployment Automation EngineerSan Francisco, CA$250,000–$300,000 / yearWe're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved - people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services. Multi-Node Scaling Validation: Design and execute large-scale validation tests across multi-node virtualized clusters to ensure linear scaling and stability of GPU workloads.
Senior Staff Hardware Qualification Engineer D-MatrixSenior Staff Hardware Qualification EngineerSanta Clara, CASignal & Power Integrity Validation: Provide senior-level oversight for the validation of high-speed SerDes (112G/224G/PCIe), LPDDR, CoWOS designs, and complex PDN (Power Delivery Network) transients. This is a high-impact, technical leadership role responsible for defining the qualification strategies that ensure our hardware-from ASIC substrates to Rack scale systems - can withstand the rigorous demands of 24/7 data center environments.
NewSenior mixed-signal electrical engineer (Board Level) ASMLSenior mixed-signal electrical engineer (Board Level)San Jose, CaliforniaLead the design, analysis, and implementation of high-performance analog and mixed-signal circuits on board level , including: High-speed (>50 MSPS), high-precision (16-bit) DAC/ADC signal chains. Individual pay is determined through interviews and an assessment of several factors that that are unique to each candidate, including but not limited to job-related skills, relevant education and experience, certifications, abilities of the candidate and pay relative to other team members.
Senior Robotics Systems Engineer ATOMSSenior Robotics Systems EngineerSan Francisco, CaliforniaYou’ll develop, test, and deploy features for our autonomous trucks, and you’ll serve as a technical lead during deployments, owning bring-up, triage, and cross-team coordination in mission-critical environments. Our systems are designed to understand, predict, and control the real world with precision, turning complex physical operations into something more reliable, more scalable, and more productive.
Senior Site Reliability Engineer Andromeda ClusterSenior Site Reliability EngineerSan Francisco, CaliforniaIncident Management: Proven track record leading incident response for complex distributed systems where the failure could be in hardware, firmware, networking, drivers, orchestration, or application code and you need to narrow it down fast. Observability & Monitoring: Hands-on experience building monitoring and alerting for GPU infrastructure, not just Prometheus/Grafana basics, but GPU-specific telemetry (DCGM, nvidia-smi, fabric manager metrics) integrated into actionable dashboards.
NewSenior Robotics Test Engineer ATOMSSenior Robotics Test EngineerSan Francisco, California$176,000–$242,000 / yearSimulation-Based Testing — Build the simulation infrastructure that lets us run autonomy software against virtual scenarios, with deterministic playback and meaningful coverage of real-world conditions. CI/CD for Robotics — Build and maintain the testing pipeline that gates code changes — unit, integration, simulation, and replay-based regression — and integrate it cleanly into the developer workflow.
Software Engineer (SE / Sr Se), Data Infrastructure PlusAI Inc.Software Engineer (SE / Sr Se), Data InfrastructureSanta Clara, CA$120,000–$200,000 / yearIn this role, you will own critical parts of the path that makes this data reliable and useful-from high-throughput recording on the vehicle to transfer, validation, cataloging, and efficient access for replay, analytics, and machine learning. Responsibilities: Design and evolve high-throughput C++ systems that continuously record vehicle sensor and runtime data, selectively capture events, and collect system telemetry while operating safely under CPU, memory, disk, and network constraints.
NewStaff+ Software Engineer, Product Sandboxing AnthropicStaff+ Software Engineer, Product SandboxingSan Francisco, CAYou might be a good fit if you: Have a minimum of 8 years of practical experience as a backend or platform engineer building scalable distributed systems, ideally operating at tech lead level or equivalent distributed systems, cloud-native products, developer tools, or external developer facing products. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences.
Co-Op, DevOps Engineer LiveRampCo-Op, DevOps EngineerSan Francisco, CaliforniaHundreds of global innovators, from iconic consumer brands and tech giants to banks, retailers, and healthcare leaders turn to LiveRamp to build enduring brand and business value by deepening customer engagement and loyalty, activating new partnerships, and maximizing the value of their first-party data while staying on the forefront of rapidly evolving compliance and privacy requirements. LiveRamp offers complete flexibility to collaborate wherever data lives to support the widest range of data collaboration use cases—within organizations, between brands, and across its premier global network of top-quality partners.
NewStaff Slurm Cluster & HPC Engineer BitdeerStaff Slurm Cluster & HPC EngineerSan Jose, CACluster health and reliability engineering- Build the passive and active health-check system expected of a top-tier GPU cloud: prolog/epilog checks, LBNL NHC or equivalent, DCGM diagnostics, and detection of XID/SXID errors, ECC faults, PCIe errors, GPUs falling off the bus, IB/RoCE link flaps, and NCCL stalls - with automatic drain and job requeue. Slurm cluster architecture and lifecycle- Design, deploy, and operate production Slurm clusters on bare metal and VMs: slurmctld/slurmdbd high availability, slurmrestd, configless slurmd, SACK/MUNGE and JWT authentication, and rolling version upgrades on live clusters without losing running jobs.
NewStaff Optical Design Engineer Lumentum Inc.Staff Optical Design EngineerSan Jose, CAWith our continual goal of making Lumentum a best place to work for our employees, we strive to offer employees competitive total compensation packages, which may include annual bonus, commission for certain sales roles, equity, and health and welfare benefits. Coordinate with internal teams, including Process, Testing, and Automation, to review and respond to their DFM items, optimize designs, and establish appropriate assembly and test workflows.
NewStaff Maas Backend Engineer BitdeerStaff Maas Backend EngineerSan Jose, CAWe are seeking a Staff Backend Engineer to take our Model-as-a-Service (MaaS) working system and re-architect it into a commercial, globally distributed, multi-tenant token service - one that sustains at least millions of monthly active users, high sustained token throughput per GPU, and invoice-grade accounting, with scalability, reliability, and observability engineered in deliberately rather than absorbed under load. Global Topology and Reliability (SLOs): Scale the platform to a globally distributed architecture featuring regional inference pools, capacity-aware failovers, and an active-active control plane.
NewSr. Cloud Support Engineer - Weekend Shift Crusoe EnergySr. Cloud Support Engineer - Weekend ShiftSan Francisco, CA$145,000–$175,000 / yearWe're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved - people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services. As a Cloud Support Engineer, you'll play a crucial role in empowering our customers to leverage this technology for groundbreaking advancements in fields like AI/ML, physics simulations, and computational biology.
NewSenior GPU Systems & Fabric Engineer BitdeerSenior GPU Systems & Fabric EngineerSan Jose, CAThis role requires deep expertise in Linux kernel internals, GPU architectures, and high-speed interconnects, as you will be tasked with transforming raw, bare-metal compute resources into scalable, resilient, and multi-tenant cloud primitives. Build and manage automated hardware remediation pipelines using DCGM telemetry to proactively identify, isolate, and reset degraded GPU/NIC components before they impact production jobs.
Hardware Engineer - Test Development Artech LLCHardware Engineer - Test DevelopmentFoster City, CA$60–$79.50 / hourFamiliarity with Data Acquisition systems (UEI and/or Keysight preferred) and Requirement-Test Case management tools (Polarion preferred). Test Infrastructure Development including custom test assets such as load simulators, test harnessing, and python scripts.
Lead Performance Engineer Artech LLCLead Performance EngineerSunnyvale, CA$40–$60 / hourThe ideal candidate has hands-on experience with JMeter, Load runner, K6, Gatling, Data Dog, Dynatrace skills, and solid expertise in API, Web and automation testing. •Design, develop, and maintain performance test scripts using Apache JMeter or similar tools for web, API, and backend systems.
Cloud Engineer Artech LLCCloud EngineerFoster City, CA$80–$106.50 / hourTroubleshoot and resolve day-to-day AWS support tickets across compute, storage, and networking for cross-functional customers (engineering tools, HPC/simulation workloads, mission control systems, data platform, manufacturing systems). We're looking for a Cloud Engineer to help operate and support our AWS environment, working closely with the Cloud Engineering team and partnering with engineering teams across IT & Applications, Infrastructure, Product Software, Data Platform, and Manufacturing Operations.
RF Hardware Engineer - Multi-Radio Coexistence Artech LLCRF Hardware Engineer - Multi-Radio CoexistenceCupertino, CA$100–$100.45 / hourKnowledge of RF component characteristics and experience using lab equipment including spectrum analyzer, network analyzer, signal generator, oscilloscopes, call box. Design bench experiments, collect performance data & analyze wireless logs across various RF parameters to derive HW specification for various multi-radio coexistence use-cases.