NewDirector of Global Category Management: Workplace, Infrastructure & Lab Operations Portfolio - (M6) Applied Materials IncDirector of Global Category Management: Workplace, Infrastructure & Lab Operations Portfolio - (M6)Santa Clara, CA$160,000–$220,000 / yearThe Director is responsible for defining portfolio strategy, building organizational capability, establishing governance, developing talent, leading supplier relationship management programs, driving business partnership excellence, and delivering measurable value across Workplace, Infrastructure, Real Estate, Construction, Laboratory Operations, Technical Facilities, Engineering Services, Energy, Security, Facilities Services, MRO, and related categories. Operating within Applied Materials' center-led category management model, the Director owns portfolio outcomes while leading a global organization of category leaders, category managers, and supporting sourcing organizations to deliver value, improve stakeholder experience, reduce risk, and enable enterprise growth.
Senior Systems Engineer - AI Infrastructure Clockwork.ioSenior Systems Engineer - AI InfrastructurePalo Alto, CA$150,000–$230,000 / yearClockwork is pioneering a software-driven approach to AI fabrics by delivering cross-stack observability to catch and quickly resolve problems, workload fault tolerance to keep jobs running through failures, and performance acceleration that dynamically routes and paces traffic to avoid congestion. In addition to cash compensation, this role is eligible to participate in the company's equity program , which may include stock options granted in accordance with the company's equity plan and subject to approval and applicable vesting schedules.
Senior Mechanical Engineer, Test Infrastructure Form EnergySenior Mechanical Engineer, Test InfrastructureBerkeley, CAYou will be invested in developing and deploying novel, state-of-the-art test equipment including the design of mechanical and fluid handling systems, and electronics test stands supporting Hardware In the Loop (HIL) testing, while collaborating with the electrical subteam and supporting our operations team in keeping our tests online. Activities including designing, drafting, and manufacturing components & assemblies; engineering analysis including FEA, heat transfer, fluid mechanics, material selection, and manufacturability; communicating with stakeholders.
ML Infrastructure Engineer AppLovinML Infrastructure EngineerPalo Alto, CA$124,000–$250,000 / yearTo deliver on this mission, our global team is composed of team members with life experiences, backgrounds, and perspectives that mirror our developers and customers around the world. As a member of our software engineering infra team, you'll solve technical challenges, including upgrading and implementing state-of-the-art software infrastructure.
Staff Infrastructure Engineer IvoStaff Infrastructure EngineerSan Francisco, CaliforniaRun Kubernetes like it's your own startup within the startup — own multi-cluster, multi-region deployments across AWS/GCP/Azure, with failover and disaster recovery that actually works when it matters (not just on paper). Office Perks: Enjoy a vibrant Downtown San Francisco office with catered lunch five days a week, premium snacks and coffee, an in-building gym, and a dog-friendly environment.
Senior Infrastructure Engineer IvoSenior Infrastructure EngineerSan Francisco, CaliforniaRun Kubernetes like it's your own startup within the startup — own multi-cluster, multi-region deployments across AWS/GCP/Azure, with failover and disaster recovery that actually works when it matters (not just on paper). Office Perks: Enjoy a vibrant Downtown San Francisco office with catered lunch five days a week, premium snacks and coffee, an in-building gym, and a dog-friendly environment.
Senior/Staff Software Engineer - Infrastructure And Devops (Bay Area) FortanixSenior/Staff Software Engineer - Infrastructure And Devops (Bay Area)Santa Clara, CA$155,000–$230,000 / yearOur unified data security platform addresses vulnerabilities in hybrid multicloud environments, defends against threats, and makes it easier to discover, assess, and fix data exposure risks. Our commitment to solving the world’s toughest data security challenges has earned Fortanix multiple Cybersecurity Excellence and Innovation Awards, as well as recognition from industry giants such as NVIDIA, Microsoft, Intel, ServiceNow, and Snowflake.
Machine Learning Infrastructure Engineer Institute of Foundation ModelsMachine Learning Infrastructure EngineerSunnyvale, CaliforniaYou’ll work side-by-side with world-class researchers and engineers to: • Extend distributed training frameworks (e.g., DeepSpeed, FSDP, FairScale, Horovod) • Implement distributed optimizers from mathematical specs • Build robust config + launch systems across multi-node, multi-GPU clusters • Own experiment tracking, metrics logging, and job monitoring for external visibility • Improve training system reliability, maintainability, and performance • While much of the work will support large-scale pre-training, pre-training experience is not required. Strategic and innovative problem-solving skills will be instrumental in establishing MBZUAI as a global hub for high-performance computing in deep learning, driving impactful discoveries that inspire the next generation of AI pioneers.
Member of Technical Staff: Data Infrastructure & Data Operations Walden RoboticsMember of Technical Staff: Data Infrastructure & Data OperationsSan Francisco, CaliforniaAs a senior IC, you'll own the architecture and reliability at the core of that platform—ingestion, storage, transformation, and the data operations that keep quality high at scale—working hand in hand with the teams who produce and depend on that data. Participation in E-Verify does not limit your right to work and verification will only be completed after you become an employee with Walden Robotics.
NewSenior Research Engineer, Foundation Model Training Infrastructure NvidiaSenior Research Engineer, Foundation Model Training InfrastructureSanta Clara, CAWays to stand out from the crowd: Master's or PhD's degree in Computer Science, Robotics, Engineering, or a related field; Demonstrated Tech Lead experience, coordinating a team of engineers and driving projects from conception to deployment; Strong experience at building large-scale LLM and multimodal LLM training infrastructure; Contributions to popular open-source AI frameworks or research publications in top-tier AI conferences, such as NeurIPS, ICRA, ICLR, CoRL. What we need to see: Bachelor's degree in Computer Science, Robotics, Engineering, or a related field; 10+ years of full-time industry experience in large-scale MLOps and AI infrastructure; Proven experience designing and optimizing distributed training systems with frameworks like PyTorch, JAX, or TensorFlow.
Mechanical Engineer, Test Infrastructure BelcanMechanical Engineer, Test InfrastructureBerkeley, caDerisking & Validation: Work closely with the Design Engineering and Test teams to translate high-level testing requirements into functional mechanical hardware that identifies failure modes early in the development cycle. * System-Level Integration: Create test environments that accurately simulate real-world conditions for our battery packs and auxiliary systems (pumps, manifolds, and power electronics).
Site Reliability Engineer - Hardware Infrastructure NvidiaSite Reliability Engineer - Hardware InfrastructureSanta Clara, CAAssist teams in responding to high severity incidents, driving root cause analysis, crafting high-quality postmortems, and developing post-incident corrective actions. At NVIDIA, Site Reliability Engineering provides a rare chance to define, develop, and support large-scale production systems with high efficiency and availability.
Senior Technical Program Manager - Infrastructure/ Platform/ Product Artech LLCSenior Technical Program Manager - Infrastructure/ Platform/ ProductFoster City, CA$70–$78.55 / hourYou will partner with senior engineering and product leadership to drive strategic initiatives that span multiple teams, systems, and business lines. Serve as a strategic partner to VP- and Director-level engineering and product leaders, influencing roadmaps, resourcing decisions, and organizational priorities.
Technical Program Manager, AI Infrastructure Character AITechnical Program Manager, AI InfrastructureRedwood City, CAThis role is a good match for someone who thrives in technically complex environments, enjoys bringing structure to ambiguous problem spaces, and can drive alignment across deeply cross-functional teams. In this role, you'll partner closely with engineering, research, and product teams to shape infrastructure strategy, align roadmaps, and drive end-to-end execution for critical initiatives across training, evaluation, and inference.
Senior Software Engineer, Backend (Infrastructure) Otter.aiSenior Software Engineer, Backend (Infrastructure)Mountain View, CaliforniaUsing artificial intelligence, Otter generates real-time automated meeting notes, summaries, and other insights from in-person and virtual meetings - turning meetings into accessible, collaborative, and actionable data that can be shared across teams and organizations. As part of our dynamic engineering team, you'll collaborate closely with AI researchers, product managers, and technologists to deliver innovative solutions that transform the experiences of end-user professionals and enterprise clients across diverse industries.
NewSenior/Staff Software Engineer, Platform Infrastructure Verkada IncSenior/Staff Software Engineer, Platform InfrastructureSan Mateo, CA$180,000–$280,000 / yearIdentify and lead critical efforts related to scalability, reliability and efficiency; Influence the features and direction of our platform with your own ideas; Provide technical support for engineers on team; Align with product and org objectives, and coordinate with cross-functional teams on delivering key results. We've got serious momentum in the market: more than 30,000 customers (including 100+ of the Fortune 500), a $5.8B valuation, more than $1 billion in annualized bookings, and backing from CapitalG, Sequoia Capital, General Catalyst, Felicis Ventures, Next47 and more.
Commissioning Program Manager, Infrastructure Delivery OpenAICommissioning Program Manager, Infrastructure DeliverySan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. You will work closely with commissioning and construction teams to translate field execution needs into practical playbooks, workflows, templates, metrics, and reporting mechanisms that teams can use from construction readiness through testing and turnover.
Infrastructure Architect Cloud Senior BluZincInfrastructure Architect Cloud SeniorSan Jose, CaliforniaNew versions of the application are being developed and these cloud versions would allow managed service providers to support multiple customers, allow enterprise customers to manage software through a single pane of glass across many locations, and support customers running in cloud environments like AWS and Azure. You will work with the Product Management organization to develop a cloud strategy and then work with the rest of the Engineering team to develop and implement an architecture and new management plane to fulfil that strategy.
Senior Technical Product Manager - DGX Enterprise Infrastructure And Cloud-Native Operations NvidiaSenior Technical Product Manager - DGX Enterprise Infrastructure And Cloud-Native OperationsSanta Clara, CAWhat We Need to See: Enterprise Data Center DNA: 12+ years demonstrated ability in Product Management,,with specific around on-premise infrastructure, private cloud, or large-scale systems management. As you define the NVIDIA Datacenter Experience, you will be positioned on a direct leadership track, with the explicit expectation to transition into formal people management as the team expands.
Staff Technical Program Manager, AI Infrastructure GMStaff Technical Program Manager, AI InfrastructureSunnyvale, CaliforniaIn this role, you will own strategy and execution for large-scale ML infrastructure – including training pipelines, model lifecycle management, compute orchestration, and platform reliability – that power next-generation autonomy models. Benefit options include medical, dental, vision, Health Savings Account, Flexible Spending Accounts, retirement savings plan, sickness and accident benefits, life insurance, paid vacation & holidays, tuition assistance programs, employee assistance program, GM vehicle discounts and more.