Research Infrastructure - Member of Technical Staff SimileResearch Infrastructure - Member of Technical StaffSan Francisco, CaliforniaKeep a multi-node fleet healthy and saturated: node bring-up, topology-aware NCCL and RDMA configuration, scheduling and queue depth, storage lifecycle and checkpoint capacity, autoscaling of serving capacity alongside training jobs, and the alerting that tells us when GPUs are sitting idle. The world's leading companies use Simile to make business-critical decisions — from consumer leaders like CVS Health and Wealthfront to professional services organizations like Deloitte and Gallup — strategizing product launches, entering new markets, and forecasting earnings calls.
Engineering Manager, Drive OS Communication Infrastructure NvidiaEngineering Manager, Drive OS Communication InfrastructureSanta Clara, CAIt enables developers to build software frameworks and distributed applications that fully leverage NVIDIA hardware acceleration, seamlessly port across NVIDIA-powered platforms - from data center to vehicle - and scale effortlessly from simple deployments to complex, multi-system architectures, all while maintaining strong isolation guarantees and real-time monitoring across the entire processing pipeline. NvStreams is a high-performance inter-process and inter-chip communication library within NVIDIA DRIVE OS, providing the sophisticated, high-integrity data and sensor processing infrastructure that forms a core element of the NVIDIA DRIVE software platform.
Solutions Architect, OEM AI Factory Infrastructure NvidiaSolutions Architect, OEM AI Factory InfrastructureSanta Clara, CAAbility and eagerness to dig into unfamiliar territories to take on problems relying on experience from previous work with data center infrastructure experience, from hardware up through technology stack. Our work encompasses MEP (Mechanical Electrical Plumbing), Ethernet and Infiniband networking, DevOps, HPC/AI workloads, Cluster Administration and Site Reliability Engineering.
Solutions Architect - OEM AI Factory Infrastructure NvidiaSolutions Architect - OEM AI Factory InfrastructureSanta Clara, CAAbility and eagerness to dig into unfamiliar territories to take on problems relying on experience from previous work with data center infrastructure experience, from hardware up through technology stack. Our work encompasses MEP (Mechanical Electrical Plumbing), Ethernet and Infiniband networking, DevOps, HPC/AI workloads, Cluster Administration and Site Reliability Engineering.
NewSenior Solutions Architect, AI Infrastructure NvidiaSenior Solutions Architect, AI InfrastructureSanta Clara, CAPractical knowledge of sophisticated networking for AI data centers, including, GPUs, CPUs, InfiniBand and Ethernet fabric topologies, PCIe, host networking and switches. What you'll be doing: Working with Cloud Providers and Hyperscalers to develop, build and deploy compute and networking solutions based on NVIDIA groundbreaking AI infrastructure hardware.
Principal Hardware Product Manager – PCIe Accelerators & Server Infrastructure Bolt GraphicsPrincipal Hardware Product Manager – PCIe Accelerators & Server InfrastructureSunnyvale, CaliforniaThermal & Power Efficiency: Manage system trade-offs for highly efficient power targets (ranging from 120W single-slot boards up to 500W multi-chip server environments), ensuring optimal balance between air-cooled density and deployment in legacy data center infrastructures. Native Network Integration: Define the roadmap for our integrated, high-bandwidth communication fabrics, utilizing native 400GbE / 800GbE interfaces built directly into the accelerator to eliminate the need for discrete NICs, lowering latency and TCO.
Manager, Workplace Projects & Infrastructure The Voleon GroupManager, Workplace Projects & InfrastructureBerkeley, CaliforniaReporting to the Managing Director, Strategic Projects, the ideal candidate is a highly organized, execution-oriented project manager with experience managing complex workplace initiatives and driving results across multiple stakeholders. Proven ability to independently manage complex projects involving multiple stakeholders, budgets, timelines, and vendors, and drive successful outcomes in fast-paced and evolving environments.
Technical Marketing Manager - Data Center Infrastructure NvidiaTechnical Marketing Manager - Data Center InfrastructureSanta Clara, CAWe are looking for an exceptionally skilled Technical Marketing Manager - Data Center Infrastructure Specialist to join our dynamic Accelerated Computing team in Santa Clara, CA. The base salary range is 128,000 USD - 201,250 USD for Level 3, and 148,000 USD - 235,750 USD for Level 4. You will also be eligible for equity and benefits.
Senior Technical Product Marketing Manager - AI Infrastructure NvidiaSenior Technical Product Marketing Manager - AI InfrastructureSanta Clara, CAMarket Positioning- Be a go-to market positioning authority for NVIDIA product marketing, campaign marketing, investor relations, and sales teams, providing positioning mentorship and data-based evidence to support NVIDIA's best-in-class platform. Market Intelligence- Analyze and distill industry disclosures and market signals into succinct reports and actionable strategies to promote NVIDIA value propositions, and then deliver to audiences across NVIDIA.
NewSr. Product Marketing Manager- AI Infrastructure NetApp IncSr. Product Marketing Manager- AI InfrastructureSan Jose, CA$200,000–$225,000 / yearTechnical fluency in AI infrastructure concepts, including generative AI, RAG, data pipelines, model development workflows, hybrid multicloud architectures, data governance, and enterprise storage or data management platforms. As the only enterprise-grade storage service natively embedded in Google Cloud, AWS, and Microsoft Azure, we empower customers to run everything from traditional workloads to enterprise AI with unmatched performance, resilience, and security.
NewSenior Software Engineer, Data Infrastructure DecagonSenior Software Engineer, Data InfrastructureSan Francisco, California$200,000–$400,000 / yearOur technology enables industry-defining enterprises like Avis Budget Group, Block’s Cash App and Square, Chime, Oura Health, and Hunter Douglas to deploy AI agents that power personalized, deeply satisfying interactions across voice, chat, email, SMS, and every other channel. You'll own critical data pipelines and storage layers end‑to‑end, improve reliability and performance, and create paved paths that let every Decagon engineer work confidently with data at scale.
Staff Software Engineer, Core Infrastructure HarveyStaff Software Engineer, Core InfrastructureSan Francisco, CaliforniaAs a Staff Software Engineer on the Core Infrastructure team at Harvey, you'll play a critical role in designing and building new infrastructure systems while equally scaling and strengthening our existing infrastructure. You'll work in an environment balanced between innovation — building new systems — and operational excellence, ensuring that Harvey remains resilient and efficient as it scales products, regions, customers, and usage.
Senior Applied AI And AI Infrastructure Engineer - Chip Design And DFX NvidiaSenior Applied AI And AI Infrastructure Engineer - Chip Design And DFXSanta Clara, CADesign-for-X Engineering at NVIDIA works on groundbreaking innovations involving crafting creative solutions in AI for Chip Design and AI for Predictions in various use cases in manufacturing testing on some of the industry's most complex semiconductor chips. You will work on hard-to-solve problems in the Design For Test space which will involve application of algorithm design, using statistical tools to analyze and interpret complex datasets and explorations using Applied AI methods.
NewPrincipal Software Engineer - Infrastructure NVIDIA CorpPrincipal Software Engineer - InfrastructureSanta Clara, CAYou will establish the multi-year technical direction, identify organization-wide challenges, and drive large-scale transformations from initial strategy through architecture, implementation, production adoption, and measurable outcomes. Outstanding communication and technical leadership skills, with a proven ability to build consensus, influence senior leaders, mentor experienced engineers, and lead through ambiguity.
NewSoftware Engineer, Infrastructure, PhD, Early Career, 2027 Start Google LLCSoftware Engineer, Infrastructure, PhD, Early Career, 2027 StartSan Bruno, CAWe"re looking for engineers who bring fresh ideas from all areas, including information retrieval, distributed computing, large-scale system design, networking and data storage, security, artificial intelligence, natural language processing, UI design and mobile; the list goes on and is growing every day. XNote: By applying to this position you will have an opportunity to share your preferred working location from the following: Sunnyvale, CA, USA; Atlanta, GA, USA; Austin, TX, USA; Kirkland, WA, USA; Los Angeles, CA, USA; Madison, WI, USA; Mountain View, CA, USA; New York, NY, USA; Raleigh, NC, USA; Durham, NC, USA; San Bruno, CA, USA; Seattle, WA, USA.
Sr. Infrastructure Security Architect (Cio) Early Warning Services, LLCSr. Infrastructure Security Architect (Cio)San Francisco, CA$160,000–$200,000 / yearEarly Warning Services takes into consideration a variety of factors when determining a competitive salary offer, including, but not limited to, the job scope, market rates and geographic location of a position, candidate's education, experience, training, and specialized skills or certification(s) in relation to the job requirements and compared with internal equity (peers). Works with architecture teams to ensure that all newly developed and legacy infrastructure implementations are in line with security policy and are compliance to the required frameworks (ISO, PCI, NIST 800-53, etc.).
Forward Deployed Engineer, Infrastructure Specialist ReductoForward Deployed Engineer, Infrastructure SpecialistSan Francisco, CaliforniaWe’ve grown rapidly, increasing revenue 8x year over year and partnering with hundreds of companies, from leading AI teams like Harvey, Vanta, and Scale, to enterprise customers across FAANG and top trading firms. We provide a comprehensive toolkit for working with documents the way a human would, combining custom in-house and leading frontier models to power efficient and accurate document workflows.
Lead Principal Core Infrastructure Engineer Oracle CorpLead Principal Core Infrastructure EngineerSanta Clara, CA$135,200–$306,400 / yearOracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle''s differing products, industries and lines of business. Our team develops the core infrastructure that powers OCI Networking''s control and management services, solving complex challenges in network configuration management at cloud scale.
Tech Lead, Deployment & Operations — Custom Infrastructure OpenAITech Lead, Deployment & Operations — Custom InfrastructureSan Francisco, CaliforniaFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. This person will become the Directly-Responsible Individual responsible for bringing OpenAI’s custom silicon and associated systems into data center environments, ensuring successful deployment, bring-up, validation, operational readiness, and ongoing reliability at scale.
Hardware Test Infrastructure Engineer Efficient ComputerHardware Test Infrastructure EngineerSan Jose, CaliforniaThe ideal candidate must be a full-stack infrastructure engineer who is equally comfortable writing Go services, managing Kubernetes environments, debugging low-level test programs, and working with a wide variety of embedded hardware such as FPGA, Ambiq, Apollo, PIC32 and Efficient E1. ATLAS supports automated benchmarking and validation of both our company's products and competitor hardware, providing essential data that drives engineering and product decisions.