Founding AI Transformation Strategist ScribeFounding AI Transformation StrategistSan Francisco | NYC, CaliforniaRemoteLead engagements end-to-end with strategic enterprise customers: validate Optimize findings with stakeholders, build the business case, present ROI to executive sponsors, scope implementation projects, and drive outcomes through to completion. Walk into C-suite conversations with a data-backed point of view and establish credibility fast – advising on AI roadmap and the art of the possible with Scribe, Optimize, and agent-driven workflows (not demonstrating the product).
AI Inference Performance Engineer - New College Grad 2026 NVIDIAAI Inference Performance Engineer - New College Grad 2026Us, CaliforniaDrive industry benchmark results: own the end-to-end optimization pipeline, implement and integrate optimizations in quantization, scheduling, memory management, and distributed inference across TensorRT-LLM, SGLang, and vLLM. Deep understanding of LLM/VLM architectures and inference mechanics: attention, KV caching, batching strategies, decode-phase bottlenecks, speculative decoding, disaggregated serving etc.
Sr. Manager, AI Storage Solutions Advanced Micro Devices, IncSr. Manager, AI Storage SolutionsCaliforniaYou bring deep experience across storage and broad experience across data center GPUs, memory, CPUs, networking, and AI workloads, along with strong communication and strategic thinking skills. This is a hybrid work opportunity and an individual contributor role (not people management), based in Austin TX or Santa Clara CA with potential for Seattle WA as well, and relocation may be available for an exceptional alignment of background and preferred experience.
AI & Digital Workplace Engineer WMEAI & Digital Workplace EngineerCaliforniaOut of scope: workforce-facing AI literacy curriculum design and delivery (owned by the workforce AI adoption and enablement program); model fine-tuning and bespoke ML model development (this role evaluates and integrates rather than trains); broader cybersecurity policy authorship (the cybersecurity governance / GRC function); audit evidence integrity ownership (the IT compliance function); server, network, and telecom infrastructure engineering; HR-side competency frameworks and corporate L&D. In scope (AI): engineering of AI capabilities in the digital workplace stack; technical implementation of AI Acceptable Use policy controls; technical AI tool inventory and shadow-AI detection; AI-driven automation for IT itself; reference patterns, prompt libraries, and Architectural Decision Records for AI engineering; technical coaching and the engineering community of practice for AI.
AI Infrastructure DC Design Engineer II AstreyaAI Infrastructure DC Design Engineer IICaliforniaThe AI Infrastructure Datacenter Design Engineer Level 2 independently handles moderate-complexity data hall designs and infrastructure projects while supporting operational improvements and design optimization initiatives. $48,868.00 - $77,160.00 USD (Salary) Please note that the salary information provided herein is base pay only (gross); it does not include other forms of compensation which may or may not apply to this specific position, namely, performance-based bonuses, benefits-related payments, or other general incentives - none of which are guaranteed, may be subject to specific eligibility requirements, and are wholly within the discretion of Astreya to remit.
NewSenior Software Engineer, AI Storage NVIDIASenior Software Engineer, AI StorageUs, CaliforniaMore recently, GPU deep learning ignited modern deep learning — the next era of computing — with the GPU acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world. Good knowledge of Linux kernel internals, Filesystem, Object storage systems, Databases, Vector Databases.
NewSenior HPC AI Cluster Engineer NVIDIASenior HPC AI Cluster EngineerUs, CaliforniaYou will work with the latest Accelerated computing and Deep Learning software and hardware platforms, and with many scientific researchers, developers, and customers to craft improved workflows and develop new, leading differentiated solutions. Excellent knowledge of Windows and Linux (Redhat/CentOS and Ubuntu) networking (sockets, firewalld, iptables, wireshark, etc.) and internals, ACLs and OS level security protection and common protocols e.g.
AI Chip Toolchain Architect ForestownAI Chip Toolchain ArchitectCaliforniaConduct mid - and long - term planning for AI model deployment, model compression, and model quantization technologies to ensure the technical competitiveness of the AI chip toolchain in the fields of model quantization and model compression. Have an in - depth understanding of the future evolution of algorithms and application development models for autonomous driving and human - machine interaction, and have an in - depth understanding of the development models for algorithms and applications.
AI Factory CPU focused Solutions Architect NVIDIAAI Factory CPU focused Solutions ArchitectUs, CaliforniaFor this particular role, that means having a deep technical understanding of NVIDIA Reference Architectures, and using that understanding to enable customers adopting our CPU-based solutions as part of the overall NVIDIA AI Factory. Our day-to-day work involves helping our partners be successful in their adoption of end-to-end AI solutions using NVIDIA's compute, networking, and software stacks.
Senior Developer Relations Manager, Digital Health AI Research NVIDIASenior Developer Relations Manager, Digital Health AI ResearchUs, CaliforniaSkilled at distilling deeply technical concepts for audiences from researchers and engineers to product leaders and executives, with high empathy for developers and researchers and comfort working in a fast-paced, highly matrixed environment. A minimum of 12+ years of overall professional experience in AI/ML, applied research, software engineering, developer relations, technical partnerships, product management, or solution architecture, including 5+ years of direct hands-on experience in healthcare AI, digital health, or life sciences AI.
Developer Relations Manager – AI Natives NVIDIADeveloper Relations Manager – AI NativesUs, CaliforniaWork directly with startup founders and engineering teams to architect and optimize AI workloads using NVIDIA technologies including CUDA-X libraries, TensorRT-LLM, Triton Inference Server, NVIDIA NeMo, NIM microservices, and GPU-accelerated data processing frameworks. You will guide companies building agentic systems, AI copilots, developer tools, reasoning models, and multimodal AI applications, helping them accelerate training, optimize inference, and deliver world-class AI experiences to millions of users.
Senior Applied AI Engineer, Cybersecurity NVIDIASenior Applied AI Engineer, CybersecurityUs, CaliforniaExperience developing or optimizing agentic architectures, agent harnesses, orchestration systems, retrieval and context pipelines, or multi-agent approaches for cybersecurity or other complex operational use cases. Hands-on experience designing and developing modern AI systems using large language models, retrieval-augmented generation, agentic architectures, agent harnesses, or related approaches.
Senior Technical Program Manager - AI Acceleration NVIDIASenior Technical Program Manager - AI AccelerationUs, CaliforniaAs a Senior Technical Program Manager, you will play a critical role in turning this vision into reality — driving multi-functional programs, scaling systems, and delivering AI capabilities that transform how silicon is designed, verified, and brought to production. This role will also help enable AI productivity across hardware teams by making both internal and third-party AI tools available, understanding user requirements, and translating feedback back to internal platform teams and external partners.
Senior Technical Marketing Engineer, Enterprise AI Software NVIDIASenior Technical Marketing Engineer, Enterprise AI SoftwareUs, CaliforniaRefine developer, user, and agent journeys: Understand how developers, enterprise platform teams, partners, and customers, and their respective agents, consume NVIDIA AI software, then craft clear technical journeys supported by documentation, code examples, demos, and deployment guidance. The NVIDIA Enterprise Product Group builds AI solutions that help enterprises develop, deploy, and scale generative AI, agentic AI, retrieval-augmented generation, and accelerated data workflows from developers laptops to deployed in data centers, clouds, and AI factories.
Developer Relations Manager, Higher Education and Research - Foundational AI NVIDIADeveloper Relations Manager, Higher Education and Research - Foundational AIUs, CaliforniaIn this role, you will work directly with top researchers building frontier AI systems, including large language models, multimodal models, reasoning systems, training methods, inference systems, model serving, and scalable AI infrastructure. Experience with NVIDIA AI platforms, including CUDA, CUDA-X libraries, TensorRT-LLM, Triton Inference Server, NIM, NeMo, Megatron, Transformer Engine, NCCL, DGX, NVLink, InfiniBand, or NVIDIA AI Enterprise.
Director, AI Enablement NVIDIADirector, AI EnablementUs, CaliforniaAs an engineer director on AI Enablement, you will be responsible for enabling start-of-the-art AI technologies and help transform Nvidia to an AI native company with speed of light, for SW development, HW development and beyond. NVIDIA is hiring an engineering director to enable and accelerate AI adoption and agentic developments across various eng and non-eng workflows within the company.
Developer Relations Manager, Local AI Ecosystem NVIDIADeveloper Relations Manager, Local AI EcosystemUs, CaliforniaWorking knowledge of GPU acceleration and the software layers that shape local AI performance, such as inference frameworks (vLLM, SGLang, Llama.cpp, Ollama), model formats, quantization, memory management, and hardware-aware optimization. This role will work directly with the companies and collaborative software projects building the models, runtimes, tools, agent platforms, and applications that bring AI onto PCs, workstations, and personal AI.
Senior AI Architect, Foundation Models and SoC Co-Design – Autonomous Vehicles NVIDIASenior AI Architect, Foundation Models and SoC Co-Design – Autonomous VehiclesUs, CaliforniaYou will work with world-class AI researchers, silicon architects, and AV platform teams to identify the AI workloads that will define the next decade — and ensure NVIDIA platforms are architected to lead them. We are looking for a Senior AI Architect to help define the next generation of AI model paradigms for autonomous vehicles and shape how those models co-evolve with NVIDIA’s future embedded SoC architectures.
Senior Consulting Director, AI Security, Proactive Services (Unit 42) Palo Alto NetworksSenior Consulting Director, AI Security, Proactive Services (Unit 42)CaliforniaWorking closely with the Managing Director and CRM NAM Practice leader, along with peers in North America and other theaters and regions, you will provide executive oversight and strategic direction for how we evaluate, mature, and optimize security operations through our assessments across our client base — driving greater security outcomes, product integration, and transformation at scale. In this strategic leadership role, you will serve as the executive sponsor on key engagements, guide a high-performing team of Directors and consultants, and partner across Unit 42 and Palo Alto Networks to refine services, contribute to thought leadership, and elevate AI Security capabilities across industries.
Distinguished Engineer, Storage – AI Cloud NVIDIADistinguished Engineer, Storage – AI CloudUs, CaliforniaDesign the storage layer for workloads spanning the next several GPU generations, including disaggregated inference with storage-backed KV caching, large-scale write-once-read-many inference patterns, exabyte regional object stores, and cross-DC dataset versioning and copy management. Lead the multi-year technical plan for AI Cloud Storage expansion across NCPs — determine the reference architecture, capabilities, performance and durability SLOs, qualification methodology, and roadmap for the high-performance file, object, and block storage that each NCP must offer to qualify for NVIDIA GPU allocation.