You will: Build and deploy custom AI solutions on NeoCloud platforms and NVIDIA Cloud Partners (NCPs), including distributed training, inference optimization, and MLOps pipelines, Act as a primary technical contact for internal and external customers and partners, guiding joint engagements, ensuring the success of initiatives on DGX Cloud, and solving complex problems in production, Work closely with the teams building the infrastructure software and accelerated frameworks that support today’s most compelling AI applications, Profile and tune large-scale training and inference workloads on NCP platforms, leading efforts to reduce latency, cost, and operational risk, and. In this role, you will develop innovative solutions that advance AI infrastructure capabilities, advise infrastructure experts on the demands of ML workloads, help practitioners diagnose and solve full-stack AI and ML system problems, and work on a team with direct responsibility for the success of internal and external customers’ AI and ML initiatives, including LLM performance evaluation and supporting new hardware in open-source frameworks.