Engineering Manager, AI Models Infrastructure IntercomEngineering Manager, AI Models InfrastructureDublin, CAFounded in 2011, Fin became one of the fastest growing companies and remains one of the largest private software companies in the world with nearly 30,000 global businesses using our products to transform their customer support. Fin can also be combined with our Helpdesk to become a complete solution called the Fin Customer Service Suite, which provides AI enhanced support for the more complex or high touch queries that require a human agent.
Data selection and quality evaluation for biological foundation models InceptiveData selection and quality evaluation for biological foundation modelsPalo Alto, CaliforniaOur team brings together vast expertise in molecular biology, machine learning, and software engineering, and we are all working towards becoming antedisciplinary, meaning we deepen the knowledge we have in our area of expertise while also expanding our knowledge of completely new fields. We approach our goals with a Beginner's mind, humbly and with fresh eyes, and aim to become the pioneers of a new discipline rooted in biology as much as in deep learning, whose impact will be realized together with out-of-the-box thinkers in business and entrepreneurship, defying established categorizations.
Staff Software Engineer- Foundation Model Inference DataBricksStaff Software Engineer- Foundation Model InferenceSan Francisco, CA$190,000–$265,000 / yearThe impact you will have: Build LLM infrastructure powering large-scale inference workloads for customers through partner models (OpenAI, Anthropic, Gemini) and self-hosted models (Qwen, GPT-OSS, Llama). More than 10,000 organizations worldwide - including Comcast, Condé Nast, Grammarly, and over 50% of the Fortune 500 - rely on the Databricks Data Intelligence Platform to unify and democratize data, analytics and AI.
NewMember of Technical Staff (Model Behavior) Perplexity AI IncMember of Technical Staff (Model Behavior)Palo Alto, CAContext Engineering: Design, test, and optimize the prompts, skills, tools, and memory that shape Perplexity responses across products, features, and use cases. Were hiring software engineers for the Model Behavior team to help shape how Perplexity's AI products behave: the style of their responses, and the way they use tools, skills, and memory.
Staff Software Engineer, Foundational Model Serving DataBricksStaff Software Engineer, Foundational Model ServingSan Francisco, CA$192,000–$260,000 / yearWe're looking for engineers who have owned high scale operational sensitive systems like customer facing APIs, Edge Gateways, ML Inference, or similar services and have an interest in getting deep building LLM APIs and runtimes at scale. Contribute directly to key components across the serving infrastructure - from working in systems like vLLM and SGLang to creating token based rate limiters and optimizers - ensuring smooth and efficient operations at scale.
Uncertainty Quantification for Surrogate Models Postdoctoral Researcher Lawrence Livermore National LaboratoryUncertainty Quantification for Surrogate Models Postdoctoral ResearcherLivermore, CAAn employees position within the salary range will be based on several factors including, but not limited to, specific competencies, relevant education, qualifications, certifications, experience, skills, seniority, geographic location, performance, and business or organizational needs. Our mission spans four critical national security areas: nuclear deterrence, threat preparedness, energy security, and multi-domain defense, empowering teams to take on the toughest challenges of today and tomorrow.
Associate Manager, Production Line Maintenance, Model 3, Body in White Tesla IncAssociate Manager, Production Line Maintenance, Model 3, Body in WhiteFremont, CA$128,000–$192,000 / yearThey must be capable of collaborating with cross-functional teams and department leaders to address concerns and issues that affect part quality, production efficiency, work cell ergonomics, equipment maintenance, etc. 5+ years of extensive experience in supervision and leadership, with troubleshooting and correcting industrial equipment maintenance - breakdowns and failures through Root Cause Analysis and effective Corrective Action implementation.
Engineering & Maintenance Manager, Model 3, General Assembly Tesla IncEngineering & Maintenance Manager, Model 3, General AssemblyFremont, CA$152,000–$228,000 / yearEquipment includes Siemens/Allen Bradley PLCs, Fanuc robots, automated dispensing, Atlas Copco fastening, electrical testing, among other technologies. Along with competitive pay, as a full-time Tesla employee, you are eligible for the following benefits at day 1 of hire: Medical plans > plan options with $0 payroll deduction.
Director, Frontier Model & AI Alliances Zenity LTDDirector, Frontier Model & AI AlliancesSan Francisco, CAEstablish multi-level, ongoing communication with target partners, product roadmap visibility, leadership access, and early awareness of changes that affect Zenity customers. We deliver full-lifecycle visibility, governance, detection, prevention, and response for AI agents from buildtime to runtime, across SaaS, home-grown platforms, and end-user devices.
Professional Engineering Trainee, PEAK Program, Model 3 (Summer 2026) Tesla IncProfessional Engineering Trainee, PEAK Program, Model 3 (Summer 2026)Fremont, CA$71,120–$106,680 / yearThe PEAK (Professional Engineering Accelerated Knowledge) Program provides early career engineers with the opportunity to develop their engineering skills in a challenging, fast-paced, and impactful manufacturing engineering role onsite at Tesla's Fremont, California facility. Execute Engineering Projects: Using advanced project management and statistical techniques taught in this program, you will bring significant impact to your engineering team and present your results to factory leadership.
Tool & Die Specialist, Model Y, Castings Tesla IncTool & Die Specialist, Model Y, CastingsFremont, CA$32–$60.70 / hourExperience with high-speed die-sets, including tooling material selection, and knowledge of CNC die surface machining, welding, and spotting. Along with competitive pay, as a full-time Tesla employee, you are eligible for the following benefits at day 1 of hire: Medical plans > plan options with $0 payroll deduction.
AI Inference Engineer Intern - Model Pruning quadric, IncAI Inference Engineer Intern - Model PruningBurlingame, CA$45–$60 / hourQuadric's co-optimized software and hardware is targeted to run neural network (NN) inference workloads in a wide variety of edge and endpoint devices, ranging from battery operated smart-sensor systems to high-performance automotive or autonomous vehicle systems. Unlike other NPUs or neural network accelerators in the industry today that can only accelerate a portion of a machine learning graph, the Quadric GPNPU executes both NN graph code and conventional C++ DSP and control code.
Principal Research Engineer, Model Training & Post-Training Inflection AI IncPrincipal Research Engineer, Model Training & Post-TrainingPalo Alto, CA$400,000–$550,000 / yearLead training and post-training strategy, including supervised fine-tuning, RLHF, DPO, GRPO, RLAIF, reward modeling, preference optimization, tool-use fine-tuning, distillation, synthetic data, and related methods. Inflection's models are central to our product and platform strategy, and we are looking for a hands-on technical leader to own the model-improvement loop from data and training through evals, post-training, release criteria, and production feedback.
Model Maker Artech LLCModel MakerFoster City, CA$50–$62.92 / hourYou'll work on highly dynamic projects that come to the prototyping team, interfacing with both prototyping teammates and internal clients to support rapid iteration for vehicle development. You will both bring and grow a prototyping mindset in addition to various capabilities including (but not limited to) metalwork, composites, and additive manufacturing.
DevOps Engineer - AI Model Evaluator MercorDevOps Engineer - AI Model EvaluatorSan Francisco, CaliforniaRemoteRegular use of AI coding agents like Cursor , Claude Code , Codex , Windsurf , Gemini CLI , or similar tools. Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks.
iOS Engineer - AI Model Evaluator MercoriOS Engineer - AI Model EvaluatorSan Francisco, CaliforniaRemoteRegular use of AI coding agents such as Cursor , Claude Code , Codex , Windsurf , Gemini CLI , or similar tools. Use frontier AI coding agents to complete and evaluate complex engineering tasks.
DevOps Engineer - AI Model Evaluator - AI Trainer MercorDevOps Engineer - AI Model Evaluator - AI TrainerSan Francisco, CaliforniaRemoteRegular use of AI coding agents like Cursor , Claude Code , Codex , Windsurf , Gemini CLI , or similar tools. Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks.
NewGNC Engineer, Advanced Modeling Xona Space Systems, Inc.GNC Engineer, Advanced ModelingBurlingame, CAWith Pulsar - the world's most advanced PNT satellite infrastructure in Low Earth Orbit - Xona will offer a future-proof, backwards-compatible global positioning system optimized for absolute precision, superior power, and robust protection. You will build nonlinear six-degree-of-freedom simulation environments, develop flexible-body and vehicle subsystem models, and validate those models using analytical, test, and flight data.
GNC Engineer, Advanced Modeling Xona Space SystemsGNC Engineer, Advanced ModelingBurlingame, CaliforniaWith Pulsar – the world’s most advanced PNT satellite infrastructure in Low Earth Orbit – Xona will offer a future-proof, backwards-compatible global positioning system optimized for absolute precision, superior power, and robust protection. You will build nonlinear six-degree-of-freedom simulation environments, develop flexible-body and vehicle subsystem models, and validate those models using analytical, test, and flight data.
Backend Engineer, Models MeterBackend Engineer, ModelsSan Francisco, CaliforniaTo make this possible, we don’t just need great models; we need infrastructure that gives those models clean, versioned, low-latency access to the right data, across training, evaluation, and deployment. As described on Meter.ai , we’re building models in a closed-loop system that takes (as input) real-time telemetry, logs, and events on the network to autonomously troubleshoot, improve performance, and resolve issues.
Senior Program Manager, API Business Models Autodesk Inc.Senior Program Manager, API Business ModelsSan Francisco, CA$124,000–$221,430 / yearEn tant que responsable de programme senior, Modèles économiques des API, vous travaillerez en collaboration avec les parties prenantes internes afin de définir, de mettre en place et de piloter les modèles économiques et les programmes de monétisation liés aux API d'Autodesk. Avec plus de 6,5 milliards de dollars de chiffre d'affaires et plus de 15 000 employés à travers le monde, Autodesk s'est imposé comme le leader mondial des technologies de conception et de fabrication pour les logiciels et les services qui offrent aux clients de meilleurs résultats grâce à l'automatisation et à l'analyse de données.
Enterprise Operating Model Senior Manager, Utilities Accenture PlcEnterprise Operating Model Senior Manager, UtilitiesMountain View, CAAccenture is a leading solutions and services company that helps the world's leading enterprises reinvent by building their digital core and unleashing the power of AI to create value at speed across the enterprise, bringing together the talent of our approximately 786,000 people, our proprietary assets and platforms, and deep ecosystem relationships. Our approach and our people put us at the front of the pack for architecting future-proof enterprise operating models, clean sheet organization designs, and advanced shared services - all embracing the future of work powered by technology, operations, GenAI & data & analytics.
Enterprise Operating Model Manager, Energy Accenture PlcEnterprise Operating Model Manager, EnergySan Francisco, CAAccenture is a leading solutions and services company that helps the world's leading enterprises reinvent by building their digital core and unleashing the power of AI to create value at speed across the enterprise, bringing together the talent of our approximately 786,000 people, our proprietary assets and platforms, and deep ecosystem relationships. Our approach and our people put us at the front of the pack for architecting future-proof enterprise operating models, clean sheet organization designs, and advanced shared services - all embracing the future of work powered by technology, operations, GenAI & data & analytics.
Sr. Staff ML Engineer, Autonomy Model Pretraining RivianSr. Staff ML Engineer, Autonomy Model PretrainingPalo Alto, California$265,000–$331,000 / yearFull timeRivian may use your Candidate Personal Data for the purposes of (i) tracking interactions with our recruiting system; (ii) carrying out, analyzing and improving our application and recruitment process, including assessing you and your application and conducting employment, background and reference checks; (iii) establishing an employment relationship or entering into an employment contract with you; (iv) complying with our legal, regulatory and corporate governance obligations; (v) recordkeeping; (vi) ensuring network and information security and preventing fraud; and (vii) as otherwise required or permitted by applicable law. Rivian may share your Candidate Personal Data with (i) internal personnel who have a need to know such information in order to perform their duties, including individuals on our People Team, Finance, Legal, and the team(s) with the position(s) for which you are applying; (ii) Rivian affiliates; and (iii) Rivian’s service providers, including providers of background checks, staffing services, and cloud services.
NewEnterprise Operating Model Senior Manager Accenture PlcEnterprise Operating Model Senior ManagerSan Francisco, CAAccenture is a leading solutions and services company that helps the world's leading enterprises reinvent by building their digital core and unleashing the power of AI to create value at speed across the enterprise, bringing together the talent of our approximately 786,000 people, our proprietary assets and platforms, and deep ecosystem relationships. Our approach and our people put us at the front of the pack for architecting future-proof enterprise operating models, clean sheet organization designs, and advanced shared services - all embracing the future of work powered by technology, operations, GenAI & data & analytics.
NewEnterprise Operating Model Manager Accenture PlcEnterprise Operating Model ManagerWalnut Creek, CAAccenture is a leading solutions and services company that helps the world's leading enterprises reinvent by building their digital core and unleashing the power of AI to create value at speed across the enterprise, bringing together the talent of our approximately 786,000 people, our proprietary assets and platforms, and deep ecosystem relationships. Our approach and our people put us at the front of the pack for architecting future-proof enterprise operating models, clean sheet organization designs, and advanced shared services - all embracing the future of work powered by technology, operations, GenAI & data & analytics.
Enterprise Operating Model Senior Manager, Energy Accenture PlcEnterprise Operating Model Senior Manager, EnergySan Francisco, CAAccenture is a leading solutions and services company that helps the world's leading enterprises reinvent by building their digital core and unleashing the power of AI to create value at speed across the enterprise, bringing together the talent of our approximately 786,000 people, our proprietary assets and platforms, and deep ecosystem relationships. Our approach and our people put us at the front of the pack for architecting future-proof enterprise operating models, clean sheet organization designs, and advanced shared services - all embracing the future of work powered by technology, operations, GenAI & data & analytics.
AI Engineer, Model Quality and Performance Cerebras SystemsAI Engineer, Model Quality and PerformanceSunnyvale, CAYoull use AI agents to spin up custom eval suites per customer use case, mine trajectories for representative test data, automate the repetitive parts of release qual, and help build performance datasets and benchmarking workflows for customer use cases. You will define what "good" looks like across the models we serve, building AI-driven systems to measure it at scale, and translating those signals into artifacts our customers and product team actually use.
Research Scientist - World Model LumaResearch Scientist - World ModelRedwood City, CaliforniaLuma already trains the strongest generative video models in the industry; the next step is turning those models into world models — interactive, controllable, physically faithful, and useful as a substrate for embodied reasoning. WHAT YOU'LL DO - Invent next-generation world model architectures — diffusion, transformer, autoregressive, or hybrid — with a particular focus on controllability and physical consistency.
Senior / Staff AI Research Scientist, Foundation Models RoboForceSenior / Staff AI Research Scientist, Foundation ModelsMilpitas, CaliforniaIn this role, you will develop algorithms that enable robots to understand their environment, interpret and execute tasks, and communicate seamlessly with humans — with a particular focus on building and training world models that allow robots to predict, plan, and generalize across complex physical tasks. Bonus Qualifications Experience with video generation or prediction models (e.g., diffusion-based video models, autoregressive video transformers) and their application to world modeling or synthetic data generation for robot learning.
Model Implementation Engineer SciforiumModel Implementation EngineerSan Francisco, CaliforniaBacked by multi-million-dollar funding and direct sponsorship from AMD with hands-on support from AMD engineers the team is scaling rapidly to build the full stack powering frontier AI models and real-time applications. This role is ideal for someone who thrives in fast-moving environments, enjoys working across a wide range of model architectures, and wants to play a key role in enabling rapid adoption of the latest advancements in AI.
Principal Product Manager, AI Model Security Microsoft CorpPrincipal Product Manager, AI Model SecurityMountain View, CA$139,900–$274,800 / yearOwn the model security roadmap: Define and prioritize the security hardening strategy for our frontier models across the full OWASP LLM threat surface - prompt injection (direct and indirect), data exfiltration, jailbreak resistance, system prompt leakage, training data extraction, and adversarial manipulation of agentic workflows. There is a different range applicable to specific work locations, within the San Francisco Bay area and New York City metropolitan area, and the base pay range for this role in those locations is USD $188,000 - $304,200 per year.
Senior Applied Scientist, Efficient LLM Inference & Model Optimization Nebius Group NVSenior Applied Scientist, Efficient LLM Inference & Model OptimizationPalo Alto, CA$195,200–$262,200 / yearInvent, evaluate, and productionize methods for quantization, QAT, distillation, speculative decoding, KV-cache reuse, KV-cache compression, long-context inference, MoE routing, and model/runtime co-optimization. Build high-quality prototypes in PyTorch, Triton, CUDA-adjacent tooling, or inference-serving frameworks, then work with MLEs and platform engineers to productionize them.
Model Designer OpenAI LLCModel DesignerSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.
Senior Radar Perception Engineer, Obstacle Foundation Models - Autonomous Vehicles NVIDIA CorpSenior Radar Perception Engineer, Obstacle Foundation Models - Autonomous VehiclesSanta Clara, CAEmbedded Optimization: Hands-on experience architecting and deploying DNN-based perception pipelines on embedded or real-time platforms, including optimization for latency, memory, and compute constraints, and familiarity with modern architectures (e.g., Transformers, BEV networks). NVIDIA GPUs run deep learning algorithms that simulate aspects of human intelligence, acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world.
Senior Software Engineer, Spatial Intelligence and Foundation Models NVIDIA CorpSenior Software Engineer, Spatial Intelligence and Foundation ModelsSanta Clara, CAYou will join a group of world-class robotics software and applied research engineers focused on geometric and semantic understanding, and reasoning for robots - building the perception systems that turn raw sensor data into actionable world understanding, shaping the future of physical AI! What you'll be doing: Design, implement, and deploy novel algorithms for spatial understanding, working on problems ranging from SLAM, structure-from-motion, optical flow, scene flow, and object reconstruction to training VLMs on a wide range of spatial reasoning skills.
Senior Research Scientist, Multimodal Foundation Models and Robotics NVIDIA CorpSenior Research Scientist, Multimodal Foundation Models and RoboticsSanta Clara, CADeep understanding of robot kinematics, dynamics, and sensors; Ability to safely operate robot hardware, lab equipment, and tools; Knowledge of control methods, including PID, model predictive control, and whole-body control; Familiarity with physics simulation frameworks such as MuJoCo and Isaac Sim; Robot hardware design and hands-on building experience. Hands-on training experience and publications in at least one of the following topics: LLMs; Large vision-language models; Video generative models and diffusion algorithms; or Action-based transformers.
Member of Technical Staff - Voice Model X CorpMember of Technical Staff - Voice ModelPalo Alto, CA$150,000–$450,000 / yearWork on pre-training and post-training of speech-language models, with targeted enhancements through supervised fine-tuning, reinforcement learning, and other techniques to ensure Grok Voice responses are accurate, factually grounded, natural and idiomatic in spoken style, conversational in tone, and fluent across multiple languages. Build and iterate a comprehensive evaluation framework covering objective metrics (accuracy, quality, latency, expressiveness), human preference studies, content factuality assessments, real-time interaction quality, and experimentation infrastructure to measure and improve performance.
Software Engineer, Models MeterSoftware Engineer, ModelsSan Francisco, CaliforniaIn addition to your customers, network engineers, you’ll partner closely with two research engineers who have deep ML backgrounds and a clear picture of what training data needs to look like. When a network engineer looks at a set of device stats and figures out it’s upstream packet loss — not a hardware failure, not a misconfiguration, specifically upstream packet loss — that reasoning lives in their head.
Full Stack Engineer, Scientific AI Models BenchlingFull Stack Engineer, Scientific AI ModelsSan Francisco, CAAlphaFold or Boltz2) predict structures, predict scientific properties, and generate new drug designs, acting as a design partner and a major time saver to scientists who are creating life-saving therapeutics. Projects you might work on include: adding new models as soon as they're published, improving model performance and scalability, and enabling scientists to automate their in-silico workflows by chaining models together into pipelines.
ML Scientist I / II, Foundation Models for Life Sciences Lila SciencesML Scientist I / II, Foundation Models for Life SciencesSan Francisco, CaliforniaFull-time U.S. employees receive a comprehensive benefits program including medical, dental, and vision coverage; employer-paid life and disability insurance; flexible time off with generous company wide holidays; paid parental leave; an educational assistance program; commuter benefits, including bike share memberships for office based employees; and a company subsidized lunch program. You will work on generative models spanning biological sequences, molecular structures, and multimodal experimental data, contributing to problem formulation, model design, training, evaluation, and integration into Lila's closed-loop discovery engine.
Senior / Principal ML Scientist, Foundation Models for Life Sciences Lila SciencesSenior / Principal ML Scientist, Foundation Models for Life SciencesSan Francisco, CaliforniaFull-time U.S. employees receive a comprehensive benefits program including medical, dental, and vision coverage; employer-paid life and disability insurance; flexible time off with generous company wide holidays; paid parental leave; an educational assistance program; commuter benefits, including bike share memberships for office based employees; and a company subsidized lunch program. LILA combines advanced AI models with proprietary AI Science Factory instruments into an operating system for science that executes the entire scientific method autonomously, accelerating discovery at unprecedented speed, scale, and impact across medicine, materials, and energy.
Partner Sales Director - AI Alliances - Model Providers Dynatrace IncPartner Sales Director - AI Alliances - Model ProvidersSan Francisco, CA$192,000–$240,000 / yearThis role operates within the AI Native Ecosystem Alliance charter and collaborating across AI native field sales, Product and Marketing to ensure AI partnerships generate partner-influenced pipeline and revenue, enterprise customer wins and new logos. This executive serves as Dynatrace''s primary alliance owner for a targeted list of prioritized partners including Anthropic OpenAI, Mistral, Groq, Cohere and other leading open source and commercial providers.
Research Product Manager, Model Behaviors AnthropicResearch Product Manager, Model BehaviorsSan Francisco, CAThis research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues.
Research Engineer, Production Model Post-Training AnthropicResearch Engineer, Production Model Post-TrainingSan Francisco, CAThis research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. As a Research Engineer on our Post-Training team, you'll train our base models through the complete post-training stack to deliver the production Claude models that users interact with.
Product Manager, Claude Code Model Performance AnthropicProduct Manager, Claude Code Model PerformanceSan Francisco, CAAs a Product Manager on Claude Code's model performance team, you will drive model launches end-to-end, build evals that measure what matters, and partner directly with researchers and product engineers to translate model improvements into developer-facing outcomes. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences.
Software Engineer, Model Performance Systems BaseTenSoftware Engineer, Model Performance SystemsSan Francisco, CAInfrastructure Validation: Create automated acceptance tests for new GPU clusters across x86 and ARM systems, measuring GPU memory bandwidth, networking throughput, and multi-node networking performance. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production.
Threat Intel Manager, Model Exploitation & Fraud AnthropicThreat Intel Manager, Model Exploitation & FraudSan Francisco, CAThe team includes established senior investigators who own our deepest technical casework, tracing distillation networks, reseller and proxy ecosystems, and financially motivated actors across first-party surfaces and third-party platforms; your job is to direct, resource, and amplify that work, not duplicate it. Direct, prioritize, and resource complex investigations into model distillation, unauthorized AI R&D usage, unauthorized access, coordinated account abuse, and fraud/scam networks, partnering with the senior investigators who lead the deepest technical casework and clearing blockers from their path.
Software Engineer - Model Products BaseTenSoftware Engineer - Model ProductsSan Francisco, CARESPONSIBILITIES: Design, build, and operate the Model APIs surface with focus on advanced inference capabilities: structured outputs (JSON mode, grammar-constrained generation), tool/function calling and multi-modal serving. Profile and optimize TensorRT-LLM kernels, analyze CUDA kernel performance, implement custom CUDA operators, tune memory allocation patterns for maximum throughput and optimize communication patterns across multi-GPU setups.
Senior Director, Device and SPICE Modeling NVIDIA CorpSenior Director, Device and SPICE ModelingSanta Clara, CAThe "Device & SPICE modeling" group is responsible to co-develop with foundries in advanced device technology in the following three areas: Achieve device performance targets in Speed, Leakage, and Variation, Release accurate SPICE model based on test chip data, Tape-out device/RO test structures in test chip for SPICE model validation, process readiness & product scribe line monitors. The Advanced Technology Group (ATG) at NVIDIA is an organization of process, CAD, and design engineers that works closely with key foundry partners and internal design groups.