Senior AI Inference Engineer - Model Optimization & Deployment

Zoox

  • Foster City, CA
  • 30+ days ago
    Want to know if you’re a fit?
    Upload your resume and let our AI show you.

    Skills

    • Artificial Intelligence (AI)unmatched
    • CUDA (Compute Unified Device Architecture)unmatched
    • Kernel Programmingunmatched
    • Modalityunmatched

    Description

    The Perception team is pioneering the development of a multi-modality foundation model to drive the next generation of autonomous system intelligence.

    As a Model Optimization & Deployment Engineer, you will focus on bringing highly efficient, production-ready large-scale models to our on-vehicle stack. We are looking for experts with hands-on experience in compressing, accelerating, and deploying complex models (LLMs, VLMs, or FMs) for power- and thermal-constrained vehicle SOCs. You will optimize the ML models, write custom CUDA kernels, and build highly concurrent inference code to ensure real-time, deterministic execution on edge devices.

    Numbers & Facts

    LocationFoster City, CA

    Similar Jobs

    See more jobs