Senior AI Inference Engineer - Model Optimization & Deployment

Zoox
  • Foster City, CA
    30+ days ago

    Job Description

    The Perception team is pioneering the development of a multi-modality foundation model to drive the next generation of autonomous system intelligence.

    As a Model Optimization & Deployment Engineer, you will focus on bringing highly efficient, production-ready large-scale models to our on-vehicle stack. We are looking for experts with hands-on experience in compressing, accelerating, and deploying complex models (LLMs, VLMs, or FMs) for power- and thermal-constrained vehicle SOCs. You will optimize the ML models, write custom CUDA kernels, and build highly concurrent inference code to ensure real-time, deterministic execution on edge devices.

    Numbers & Facts

    LocationFoster City, CA

    Skills

    • Artificial Intelligence (AI)unmatched
    • CUDA (Compute Unified Device Architecture)unmatched
    • Kernel Programmingunmatched
    • Modalityunmatched

    Be found by employers

    5,500+ employers search our resume database daily. Add yours to get found by recruiters looking for candidates like you.

    Level up your application

    Professional resume templates

    Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.

    Free resume templates

    Free resume builder

    Improve your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.

    Free resume builder