Sr. Field Applications Engineer, Datacenter & AI Systems Debug and Deployment Support

Advanced Micro Devices, Inc
  • Austin, Texas
    22 days ago

    Job Description

    Overview:

    ADVANCE YOUR CAREER. ADVANCE THE WORLD. 

    At AMD, we believe technology has the power to solve the world’s most important challenges. From advancing healthcare and scientific discovery to powering AI and the technologies people rely on every day, innovation at AMD is shaping the future. 

     

    Whether you’re designing next-gen processors, enabling AI breakthroughs, or bringing leading edge products to market, every role at AMD contributes to something bigger — technology that moves the world forward. Join us and, together, we’ll advance your career.

    Responsibilities:

     

    THE ROLE:

    We are seeking a motivated and technically curious Field Applications Engineer to join our Global Partner Support team. This role is ideal for an early career engineer who is motivated to grow deep expertise at the intersection of data center hardware, system software, and AI workloads.

    In this role, you will work alongside experienced FAEs and engineering teams to debug, triage, and resolve complex issues involving AMD Data Center GPUs and the AI software stack. You will gradually take on increasing ownership as your technical depth and confidence grow. This is a hands-on role designed for engineers who enjoy learning through real-world problem solving in fast-paced, high-impact environments.

     

    THE PERSON:

    The ideal candidate is an intellectually curious engineer with a passion for solving challenging technical problems and building expertise across modern computing platforms. You thrive in collaborative environments, enjoy working directly with customers and partners, and are energized by opportunities to learn new technologies.

    You possess strong analytical and debugging skills and are comfortable navigating issues that span hardware, software, systems, and AI workloads. You communicate effectively with both technical and non-technical audiences and approach problems with a customer-first mindset. Most importantly, you are motivated to grow into a trusted technical advisor supporting some of the industry's most advanced AI and accelerated computing deployments.

    KEY RESPONSIBILITIES:

    • Participate in system-level debugging of GPU, driver, networking, and AI software issues across single-node and multi-node environments, progressively taking ownership of complex debug efforts.
    • Support post-sales technical engagements with cloud providers, OEMs, ODMs, enterprise customers, and strategic partners.
    • Reproduce customer and partner issues in lab environments and analyze logs, core dumps, performance data, and system behavior to determine root causes.
    • Collaborate closely with engineering, product, and validation teams to drive issue resolution, validate fixes, and provide critical field feedback.
    • Document debug methodologies, technical findings, and solutions while contributing to internal knowledge bases and best practices.
    • Support large-scale cluster bring-up activities, new product introductions, and strategic customer deployments as experience grows.
    • Develop and deliver technical training sessions and collateral covering new products, feature enhancements, and advanced troubleshooting methodologies.

    PREFERRED EXPERIENCE:

    • Bachelor's or Master's degree in Computer Science, Electrical Engineering, Computer Engineering, or a related technical field, or equivalent practical experience.
    • Knowledge of CPU and GPU architecture concepts, memory hierarchies, and system-level behavior.
    • Experience with Linux systems, command-line tools, system administration, and low-level debugging using tools such as kernel logs and gdb.
    • Relevant experience gained through internships, co-ops, academic projects, or early-career technical roles.
    • Working understanding of GPU architecture and GPU-accelerated workloads, including experience using AMD or NVIDIA GPU platforms.
    • Foundational understanding of server and accelerator architectures, including PCIe topologies, CPU-GPU interconnects, memory hierarchies, and NUMA concepts.
    • Exposure to AI or HPC workloads and familiarity with frameworks such as PyTorch or TensorFlow.
    • Proficiency in at least one programming language, such as Python, C/C++, or Java.
    • Experience using source control systems such as Git.
    • Strong analytical, problem-solving, and communication skills with the ability to work effectively alongside highly technical customer and partner teams.

    WAYS TO STAND OUT:

    • Hands-on experience with modern data center GPU platforms, including AMD Instinct accelerators and comparable NVIDIA solutions.
    • Direct experience with GPU programming frameworks such as ROCm HIP or CUDA.
    • Experience using performance analysis and debugging tools such as rocProf, ROCgdb, AMDuProf, or PyTorch Profiler.
    • Familiarity with high-performance networking technologies including InfiniBand, RoCE, and RDMA concepts.
    • Knowledge of distributed GPU communication frameworks such as RCCL and NCCL.
    • Experience with containerization and orchestration technologies such as Docker, Kubernetes, or Slurm.
    • Experience supporting or debugging HPC clusters, AI training environments, inference deployments, or proof-of-concept systems.
    • Previous experience working with OEMs, ODMs, cloud service providers, or large enterprise customers.

    ACADEMIC CREDENTIALS:

    • Bachelor's or Master's degree in Computer Engineering, Electrical Engineering, Computer Science, or a related technical discipline preferred.

    LOCATION:

    Austin, TX

     

    This role is not eligible for visa sponsorship.

    #LI-RF1

    Qualifications:

    Benefits offered are described:  AMD benefits at a glance.

     

    AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law.   We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.

     

    AMD may use Artificial Intelligence to help screen, assess or select applicants for this position.  AMD’s “Responsible AI Policy” is available here.

     

    This posting is for an existing vacancy.

    Numbers & Facts

    LocationAustin, Texas

    Skills

    • Analysis Skillsunmatched
    • Artificial Intelligence (AI)unmatched
    • Best Practicesunmatched
    • C Programming Languageunmatched
    • C++ Programming Languageunmatched
    • CPU (Central Processing Unit)unmatched
    • CUDA (Compute Unified Device Architecture)unmatched
    • Channel Supportunmatched
    • Cloud Computingunmatched
    • Command Lineunmatched
    • Communication Skillsunmatched
    • Computer Engineeringunmatched
    • Computer Scienceunmatched
    • Computer Systemsunmatched
    • Core Loggingunmatched
    • Debugging Skillsunmatched
    • Debugging Toolsunmatched
    • Device Driversunmatched
    • Dockerunmatched
    • Electrical Engineeringunmatched
    • Environmental Impactunmatched
    • Environmental Issuesunmatched
    • Field Trialsunmatched
    • GDB (Gnu Debugger)unmatched
    • GPU (Graphics Processing Unit)unmatched
    • Gitunmatched
    • Healthcareunmatched
    • Identify Issuesunmatched
    • Javaunmatched
    • Kernel Programmingunmatched
    • Knowledge Baseunmatched
    • Laboratory Analysisunmatched
    • Linux Operating Systemunmatched
    • Memory Hardwareunmatched
    • Militaryunmatched
    • Multiplatform/Cross-Platformunmatched
    • Network Operations Centerunmatched
    • Network Softwareunmatched
    • OEM (Original Equipment Manufacturer)unmatched
    • Original Design Manufacturer (ODM)unmatched
    • PCI Express (PCI-E)unmatched
    • Performance Analysisunmatched
    • Post-Salesunmatched
    • Problem Solving Skillsunmatched
    • Product Testingunmatched
    • Product/Service Launchunmatched
    • Programming Languagesunmatched
    • Proof of Conceptunmatched
    • Python Programming/Scripting Languageunmatched
    • Recruiting/Staffing Agencyunmatched
    • Return on Capital Employed (ROCE)unmatched
    • Root Cause Analysisunmatched
    • Sales Supportunmatched
    • Sales/Support Engineering (SE)unmatched
    • Server Architectureunmatched
    • Source Code Control System (SCCS)unmatched
    • Systems Administration/Managementunmatched
    • Team Playerunmatched
    • Technical Salesunmatched
    • Technical Supportunmatched
    • Technical Trainingunmatched
    • Technical/Engineering Designunmatched
    • Topologyunmatched

    Be found by employers

    5,500+ employers search our resume database daily. Add yours to get found by recruiters looking for candidates like you.

    Level up your application

    Professional resume templates

    Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.

    Free resume templates

    Free resume builder

    Improve your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.

    Free resume builder