NVIDIA Corp logo

Senior DevOps Engineer

NVIDIA Corp
  • Santa Clara, CA
    8 days ago

    Job Description

    As a Senior DevOps Engineer, you will help lead the evolution of infrastructure operations within our Networking Software group. Building on a strong Linux systems administration foundation, you will build, automate, and operate scalable platforms that support networking software development and testing. This role offers an outstanding opportunity to work with elite technology and collaborate with ambitious engineers across global sites. If you are passionate about automation, reliability, technical leadership, and continuous improvement, this is the perfect opportunity for you!

    What you'll be doing:

    • Build, provision, configure, and maintain scalable Linux infrastructure for networking feature creation and validation, including physical servers, network switches, virtualization platforms, containers, and remote-management interfaces.

    • Develop automation for infrastructure provisioning, configuration management, software deployment, upgrades, and day-to-day operations using infrastructure-as-code and configuration-management practices.

    • Build reusable tools and self-service capabilities that simplify infrastructure operations, improve engineering efficiency, and reduce repetitive manual work.

    • Diagnose and resolve complex issues spanning hardware, firmware, operating systems, virtualization, containers, storage, network communications, and application environments.

    • Implement monitoring, observability, capacity management, and reliability practices to improve infrastructure performance, availability, and operational readiness.

    • Partner with engineering, IT, facilities, security, and network teams to define technical standards, maintain documentation and runbooks, and establish scalable operational processes.

    • Provide technical leadership, guide infrastructure initiatives, and mentor team members in automation, troubleshooting, and operational guidelines.

    What we need to see:

    • Bachelor's degree in Computer Science, Information Technology, Engineering, or equivalent experience.

    • 6+ years of experience in systems engineering, DevOps, site reliability engineering, or infrastructure operations, including significant hands-on experience coordinating production or engineering Linux environments.

    • Experience working in the semiconductor industry or a hardware-focused engineering environment, with deep hands-on expertise in bare-metal Linux systems and server components-including CPUs, GPUs, memory, PCIe devices, NICs, storage, BIOS/UEFI, BMC/IPMI/Redfish, power, and cooling.

    • Proficiency in generative AI tools and skill in applying them effectively to automation, troubleshooting, documentation, operational analysis, and engineering efficiency.

    • Skilled at diagnosing complex issues across hardware, firmware, and operating-system layers.

    • Strong data-center networking knowledge, including TCP/IP, DNS, DHCP, VLANs, routing, switching, firewalls, and network troubleshooting tools.

    • Strong analytical, problem-solving, written communication, and cross-departmental collaboration skills, with the ability to guide technical initiatives to completion.

    Ways to stand out from the crowd:

    • Experience managing Linux KVM/QEMU virtualization, Kubernetes clusters, multi-user engineering lab environments, NFS or distributed storage systems, and automated OS or cluster provisioning platforms.

    • Experience supporting fast-growing engineering labs, large-scale data-center environments, or globally distributed infrastructure.

    • Familiarity with observability platforms, including metrics, logging, tracing, alerting, incident management, and service-level objectives.

    • Experience crafting self-service infrastructure platforms and reusable automation that improves developer efficiency and theaAbility to establish clear, reliable, and scalable engineering and operational practices from evolving requirements.

    • Proven technical leadership, team leadership, or managerial experience, including mentoring engineers, prioritizing work, coordinating cross-functional initiatives, and driving projects from planning through completion.

    Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5.

    You will also be eligible for equity and benefits.

    Applications for this job will be accepted at least until September 21, 2026.

    This posting is for an existing vacancy.

    NVIDIA uses AI tools in its recruiting processes.

    NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

    Numbers & Facts

    LocationSanta Clara, CA
    IndustryComputer Software
    Company Size10,000 employees or more
    Year Founded1993
    Websitehttp://www.nvidia.com

    About Company

    Visualize your future . . . We Do
    NVIDIA is the world leader in graphics processing technologies, creating innovative, industry-changing products for computing, consumer electronics, and mobile devices. NVIDIA products are transforming visually-rich applications such as video games, film production, broadcasting, industrial design, space exploration, and medical imaging. We invest in our people and our technologies, support and fund industry research around the world, and consistently deliver high-quality products. NVIDIA's culture promotes and inspires a team of world-class employees to be at the top of their game. We've created an environment where talents are recognized and collaboration is valued. Our employees are shaping the world of tomorrow. . . today. We invite you to explore the opportunities available at NVIDIA to see what your future may hold.

    Skills

    • Analysis Skillsunmatched
    • Artificial Intelligence (AI)unmatched
    • Automationunmatched
    • Business Operationsunmatched
    • Capacity Managementunmatched
    • Computer Firmwareunmatched
    • Computer Networksunmatched
    • Computer Scienceunmatched
    • Configuration Managementunmatched
    • Continuous Improvementunmatched
    • Cross-Functionalunmatched
    • DHCP (Dynamic Host Configuration Protocol)unmatched
    • DNS (Domain Name System)unmatched
    • DevOpsunmatched
    • Distributed Computingunmatched
    • Documentationunmatched
    • Establish Prioritiesunmatched
    • Firewallsunmatched
    • Hardware Virtualizationunmatched
    • Identify Issuesunmatched
    • Incident Managementunmatched
    • Information Technology & Information Systemsunmatched
    • K Virtual Machine (KVM)unmatched
    • Linux Administrationunmatched
    • Linux Operating Systemunmatched
    • Mentoringunmatched
    • Metricsunmatched
    • Multi-User Systemsunmatched
    • NFS (Network File System)unmatched
    • Network Administration/Managementunmatched
    • Network Operations Centerunmatched
    • Network Routingunmatched
    • Network Softwareunmatched
    • Network Supportunmatched
    • Network Switchingunmatched
    • Operating Systemsunmatched
    • Operational Auditunmatched
    • Operational Improvementunmatched
    • Operations Guidelinesunmatched
    • Operations Processesunmatched
    • Performance Managementunmatched
    • Problem Solving Skillsunmatched
    • Project Planningunmatched
    • Reliability Engineeringunmatched
    • Software Administrationunmatched
    • Software Configuration Managementunmatched
    • Software Engineeringunmatched
    • Software Testingunmatched
    • Software Upgradesunmatched
    • Standards Developmentunmatched
    • Systems Engineeringunmatched
    • TCP/IP (Transmission Control Protocol/Internet Protocol)unmatched
    • Team Lead/Managerunmatched
    • Team Playerunmatched
    • Technical Leadershipunmatched
    • User Documentationunmatched
    • VLAN (Virtual Local Area Network)unmatched
    • Virtualizationunmatched

    Be found by employers

    5,500+ employers search our resume database daily. Add yours to get found by recruiters looking for candidates like you.

    Level up your application

    Professional resume templates

    Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.

    Free resume templates

    Free resume builder

    Improve your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.

    Free resume builder