Back to jobs

    About 2100 NVIDIA USA

    NVIDIA is at the forefront of groundbreaking developments in Artificial Intelligence, High-Performance Computing, and Visualization. The GPU, an NVIDIA invention, serves as the visual cortex of modern computers and is fundamental to our products and services. We pride ourselves on having some of the most forward-thinking and talented individuals globally. If you are creative, passionate, and self-motivated, we encourage you to join us.

    NVIDIA's invention of the GPU in 1999 not only fueled the growth of the PC gaming market and redefined modern computer graphics but also revolutionized parallel computing. More recently, GPU deep learning ignited the modern deep learning era, with the GPU acting as the brain for computers, robots, and self-driving cars that perceive and understand the world. Today, NVIDIA is widely recognized as "the AI computing company." We are committed to expanding our company and building teams with the most innovative minds. Are you ready to shape the next generation of computing? Join us at the forefront of technological advancement.

    About the Position

    Introduction

    NVIDIA's data center systems, including DGX and HGX, are central to the company's rapidly expanding enterprise and cloud provider businesses. These platforms integrate the full power of NVIDIA GPUs, NVLink, InfiniBand networking, Grace CPUs, and a fully optimized NVIDIA AI and HPC software stack. We are seeking a highly skilled technical architect to assume end-to-end ownership of the system software architecture for these products, encompassing firmware, kernel drivers, operating systems, and user mode drivers. In this role, you will collaborate with internal component leads and engage with leading cloud service providers to bring these innovative products to market.

    Responsibilities

    • Serve as the primary technical point of contact for major customers, leading technical discussions, defining KPIs, gathering requirements, and resolving complex technical challenges.
    • As a system software architect, drive technical innovation and foster strategic collaborations with major hyperscalers to design next-generation data center products.
    • Align NVIDIA's product roadmap with the strategic requirements of key customers through direct engagement.
    • Develop and champion the adoption of new technologies and industry-standard protocols.
    • Make critical technical decisions in ambiguous situations, proactively mitigating risks through left-shift strategies.

    Requirements

    • Deep expertise in scalable and performant server system architecture, with a significant focus on software/hardware interfaces.
    • Extensive experience with complex system software for accelerators (GPUs, DPUs, FPGAs).
    • Mastery of system firmware (SBIOS, OpenBMC), embedded systems, and Linux kernel internals.
    • Proficiency in Out-of-Band and In-Band management architectures, device management protocols (e.g., MCTP, PLDM, SPDM, RDE), and system management protocols (Redfish, IPMI).
    • Extensive knowledge of networking technologies and protocols, including TCP/IP, Ethernet, InfiniBand, as well as advanced switching and routing concepts.
    • Demonstrated experience collaborating with platform security experts to balance security requirements with ease of use.
    • Proven success in leading complex, cross-functional projects to completion, showcasing the ability to influence and achieve outcomes without direct authority in large-scale, collaborative environments.
    • Demonstrable experience in implementing left-shift strategies to de-risk program execution.
    • BS or MS degree in Computer Science, Electrical Engineering, or a related field (or equivalent practical experience).
    • 20+ years of experience in system architecture and design.

    Nice to Have

    • Knowledge of cloud and cluster-level deployment and management systems.
    • Participation and contributions in standards bodies such as OCP and DMTF.
    • Familiarity with NVIDIA HPC programming models and libraries (CUDA, cuDNN, DOCA).
    • Knowledge of enterprise storage architectures and distributed parallel processing paradigms.

    Benefits

    • Competitive base salary range of $320,000 USD - $488,750 USD.
    • Eligible for equity compensation.
    • Comprehensive benefits package.

    About Company

    NVIDIA is at the forefront of groundbreaking developments in Artificial Intelligence, High-Performance Computing, and Visualization. The GPU, an NVIDIA invention, serves as the visual cortex of modern computers and is fundamental to our products and services. We pride ourselves on having some of the most forward-thinking and talented individuals globally. If you are creative, passionate, and self-motivated, we encourage you to join us.

    NVIDIA's invention of the GPU in 1999 not only fueled the growth of the PC gaming market and redefined modern computer graphics but also revolutionized parallel computing. More recently, GPU deep learning ignited the modern deep learning era, with the GPU acting as the brain for computers, robots, and self-driving cars that perceive and understand the world. Today, NVIDIA is widely recognized as "the AI computing company." We are committed to expanding our company and building teams with the most innovative minds. Are you ready to shape the next generation of computing? Join us at the forefront of technological advancement.

    How to Apply

    Applications for this position will be accepted until at least May 17, 2026. Please apply through the NVIDIA careers portal, ensuring you highlight your relevant experience and qualifications.

    Apply Now

    Your data is only shared with 2100 NVIDIA USA

    Location

    Santa Clara, CA, United States

    Type

    Full-Time

    Keywords

    AI
    HPC
    GPU
    NVLink
    InfiniBand
    Grace CPUs
    Firmware
    Kernel Drivers
    Operating Systems
    User Mode Drivers
    System Software
    Accelerators
    DPUs
    FPGAs
    SBIOS
    OpenBMC
    Embedded Systems
    Linux Kernel
    Out-of-Band Management
    In-Band Management
    MCTP
    PLDM
    SPDM
    RDE
    Redfish
    IPMI
    Networking
    TCP/IP
    Ethernet
    System Security
    Cloud Deployment
    Cluster Management
    OCP
    DMTF
    CUDA
    cuDNN
    DOCA
    Enterprise Storage
    Distributed Parallel Processing
    Verified Company

    Distinguished Engineer, Data Center System Software Architect

    2100 NVIDIA USA