Job Title
Manager, Distinguished Engineer - DGX Systems Software
Role Summary
Lead engineering for end-to-end readiness and delivery of NVIDIA DGX compute systems, spanning firmware through the AI software stack to customer deployment. Own platform architecture, release readiness, validation, vendor management, and a multi-site engineering organization responsible for production-quality systems.
Experience Level
Senior β typically 12+ years in systems firmware/software engineering with 5+ years in engineering leadership.
Responsibilities
Accountable for DGX platform readiness, delivery lifecycle, cross-team alignment, and people leadership.
- Ensure full-stack readiness (firmware, OS, drivers, CUDA, networking, management tools) and manage GA software/firmware release process.
- Lead development and integration of manageability firmware (BMC, SBIOS/BIOS) and coordinate third-party firmware contributors.
- Define and execute validation strategy: firmware regression, NVQual, DL workload performance, OS/CUDA stack testing, multi-user and reliability validations.
- Drive platform bring-up and architecture for new DGX systems, including firmware update mechanisms and system security posture.
- Represent platform readiness in executive reviews and enable customer/CSP deployment requirements.
- Own product delivery lifecycle from architecture through GA release and field deployment; manage RCCA for field issues.
- Align GPU, CPU, networking, security, OS, and AI software teams to meet schedules and quality gates.
- Build, mentor, and lead a global engineering organization focused on technical excellence and delivery.
Requirements
Must-have technical background and leadership experience; prefer platform and standards familiarity.
- 12+ years in systems firmware/software engineering; 5+ years in engineering leadership managing distributed teams.
- Deep expertise in server system stack and system-level integration (SBIOS, BMC, OS, drivers, applications).
- Proven record delivering multi-generation server or data center platforms from architecture through customer deployment.
- Strong understanding of server hardware: CPU, GPU, interconnect, memory, PCIe, and power delivery.
- Experience owning end-to-end product quality, firmware validation, full-stack system testing, and field deployment reliability.
- Experience managing vendors and setting clear quality gates and delivery tracking.
- Nice to have: experience with DGX or GPU-accelerated server platforms, DMTF Redfish/OCP standards, and AI/DL workload validation and performance optimization.
- Nice to have: demonstrated ability influencing at VP/SVP level and driving cross-BU strategic decisions.
Education Requirements
BS or MS in Computer Science, Electrical Engineering, or a related field β or equivalent practical experience.
About the Company
Company: NVIDIA
Headquarters: Santa Clara, California, USA
NVIDIA is a global leader in accelerated computing, renowned for its innovative solutions in AI and digital twins that transform diverse industries. The company specializes in networking technologies, providing end-to-end InfiniBand and Ethernet solutions for servers and storage that optimize performance and scalability. NVIDIA serves sectors such as high-performance computing, enterprise data centers, and cloud computing, constantly reinventing its products and services to stay ahead in the market.

Date Posted: 2026-08-14