Job Title
Host and Systems Performance Senior Manager
Role Summary
Lead a team responsible for measuring, analyzing, and improving NIC, DPU, host, DOCA networking stack, and full-system performance across NVIDIA networking products and platforms. The role covers the full product lifecycle from pre-silicon modeling through post-silicon bring-up, validation, characterization, and GA readiness.
The position requires cross-functional collaboration with hardware, firmware, software, architecture, validation, and product teams to define performance goals and resolve bottlenecks for AI, HPC, cloud, and accelerated networking workloads.
Experience Level
Senior β typically 10+ years of software engineering experience, including 4+ years in a leadership or management role (as stated in the posting).
Responsibilities
Accountable for team leadership, performance engineering strategy, and execution across host and system-level networking performance.
- Lead, coach, and develop engineers focused on NIC, host, DOCA networking stack, and system-level performance.
- Define performance test plans, methodologies, metrics, and success criteria for new networking technologies.
- Guide pre-silicon and post-silicon performance planning, analysis, reporting, and readiness reviews.
- Benchmark and profile workloads across RDMA, RoCE, InfiniBand, Ethernet, MPI, NCCL, DOCA, storage, security, and host networking stacks.
- Identify system bottlenecks (NIC/DPU, CPU, memory, PCIe, firmware, drivers, OS) and lead root-cause analysis and mitigation coordination across teams.
Requirements
Must-have technical skills and experience required to perform the role; followed by desirable skills that strengthen a candidate's fit.
- 10+ years software engineering experience with 4+ years in leadership or management.
- Experience in performance analysis, systems engineering, networking, or HPC/AI infrastructure.
- Experience with high-performance networking technologies such as RDMA, RoCE, InfiniBand, Ethernet, MPI, or NCCL.
- Hands-on experience with system performance analysis, benchmarking, profiling, and root-cause analysis.
- Understanding of host architecture: CPUs, memory hierarchy, NUMA, PCIe, Linux OS, drivers, firmware, and DPU/NIC interactions.
- Experience creating performance test plans for pre-silicon and post-silicon phases.
- Programming/scripting experience with Python and Bash; C/C++ experience is helpful.
- Ability to communicate technical findings clearly and collaborate across engineering teams.
Nice-to-have:
- Experience with NIC or DPU architecture, networking offloads, host datapath behavior, networking services, or DPU offloads.
- Experience with AI/HPC cluster performance, distributed training/inference workloads, collective communication libraries, telemetry pipelines, or benchmark automation.
- Familiarity with CUDA, NCCL internals, Linux kernel networking, DPDK, OVS, storage acceleration, or security offloads.
Education Requirements
B.Sc. or M.Sc. in Computer Science, Computer Engineering, Electrical Engineering, or equivalent practical experience.
About the Company
Company: NVIDIA
Headquarters: Santa Clara, California, USA
NVIDIA is a global leader in accelerated computing, renowned for its innovative solutions in AI and digital twins that transform diverse industries. The company specializes in networking technologies, providing end-to-end InfiniBand and Ethernet solutions for servers and storage that optimize performance and scalability. NVIDIA serves sectors such as high-performance computing, enterprise data centers, and cloud computing, constantly reinventing its products and services to stay ahead in the market.

Date Posted: 2026-08-17