Job Title
Host and Systems Performance Senior Manager
Role Summary
Lead a team that measures, analyzes, and improves NIC, host, DOCA networking stack, and system-level performance across NVIDIA networking products and platforms. Work spans pre-silicon modeling, post-silicon bring-up, validation, characterization, and GA readiness.
This role requires cross-functional collaboration with hardware, firmware, software, architecture, validation, and product teams to set performance goals and drive product readiness for AI, HPC, cloud, and accelerated networking workloads.
Experience Level
Senior β requires at least 10 years of software engineering experience, including 4+ years in a leadership or management role.
Responsibilities
You will manage a performance engineering team and lead performance activities across the product lifecycle.
- Lead, coach, and develop engineers focused on NIC, host, DOCA stack, and system-level performance.
- Define performance test plans, methodologies, metrics, and success criteria for new networking technologies.
- Plan and guide pre-silicon and post-silicon performance analysis, benchmarking, profiling, and readiness reviews.
- Benchmark and profile workloads across RDMA, RoCE, InfiniBand, Ethernet, MPI, NCCL, storage, security, and host networking stacks.
- Identify bottlenecks across NIC/DPU, CPU, memory, PCIe, firmware, drivers, Linux networking, and system architecture; lead root-cause analysis and mitigation planning.
- Coordinate cross-team alignment on performance goals and report technical findings to stakeholders.
Requirements
Must-have technical skills and leadership experience. Education requirements are summarized separately below.
- Proven experience in performance analysis, systems engineering, networking, or HPC/AI infrastructure.
- 10+ years of software engineering experience with 4+ years in a management or leadership role.
- Hands-on experience with system performance analysis, benchmarking, profiling, and root-cause analysis.
- Familiarity with high-performance networking technologies: RDMA, RoCE, InfiniBand, Ethernet, MPI, NCCL.
- Strong understanding of host architecture: CPUs, memory hierarchy, NUMA, PCIe, Linux OS, drivers, firmware, and DPU/NIC interactions.
- Experience creating performance test plans for pre-silicon and post-silicon phases.
- Programming/scripting proficiency in Python and Bash; C/C++ experience is beneficial.
- Clear technical communication and ability to work across engineering teams.
Nice-to-have:
- Experience with NIC or DPU architecture, networking offloads, or host datapath behavior.
- Familiarity with AI/HPC cluster performance, distributed training/inference, collective communication libraries, CUDA, or NCCL internals.
- Experience with Linux kernel networking, DPDK, OVS, storage acceleration, or security offloads.
Education Requirements
B.Sc. or M.Sc. in Computer Science, Computer Engineering, Electrical Engineering, or equivalent practical experience; equivalent experience is explicitly accepted.
About the Company
Company: NVIDIA
Headquarters: Santa Clara, California, USA
NVIDIA is a global leader in accelerated computing, renowned for its innovative solutions in AI and digital twins that transform diverse industries. The company specializes in networking technologies, providing end-to-end InfiniBand and Ethernet solutions for servers and storage that optimize performance and scalability. NVIDIA serves sectors such as high-performance computing, enterprise data centers, and cloud computing, constantly reinventing its products and services to stay ahead in the market.

Date Posted: 2026-07-17