Deep Learning Performance Architect
Work on performance modelling, analysis, and optimization for deep learning (including LLM) workloads on current and next-generation NVIDIA hardware. The role partners with architecture, software, and product teams to influence processor and system design for inference performance and efficiency.
Mid-level β typically 5+ years of relevant industry experience.
Analyze models and define measurable opportunities to improve DL inference performance across software and hardware.
Must-have technical experience and skills.
BS, MS, or PhD in Computer Science, Electrical Engineering, Mathematics, or a related technical field, or equivalent practical experience.
Company: NVIDIA
Headquarters: Santa Clara, California, USA
NVIDIA is a global leader in accelerated computing, renowned for its innovative solutions in AI and digital twins that transform diverse industries. The company specializes in networking technologies, providing end-to-end InfiniBand and Ethernet solutions for servers and storage that optimize performance and scalability. NVIDIA serves sectors such as high-performance computing, enterprise data centers, and cloud computing, constantly reinventing its products and services to stay ahead in the market.
