High Performance and Scientific Computing Fellow
Advanced Micro DevicesJob Title
High Performance and Scientific Computing Fellow
Role Summary
Senior individual-contributor technical leader responsible for defining strategy and driving the technical vision for AMD ROCm high-performance computing (HPC) GPU libraries across multiple GPU generations. Focuses on algorithmic development, GPU performance analysis and tuning, distributed systems, and external technical engagements.
Works across software layers (kernels, runtimes, libraries, distributed strategies) and represents AMD in benchmarks, customer interactions, and industry forums. Location: Austin, TX or remote; this role is not eligible for visa sponsorship.
Experience Level
Senior β Fellow-level individual contributor. No specific years-of-experience listed; the role expects extensive, established experience in HPC software, GPU optimization, and technical leadership.
Responsibilities
Primary responsibilities include technical leadership, performance optimization, and external representation.
- Define technical vision and drive software strategy for ROCm HPC libraries.
- Lead performance analysis, tuning, and algorithmic innovation across HPC libraries.
- Provide technical mentorship to senior engineers and influence engineering best practices.
- Communicate complex technical findings and recommendations to senior leadership and stakeholders.
- Represent AMD in external technical forums, industry benchmarks, and customer engagements.
- Collaborate with internal experts, partners, and customers to improve ROCm HPC and AI applications, libraries, and tools.
Requirements
Key technical skills and abilities expected for successful candidates. Educational credentials are summarized separately below.
- Must-have: Deep expertise in mathematical algorithms for dense and sparse linear algebra and in floating-point numerical behavior.
- Must-have: Proven experience designing, implementing, debugging, and optimizing parallel algorithms on large-scale supercomputers and distributed systems.
- Must-have: Strong background in developing libraries and applications in C++, C, and Fortran; experience with GPU programming and optimization (HIP, CUDA, or OpenCL) and distributed programming (MPI and/or SHMEM).
- Must-have: Strong GPU performance analysis and low-level optimization skills, including assembly programming and vectorization techniques.
- Must-have: Demonstrated ability to drive impactful optimizations, define software architecture, and influence technical direction across teams.
- Nice-to-have: Experience with industry-standard HPC benchmarks and public technical engagements; familiarity with testing, profiling, debugging, documentation, version control, and issue tracking best practices.
- Nice-to-have: Experience applying agentic AI workflows to engineering problems.
Education Requirements
B.Sc. or B.Eng. in Computer Science, Software Engineering, Electrical Engineering, Applied Mathematics, or equivalent. Advanced degrees (M.Sc., M.Eng., Ph.D.) are preferred. Equivalent practical experience is acceptable.
About the Company
Company: Advanced Micro Devices
Headquarters: Sunnyvale, California, USA
Advanced Micro Devices, or AMD, is a global semiconductor company that designs and manufactures microprocessors, graphics processors, and related technologies for a variety of computing devices. Known for pushing the boundaries of innovation, AMD's mission is to deliver high-performance computing solutions for AI, data centers, gaming, and embedded applications. They foster a collaborative, inclusive culture focused on creativity and problem-solving, aiming to drive progress and excellence in technology.
