Senior Software Engineer - SOC Platforms
NVIDIAJob Title
Senior Software Engineer - SOC Platforms
Role Summary
The Senior Software Engineer owns platform and system software for NVIDIA GPU and SOC products, focusing on kernel/device drivers, firmware interactions, observability, performance, and system-level features.
The role partners with software, driver, product, and cloud teams to launch GPU platforms, resolve system-level issues, and improve performance and reliability across configurations.
Experience Level
Senior - requires substantial experience; the posting specifies at least 10 years developing SOC systems and platform software.
Responsibilities
Key responsibilities include ownership of platform software, design and delivery of observability and production systems, and collaboration across hardware and software teams.
- Develop and maintain kernel device drivers and system-level platform software for GPU/SOC products.
- Design and implement high-performance, distributed observability platforms for metrics, logs, and traces at scale.
- Build production backend pipelines, analytics, monitoring, logging, and alerting systems focused on performance and scale.
- Implement full-stack production applications and features for SOC platforms.
- Partner with cross-functional teams (software, drivers, product, data center/cloud) to ship platforms and troubleshoot complex system issues.
- Participate in code reviews and contribute substantial production code in C/C++ and related stacks.
- Analyze how AI workloads map to hardware (SMs, Tensor Cores, memory hierarchy, NVLink, PCIe, MIG) and optimize for latency, throughput, and cost.
- Profile and tune OS and system software focusing on Linux internals, concurrency, and performance.
- Manage CPU–GPU interactions, PCIe, interrupts, firmware/bootloader integration, and system reliability features.
Requirements
Core qualifications and skills expected for the role.
- Must-have: At least 10 years of experience in SOC systems and platform software development; strong hands-on experience with Linux internals, device drivers, kernel/user boundaries, concurrency, and performance profiling.
- Must-have: Strong proficiency in C/C++ and experience contributing to large production codebases.
- Must-have: Experience designing and building observability architectures and distributed systems (metrics, logs, traces, analytics stacks).
- Must-have: Backend systems programming and production-grade, secure software development experience.
- Must-have: Solid understanding of operating systems, computer architecture, distributed systems, and databases.
- Must-have: Experience with AI/ML applications and how workloads interact with hardware.
- Nice-to-have: Experience across x86 and ARM architectures and on-device programming/debugging for SOC platforms.
- Nice-to-have: Deep knowledge of networking, virtualization, and familiarity with Windows OS internals in addition to Linux.
Education Requirements
BS, MS, or PhD in Computer Science or a related field, or equivalent practical experience. The posting explicitly allows equivalent experience in lieu of a degree.
About the Company
Company: NVIDIA
Headquarters: Santa Clara, California, USA
NVIDIA is a global leader in accelerated computing, renowned for its innovative solutions in AI and digital twins that transform diverse industries. The company specializes in networking technologies, providing end-to-end InfiniBand and Ethernet solutions for servers and storage that optimize performance and scalability. NVIDIA serves sectors such as high-performance computing, enterprise data centers, and cloud computing, constantly reinventing its products and services to stay ahead in the market.
