Hardware Systems Engineer, NPI AI
Meta PlatformsJob Title
Hardware Systems Engineer, NPI AI
Role Summary
Support new product introduction (NPI) for next-generation AI and high-performance computing infrastructure used in large-scale data centers. Work across hardware design, firmware, software, networking, and capacity teams to validate and scale AI server systems from early bring-up through production readiness.
US salary range: $118,000β$170,000/year (plus bonus, equity, benefits).
Experience Level
Mid-level - typically requires 2+ years of relevant experience in hardware systems engineering, silicon or system bring-up, or firmware validation.
Responsibilities
Primary duties include system validation, bring-up, debugging, and cross-functional coordination during NPI.
- Drive system validation strategies for AI and HPC hardware platforms in data center environments.
- Perform hands-on bring-up, characterization, and validation of AI server systems and components (PCIe, NVLink, DRAM, high-speed networking fabrics).
- Develop and maintain test specifications, validation procedures, and debug guides for NPI programs.
- Investigate and root-cause complex failures spanning silicon, firmware, software, and hardware layers.
- Triage and track hardware and firmware defects to resolution while keeping NPI milestones on track.
- Identify gaps in test coverage and improve test methodologies, tooling, and automation frameworks.
- Partner with platform and capacity engineering to define acceptance criteria and deployment readiness standards.
- Collect, analyze, and report hardware quality trends; communicate validation status and technical findings to engineering teams and vendors.
- Collaborate with firmware and software teams on hardware-software interface requirements for telemetry, diagnostics, and remote management.
Requirements
Must-have technical experience and skills; preferred items noted as nice-to-have.
- 2+ years experience in hardware systems engineering, silicon validation, firmware validation, or system-level bring-up for AI servers, GPUs, TPUs, or AI accelerator platforms.
- Experience in one or more domains: ASIC bring-up and characterization, board-level debug, firmware validation, or large-scale system validation in data center environments.
- Experience developing test specifications, validation procedures, and debug methodologies for complex hardware systems.
- Proven root-cause analysis and troubleshooting of system-level failures across hardware, firmware, and software stacks.
- Experience with high-speed interconnects or memory subsystems such as PCIe, NVLink, DDR5, or HBM in AI or HPC validation contexts.
- Nice-to-have: integration of lab instrumentation and automation frameworks; proficiency in Linux and server system management; familiarity with SoC debug tools (JTAG, GDB, Trace32) and bus protocols (I2C, SPI, USB); experience defining telemetry and out-of-band management interfaces.
Education Requirements
Bachelor's degree in Computer Science, Computer Engineering, or a relevant technical field is required, or equivalent practical experience. Degree must be completed prior to joining Meta.
About the Company
Company: Meta Platforms
Headquarters: Menlo Park, California, United States
American technology company that develops social networking products (Facebook, Instagram, WhatsApp) and invests in virtual/augmented reality hardware and software through Reality Labs, focusing on connectivity, advertising, and immersive computing experiences.
