Job Title
Sr Staff Engineer, AI/ML Compiler & Runtime Software Engineer
Role Summary
Senior software engineer responsible for compiler and runtime software enabling efficient AI execution on RISC-V, NPU, and SoC platforms. The role spans MLIR/LLVM-based compiler flows, IREE runtime integration, quantization and code generation, and hardware-aware optimization for edge AI deployment.
Experience Level
Senior. The posting specifies 3β12 years of hands-on software engineering experience; title indicates a senior-level position.
Responsibilities
Design, implement, and validate compiler and runtime components to enable AI workloads on edge silicon and accelerators.
- Architect and develop AI/ML compiler and runtime software targeting RISC-V, NPU, and SoC platforms.
- Develop and extend IREE-based compiler flows including MLIR lowering, code generation, runtime integration, and deployment paths for edge workloads.
- Create and maintain custom MLIR dialects, compiler passes, lowering pipelines, and transformation flows for accelerator mapping.
- Support framework import paths (PyTorch, ONNX, TFLite) and lowering via torch-mlir, TOSA, Linalg, and related MLIR dialects.
- Optimize neural network workloads for edge deployment: operator fusion, tiling, memory planning, quantization, layout transforms, and accelerator-aware scheduling.
- Enable execution across CPU, vector/matrix units, and NPUs, balancing latency, throughput, memory, and power.
- Collaborate with architecture, hardware, firmware, FPGA, validation, and product teams for simulator/FPGA/silicon bring-up.
- Analyze model performance, identify compiler/runtime bottlenecks, and implement optimizations at graph, operator, and kernel levels.
- Define software architecture and technical direction for AI SDK components, build test infrastructure, benchmarks, and CI for correctness and performance.
- Provide technical leadership, mentoring, and cross-team coordination for compiler/runtime and platform enablement.
Requirements
Core qualifications and technologies required for the role.
- Proven experience with IREE, LLVM, and MLIR; developing MLIR dialects, passes, lowering pipelines, and backend integration for custom hardware.
- Practical knowledge of IREE code generation flow, dispatch formation, executable generation, HAL/runtime concepts, and target-specific lowering.
- Experience with AI model formats and frameworks: PyTorch, ONNX, TensorFlow Lite/TFLite and related import/conversion flows.
- Strong understanding of neural network execution and optimization: quantization, operator fusion, tensor layouts, memory planning, tiling, vectorization, and kernel selection.
- Experience enabling or optimizing workloads for accelerators (NPUs, DSPs, vector/matrix engines, or custom SoC IP).
- Strong C/C++ skills and Python for tooling, testing, automation, and model workflow integration.
- Linux development experience including cross-compilation, debugging, profiling, build systems, and runtime bring-up.
- Proven ability to lead technical projects and mentor engineers across cross-functional teams.
Nice-to-have
- Experience with RISC-V, RISC-V Vector or custom instruction code generation, ARM, x86, DSP, or GPU accelerator stacks.
- Familiarity with FPGA prototyping, board bring-up, emulation, or early silicon validation.
- Exposure to LLM/edge inference stacks (llama.cpp, GGML/GGUF, ONNX Runtime, TVM, XNNPACK) and related quantization/benchmarking techniques.
- Experience with CI/CD and engineering infrastructure tools like Jenkins, CMake, Bazel, Git, or Jira.
Education Requirements
Not specified.
About the Company
Company: GlobalFoundries
Headquarters: Saratoga Springs, New York, USA
GlobalFoundries is a leading contract manufacturer for the global semiconductor industry, with facilities in multiple countries, including the USA. The company develops a broad portfolio of semiconductor technologies and employs around 13,000 people worldwide. GlobalFoundries focuses on enhancing competitiveness in specialized application solutions and fostering innovation in mobile communications, consumer electronics, and automotive applications.

Date Posted: 2026-08-21