Job Title
EDA Computing / HPC Infrastructure MTS
Role Summary
Senior hands-on technical role providing ownership and operational support for EDA/HPC compute, storage, scheduling, and application services used for semiconductor design, verification, and tapeout. The role supports hybrid on-premises and cloud (AWS) environments and focuses on performance, reliability, security, and cost-effective operations.
Experience Level
Senior. Minimum 8+ years relevant experience for MTS level; 10+ years for Senior MTS (experience in EDA computing, HPC, Linux infrastructure, storage, cloud, or production engineering support).
Responsibilities
Primary technical and operational responsibilities include:
- Operate, optimize, and troubleshoot EDA compute clusters and batch scheduling (preferably IBM LSF) for simulation, physical verification, OPC, lithography, signoff, and tapeout workloads.
- Analyze job failures, hung jobs, queue bottlenecks, license constraints, compute node issues, and workload performance problems.
- Provide technical ownership for hybrid EDA HPC platforms across on-premises infrastructure and AWS-based compute environments.
- Support and optimize EDA storage services (NFS, NetApp ONTAP, AWS FSx, S3, and related high‑performance file services).
- Design, develop, and maintain automation tools to improve operations, observability, reliability, and support efficiency.
- Own and resolve complex ServiceNow incidents, requests, problems, and changes related to EDA computing services.
- Lead root-cause analysis, coordinate planned system changes, and provide concise technical status updates during critical incidents.
- Participate in team documentation, knowledge sharing, cross-training, and continuous improvement; follow EHS and security requirements.
Requirements
Must-have technical skills and experience; preferred items listed separately.
- Extensive experience supporting EDA computing, HPC, or large-scale Linux production environments.
- Hands-on experience with batch schedulers, preferably IBM LSF; ability to tune and troubleshoot scheduler behavior.
- Strong Linux system administration skills on RHEL-based platforms.
- Bash scripting and automation experience; ability to develop tools that improve operational efficiency.
- Ability to analyze complex incidents, perform root-cause analysis, and implement sustainable fixes.
- Experience with high-performance storage and file services, troubleshooting performance and capacity issues.
- Familiarity with ServiceNow or similar ITSM tooling for incident/request/change management.
- Fluency in English (written and verbal); travel up to 10%.
Nice-to-have / preferred:
- AWS or other cloud platform experience for HPC workloads; hybrid cloud operations.
- Python automation, monitoring/logging/dashboarding, capacity reporting, and operational analytics tools.
- Experience supporting semiconductor design and tapeout flows (Calibre, ASML Brion, Cadence, Synopsys, or similar EDA applications).
- NetApp ONTAP, NFS, FSx, S3 experience.
Education Requirements
Required: Bachelor's degree in Computer Science, Engineering, Information Technology, or equivalent practical experience. Preferred: Master’s degree in Computer Science, Engineering, Information Technology, or related technical field. Equivalent practical experience is explicitly accepted.
Expected Salary Range: $104,000 - $175,000 (exact salary determined by qualifications, experience, and location).
About the Company
Company: GlobalFoundries
Headquarters: Saratoga Springs, New York, USA
GlobalFoundries is a leading contract manufacturer for the global semiconductor industry, with facilities in multiple countries, including the USA. The company develops a broad portfolio of semiconductor technologies and employs around 13,000 people worldwide. GlobalFoundries focuses on enhancing competitiveness in specialized application solutions and fostering innovation in mobile communications, consumer electronics, and automotive applications.

Date Posted: 2026-07-27