Senior HPC Infrastructure SRE — Large-Scale Compute

Synopsys, Inc.

Hinoba-an

On-site

PHP 1,674,000 - 2,567,000

Full time

3 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Synopsys, Inc. is seeking an experienced HPC Infrastructure Engineer to design, build, and optimize large-scale compute farms that power flagship EDA workloads.

You will administer schedulers, lead capacity planning, and drive automation across global teams. You will collaborate with R&D, Cloud, Infrastructure, and Security to enable cloud-integrated HPC and AI/ML workloads while mentoring engineers and delivering high-availability platforms.

Qualifications

  • 8+ years of Linux/UNIX systems administration with performance tuning and large-scale operations.
  • 5+ years in HPC/compute farm administration including workload scheduling, resource management, and capacity planning.
  • Deep expertise with IBM Spectrum LSF, Slurm, or equivalent schedulers and policy configuration.
  • Strong knowledge of LDAP, NFS, DNS, enterprise storage, and networking in distributed compute environments.
  • Proven automation experience with Python and shell scripting for orchestration and monitoring.
  • Proficient with Grafana, Prometheus, Elastic, or Splunk for proactive incident detection and analysis.
  • Familiarity with cloud-integrated HPC on Azure/AWS, Kubernetes, Docker, Ansible, or Terraform.

Responsibilities

  • Design, build, and optimize large-scale HPC compute farm platforms supporting thousands of workloads across global sites.
  • Administer and tune IBM Spectrum LSF, Slurm, or equivalents to maximize resource utilization and minimize queue times.
  • Lead capacity planning and translate engineering demand into infrastructure requirements and timelines.
  • Drive automation using Python and Shell scripting to reduce toil, improve reliability, and accelerate incident response.
  • Lead complex troubleshooting and root cause analysis across storage, networking, LDAP, NFS, and scheduler layers.
  • Collaborate with R&D, Cloud, Infrastructure, and Security on cloud-integrated HPC and AI/ML workload enablement.
  • Mentor junior engineers and provide technical leadership across global teams, guiding operational standards.

Skills

Linux/UNIX
HPC/Compute
Python
Shell scripting
LDAP
NFS
DNS
Monitoring
Cloud platforms
Kubernetes
Docker
Ansible
Terraform
Azure
AWS
LSF
Slurm

Tools

IBM Spectrum LSF
Slurm
Azure
AWS
Kubernetes
Docker
Ansible
Terraform

Job description

Synopsys, Inc. is seeking an experienced HPC Infrastructure Engineer to design, build, and optimize large-scale compute farms that power flagship EDA workloads.

You will administer schedulers, lead capacity planning, and drive automation across global teams. You will collaborate with R&D, Cloud, Infrastructure, and Security to enable cloud-integrated HPC and AI/ML workloads while mentoring engineers and delivering high-availability platforms.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability, Staff / HPC Infrastructure Engineer
Site Reliability, Staff / HPC Infrastructure Engineer

Synopsys, Inc. • Hinoba-an

On-site
PHP 1,674,000 - 2,567,000
Staff Platform Engineer: Scalable AI Infra & Kubernetes
Staff Platform Engineer: Scalable AI Infra & Kubernetes

Ansys • España

On-site
PHP 7,528,000 - 10,038,000
Senior EDA Tools & Infrastructure Engineer
Senior EDA Tools & Infrastructure Engineer

Syntiant • Hinoba-an

On-site
PHP 7,528,000 - 11,292,000
Senior Staff SRE - Global Infra & Cloud Lead
Senior Staff SRE - Global Infra & Cloud Lead

NVIDIA Corporation • Hinoba-an

Hybrid
PHP 2,000,000 - 4,500,000
Hybrid work model
Senior SRE: Scale, Automate & Self-Healing Systems
Senior SRE: Scale, Automate & Self-Healing Systems

Replit • España

On-site
PHP 2,461,000 - 4,308,000
Competitive Salary & Equity
Health, Dental, Vision and Life Insurance
Flexible Time Off (FTO) + Holidays
+2
Senior SRE: Cloud Compute Reliability & Automation
Senior SRE: Cloud Compute Reliability & Automation

JobCubby • Hinoba-an

Hybrid
PHP 1,200,000 - 1,800,000
FlexBase
Hybrid work option
Senior Devops Engineer
Senior Devops Engineer

V2 Solutions • Hinoba-an

Hybrid
PHP 1,200,000 - 2,400,000
Platform SRE: Scale, Automate & Reliability (Hybrid)
Platform SRE: Scale, Automate & Reliability (Hybrid)

Broadridge Financial Solutions • Manila, Hinoba-an

Hybrid
PHP 1,200,000 - 1,800,000
Senior Datacenter Engineer - Remote, Night Shift Lead
Senior Datacenter Engineer - Remote, Night Shift Lead

SCALABLE OS CORP. • Metro Manila

On-site
PHP 1,200,000 - 2,400,000
Senior Staff Site Reliability Engineer
Senior Staff Site Reliability Engineer

NVIDIA Corporation • Hinoba-an

Hybrid
PHP 2,000,000 - 4,500,000
Hybrid work model