Systems Engineer (HPC/Server Farm)

Staffing Technologies

United States

On-site

USD 150,000 - 220,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Staffing Technologies is seeking a Systems Engineer to support and optimize a large-scale HPC/server farm. The role focuses on on-premises compute infrastructure in San Jose, with exposure to cloud (AWS) and collaboration with R&D teams on software optimization.

The candidate should bring 8+ years of Linux-based compute farm experience, hands-on LSF/RTM, scripting in Python/Shell/Perl, and a track record of cross-geography coordination. BS/MS in CS preferred.

Qualifications

  • 8+ years of experience in a compute farm environment running Linux.
  • 5+ years in a global/large compute farm or hybrid cloud with 1,000+ servers and remote reports.
  • 3+ years coordinating across geographies in a team setting.
  • Extensive scripting with Python, Shell, Perl; IBM LSF and RTM experience; Farm to Cloud familiarity.
  • Experience with R&D software teams to optimize workloads and develop KPI.
  • Proven process focus with documentation, change, incident, and problem-resolution activities.
  • BS or MS in computer science or related field.

Responsibilities

  • Support multiple locations across North America, Europe, and Asia.
  • Improve R&D productivity and drive customer success.
  • Lead the three-year compute roadmap and capacity growth for the San Jose on-premises server farm.
  • Operate, manage, and enhance the internal compute farm and cloud (AWS).
  • Maintain and improve efficiency with monitoring and reporting.

Skills

Linux
LSF
RTM
Python
Shell
Perl
Compute farms
Cloud

Education

BS/MS in computer science

Tools

IBM LSF

Job description

Systems Engineer (HPC/Server Farm)

Location: San Jose, CA

Posted On: 06/16/2026

Requirement Code: 73852

Requirement Detail

Seeking a Server Farm Engineer to join our team to support, manage, and improve the compute farm environment. The candidate should have hands-on experience with cloud solutions and proven expertise in working directly with R&D software development teams to develop solutions to optimize their working environment collaboratively.

30-year history of applying leading-edge optimization and analysis algorithms to highly complex problems in semiconductor and electronic design, verification, and analysis. We are looking for a recent graduate software engineer to join our team of collaborative EDA professionals to deliver the best-in-class next-generation software for physical IC applications. The software engineer will work on complex problems where data analysis requires an evaluation of intangible variance factors to develop leading-edge software for the physical design and verification of products at advanced nodes. Each day with offers exciting opportunities to create a better, more connected world. We are leading the charge to solve technology's toughest challenges. Working here means working alongside the industry's brightest people and innovating for the biggest, most innovative companies around the globe.

Responsibilities
  • Supporting multiple geological locations to serve user communities across North America, Europe, and Asia sites.
  • Focusing on improving R&D productivity and committing to customer success.
  • Driving the overall operational strategy for internal High-Performance Compute (HPC) farms in all locations.
  • Developing and executing the three-year compute roadmap and planning annual capacity growth for on-premises server farm in San Jose.
  • Operating, managing, and enhancing the internal compute farm and associated cloud (AWS).
  • Maintaining, enhancing, monitoring, reporting, and improving its efficiency.
Requirements
  • 8+ years of technical experience architecting, managing, and improving a compute farm environment running Linux.
  • At least 5 years of direct hands-on experience in a global or regional compute farm and/or hybrid cloud environment consisting of 1,000 or more servers with some remote direct reports
  • At least 3 years working in a global group, coordinating support, strategies, projects, and operations across multiple geographies in a team-oriented approach
  • Extensive technical experience managing IBM LSF and RTM and scripting using Python, shell, Perl, etc., in a Farm environment and knowledge of LSF spanning Farm to Cloud is highly desirable.
  • Solid understanding and proven operational experience with compute farms, job submission/management technologies, cloud, and associated management tools.
  • Proven experience working directly with R&D software development teams to collaboratively develop solutions to optimize their working environment (Direct EDA experience desired)
  • Proven experience in capacity and performance management, optimizing performance, ensuring adequate capacity, working with R&D on optimization of their workloads, and development and maintenance of key performance indicators
  • A proven process focus shown through documentation, change management, incident management and problem-resolution activities
  • Education: BS / MS in computer science or related field
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

HPC Systems Engineer — Cloud & On-Prem Compute Lead
HPC Systems Engineer — Cloud & On-Prem Compute Lead

Staffing Technologies • United States

On-site
USD 150,000 - 220,000
Data Center Compute Engineer
Data Center Compute Engineer

Blue Signal Search • United States

Hybrid
USD 120,000 - 180,000
Competitive compensation
Equity opportunity
Comprehensive benefits
Data Center Compute Engineer
Data Center Compute Engineer

Blue Signal Search • San Francisco (CA)

Hybrid
USD 150,000 - 210,000
Competitive compensation
Equity opportunity
Comprehensive benefits
+2
Business Systems Analyst
Business Systems Analyst

Insight Global • Sunnyvale (CA)

On-site
USD 103,320 - 151,536
Software Engineer Sr Staff
Software Engineer Sr Staff

Signature Federal Systems , LLC • Colorado Springs (CO)

On-site
USD 120,000 - 160,000
HPC System Administrator
HPC System Administrator

Cybotic System • Savannah (GA)

On-site
USD 90,000 - 150,000
Compute Platform Engineer
Compute Platform Engineer

NorthMark Compute & Cloud • Dallas (TX)

On-site
USD 120,000 - 180,000
HPC Software Engineer
HPC Software Engineer

Signature Federal Systems , LLC • Colorado Springs (CO)

On-site
USD 140,000 - 190,000
Server Infrastructure Engineer
Server Infrastructure Engineer

Info Services • Dearborn (MI)

On-site
USD 120,000 - 180,000
Compute Platform Engineer
Compute Platform Engineer

NMC2 • Dallas (TX)

On-site
USD 100,000 - 130,000