CAE HPC System Administrator

Toyota Tsusho Systems Corporation

Saline (MI)

Hybrid

USD 110,000 - 170,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Toyota Tsusho Systems US, Inc. is seeking a CAE HPC System Administrator with 7+ years of Linux experience to join our digital solution manufacturing team.

You will administer and optimize HPC clusters, manage job scheduling, and support CAE applications and licensing. The role emphasizes automation, performance monitoring, and collaborative support across CAE tools such as Ansys or LS-DYNA, with familiarity in InfiniBand networking and Rack/Ethernet hardware.

Qualifications

  • 7+ years of Linux system administration experience (preferably RHEL environments).
  • Bachelor’s degree in mechanical engineering, electrical engineering, computer engineering, computer science, or related field; and/or commensurate work experience
  • Hands-on experience managing HPC clusters and job schedulers (LSF, Slurm, PBS, or similar).
  • Proven experience in CAE application support and integration.
  • Strong scripting skills (Bash, Shell, Perl, or equivalent).
  • Experience with OS deployment, patching, and system automation.
  • Solid understanding of enterprise server hardware, storage, and networking fundamentals.
  • Experience with CAE tools such as Ansys, LS-DYNA, Nastran, or similar.
  • Familiarity with high-performance networking technologies is plus (e.g., InfiniBand).
  • Experience developing internal tools or dashboards are plus (e.g., PHP or web-based tooling).

Responsibilities

  • Administer, configure, and optimize HPC job scheduling environments, including IBM Spectrum LSF, Open PBS, or equivalent schedulers.
  • Design and tune job queues, resource allocation policies, and scheduling strategies to support diverse CAE workloads.
  • Monitor system performance and utilization trends and implement improvements to maximize efficiency and throughput.
  • Install, upgrade, test, and support CAE applications and simulation tools in production environments.
  • Provide integration support between CAE applications and HPC scheduling systems.
  • Manage CAE software licensing systems (e.g., FlexLM, RLM) and ensure availability.
  • Troubleshoot application-related issues and ensure minimal disruption to engineering activities.
  • Administer and maintain Red Hat Enterprise Linux (RHEL) environments across HPC clusters.
  • Perform OS deployment, deployment, and patch management using automated tools (e.g., PXE, or configuration management solutions).
  • Develop and maintain scripts (Bash, Korn shell, C Shell, Perl, Awk, or equivalent) to automate system monitoring, health checks, and routine administrative tasks.
  • Maintain system logs, monitoring processes, and standard operating procedures.
  • Troubleshoot and resolve issues related to servers, storage systems, and high-performance networking (e.g., InfiniBand, high-speed Ethernet).
  • Support hardware lifecycle activities including installation, maintenance, and upgrades.
  • Conduct capacity planning based on system utilization trends and future demand.
  • Perform system health checks, monitoring, and incident tracking for HPC and CAE environments.
  • Document system configurations, procedures, incidents, and best practices.
  • Track outages, analyze root causes, and implement preventive measures.
  • Follow change management processes for system updates and deployments.
  • Provide accurate reporting (e.g., utilization, incidents, system performance) and support project initiatives.

Skills

Linux system administration
HPC
Job scheduling
CAE applications
Scripting
OS deployment
Networking fundamentals

Education

Bachelor’s degree in engineering or CS

Tools

LSF
Slurm
PBS
OpenPBS
RHEL

Job description

About TTS-US:

Founded in 2011, Toyota Tsusho Systems US, Inc. (TTS-US) is a Toyota group company, that develops IT solutions wherever global businesses operate. Transforming into a technology and mobility company, TTS-US, with its 8 TTS affiliates worldwide is establishing a secure and resilient Toyota global value chain. The creative capacity to forge such limitless business opportunities is one of the strengths of Toyota Tsusho Systems.


Position Summary

We are seeking a highly motivated and experienced CAE HPC System Administrator with more than 7 years of experience to join our dynamic digital solution manufacturing team. This position is ideal for a candidate with strong Linux system administration experience and hands-on expertise managing HPC environments for CAE workloads, including job schedulers, system automation, and engineering application support. The successful candidate will be responsible for administering and optimizing HPC clusters, managing job scheduling systems, supporting CAE applications and licensing, automating Linux operations, maintaining infrastructure performance, and ensuring system stability, scalability, and efficient workload execution.


Essential Functions

HPC Job Queuing & Workload Management


  • Administer, configure, and optimize HPC job scheduling environments, including IBM Spectrum LSF, Open PBS, or equivalent schedulers.

  • Design and tune job queues, resource allocation policies, and scheduling strategies to support diverse CAE workloads.

  • Monitor system performance and utilization trends and implement improvements to maximize efficiency and throughput.


CAE Application and Licensing Support


  • Install, upgrade, test, and support CAE applications and simulation tools in production environments.

  • Provide integration support between CAE applications and HPC scheduling systems.

  • Manage CAE software licensing systems (e.g., FlexLM, RLM) and ensure availability.

  • Troubleshoot application-related issues and ensure minimal disruption to engineering activities.


Linux Systems Administration & Automation


  • Administer and maintain Red Hat Enterprise Linux (RHEL) environments across HPC clusters.

  • Perform OS provisioning, deployment, and patch management using automated tools (e.g., PXE, or configuration management solutions).

  • Develop and maintain scripts (Bash, Korn shell, C Shell, Perl, Awk, or equivalent) to automate system monitoring, health checks, and routine administrative tasks.

  • Maintain system logs, monitoring processes, and standard operating procedures.


Hardware & Infrastructure Management


  • Troubleshoot and resolve issues related to servers, storage systems, and high-performance networking (e.g., InfiniBand, high-speed Ethernet).

  • Support hardware lifecycle activities including installation, maintenance, and upgrades.

  • Conduct capacity planning based on system utilization trends and future demand.


Operations, Monitoring & Continuous Improvement


  • Perform system health checks, monitoring, and incident tracking for HPC and CAE environments.

  • Document system configurations, procedures, incidents, and best practices.

  • Track outages, analyze root causes, and implement preventive measures.

  • Follow change management processes for system updates and deployments.

  • Provide accurate reporting (e.g., utilization, incidents, system performance) and support project initiatives.


Minimum qualifications

Required Education & Experience:


  • 7+ years of Linux system administration experience (preferably RHEL environments).

  • Bachelor’s degree in mechanical engineering, electrical engineering, computer engineering, computer science, or related field; and/or commensurate work experience

  • Hands-on experience managing HPC clusters and job schedulers (LSF, Slurm, PBS, or similar).

  • Proven experience in CAE application support and integration.

  • Strong scripting skills (Bash, Shell, Perl, or equivalent).

  • Experience with OS deployment, patching, and system automation.

  • Solid understanding of enterprise server hardware, storage, and networking fundamentals.

  • Experience with CAE tools such as Ansys, LS-DYNA, Nastran, or similar.

  • Familiarity with high-performance networking technologies is plus (e.g., InfiniBand).

  • Experience developing internal tools or dashboards are plus (e.g., PHP or web-based tooling).


Position Type/Expected Hours of Work


  • Hybrid Full-time contract: Standard business hours with flexibility required to support maintenance windows and critical production issues.

  • Occasional after-hours or weekend work may be required based on business needs.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

CAE HPC System Administrator
CAE HPC System Administrator

Toyota Tsusho Systems • Kansas

On-site
USD 90,000 - 130,000
CAE HPC Systems Admin - Linux, Scheduling & CAE Apps
CAE HPC Systems Admin - Linux, Scheduling & CAE Apps

Toyota Tsusho Systems Corporation • Saline (MI)

Hybrid
USD 110,000 - 170,000
CAE HPC Systems Engineer - Linux, Scheduling & Automation
CAE HPC Systems Engineer - Linux, Scheduling & Automation

Toyota Tsusho Systems • Kansas

Hybrid
USD 90,000 - 130,000
CAE Engineer
CAE Engineer

Saigepartners • San Jose (CA)

On-site
USD 90,000 - 130,000
HPC Engineer
HPC Engineer

Tata Consultancy Services • Indianapolis (IN)

On-site
USD 75,000 - 80,000
Discretionary annual incentive
Comprehensive medical coverage (M/D/V,
Family support leaves
+6
Senior HPC Systems Administrator
Senior HPC Systems Administrator

RedLine Performance Solutions, LLC. • Berkeley (CA), Northern (KY)

Hybrid
USD 140,000 - 190,000
paid time off
401k match
health care benefits
HPC Linux Administrator
HPC Linux Administrator

TotalCAE • Plymouth (MI)

On-site
USD 95,000 - 130,000
CAE IT Support & Ticket Management Specialist - Columbus OHIO
CAE IT Support & Ticket Management Specialist - Columbus OHIO

aspwebsolutions • Columbus (OH)

On-site
USD 52,000 - 76,000
Two Week Vacation
Paid Medical/Dental/Vision
401k
+1
Lead HPC and Systems Engineer
Lead HPC and Systems Engineer

EPAM Systems • United States

On-site
USD 150,000 - 190,000
Sr. HPC Systems Engineer (High Performance Computing)
Sr. HPC Systems Engineer (High Performance Computing)

SpaceX • El Segundo (CA)

On-site
USD 165,000 - 265,000
Medical, Vision & Dental
401(k) plan
Disability insurance
+4