HPC Engineer

Clovertex

Hyderabad

On-site

INR 1,200,000 - 2,400,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Clovertex is seeking an HPC Engineer to deploy, manage, and support Linux-based HPC environments on AWS. You will work with customers and internal teams to ensure reliable, secure, and high-performing HPC infrastructure while helping users optimize workloads.

Responsibilities include administering HPC clusters, monitoring performance, and implementing automation with Bash or Python. Prior experience with SLURM/PBS/SGE and cloud HPC platforms is highly valued.

Qualifications

  • 2-5 years of experience in HPC or Linux System Administration.
  • Hands-on experience with HPC schedulers such as SLURM, PBS, or SGE.
  • Experience troubleshooting HPC clusters, storage, networking, and job scheduling issues.
  • Shell scripting (Bash) or Python for automation.
  • Understanding of cluster monitoring and performance tuning.
  • Good communication and customer-facing skills.

Responsibilities

  • Deploy, administer, and support cloud-based HPC environments on AWS.
  • Manage HPC clusters, Linux servers, and workload schedulers (SLURM/PBS).
  • Monitor cluster health, troubleshoot infrastructure and application issues, and optimize system performance.
  • Provide L2/L3 support for HPC environments, including scheduler, storage, networking, and compute-related issues.
  • Support installation and maintenance of HPC applications and user environments.
  • Automate routine operational tasks using Bash or Python scripting.
  • Collaborate with customers and engineering teams during planning, deployment, and production support.
  • Create and maintain technical documentation, SOPs, and deployment runbooks.
  • Drive continuous improvements in reliability, performance, and operational efficiency.

Skills

HPC
Linux
Shell scripting
Python
SLURM
PBS
SGE
AWS
Terraform

Tools

AWS ParallelCluster
Terraform

Job description

At Clovertex, we help organizations accelerate scientific innovation by designing, deploying, and managing cloud-native High-Performance Computing (HPC) platforms on AWS. Our solutions enable customers in Healthcare, Life Sciences, and AI to run large-scale compute workloads efficiently, securely, and at scale.

Job Summary

We are looking for a motivated HPC Engineer to deploy, manage, and support Linux-based HPC environments on AWS. You will work closely with customers and internal engineering teams to ensure reliable, secure, and high-performing HPC infrastructure while helping users get the best out of their workloads.

Key Responsibilities
  • Deploy, administer, and support cloud-based HPC environments on AWS.
  • Manage HPC clusters, Linux servers, and workload schedulers (SLURM/PBS).
  • Monitor cluster health, troubleshoot infrastructure and application issues, and optimize system performance.
  • Provide L2/L3 support for HPC environments, including scheduler, storage, networking, and compute-related issues.
  • Support installation and maintenance of HPC applications and user environments.
  • Automate routine operational tasks using Bash or Python scripting.
  • Collaborate with customers and engineering teams during planning, deployment, and production support.
  • Create and maintain technical documentation, SOPs, and deployment runbooks.
  • Drive continuous improvements in reliability, performance, and operational efficiency.
Required Skills
  • 2-5 years of experience in High-Performance Computing (HPC) or Linux System Administration.
  • Hands-on experience with HPC schedulers such as SLURM, PBS, or SGE.
  • Experience troubleshooting HPC clusters, storage, networking, and job scheduling issues.
  • Shell scripting (Bash) or Python for automation.
  • Understanding of cluster monitoring and performance tuning.
  • Good communication and customer-facing skills.
Nice to Have
  • Experience with AWS ParallelCluster, AWS PCS, or other cloud HPC platforms.
  • Knowledge of Terraform, Infrastructure as Code, or DevOps practices.
  • Exposure to AI/ML workloads or scientific computing.
  • Experience in Healthcare, Pharma, or Life Sciences.
What We're Looking For
  • A strong learning attitude and passion for new technologies.
  • A proactive, self-driven individual who takes ownership.
  • A go-getter who enjoys solving complex technical challenges.
  • Strong analytical and troubleshooting skills.Excellent collaboration and communication abilities.

If you're passionate about Cloud, AI, and High-Performance Computing and want to build solutions that power scientific discovery, we'd love to hear from you.

Required Skills

DevOps Agile Methodologies Shell Scripting Artificial Intelligence Cloud Computing AWS Programming Languages Terraform

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior HPC Engineer
Senior HPC Engineer

Clovertex • Hyderabad

On-site
INR 2,400,000 - 4,200,000
Lead HPC Engineer
Lead HPC Engineer

Clovertex • Hyderabad

On-site
INR 2,000,000 - 3,000,000
HPC Engineer
HPC Engineer

PlusWealth Group • Gurugram District

On-site
INR 900,000 - 1,300,000
Medical insurance
Meals at office
Generous paid time off
AWS Senior HPC Engineer SME
AWS Senior HPC Engineer SME

Tata Consultancy Services • Hyderabad, Chennai District, Bengaluru

On-site
INR 1,800,000 - 3,200,000
Linux System Administrator
Linux System Administrator

SISL Global • Chennai District

On-site
INR 800,000 - 1,200,000
HPC Engineer
HPC Engineer

Whiteblue • Chennai

On-site
INR 1,500,000 - 2,500,000
HPC Engineer
HPC Engineer

Yotta Data Services Private Limited • Mumbai

On-site
INR 400,000 - 700,000
Senior HPC Engineer
Senior HPC Engineer

Netweb Technologies India Ltd. • Faridabad District

On-site
INR 1,500,000 - 2,100,000
HPC Admin
HPC Admin

SHI Solutions India Pvt. Ltd. • Maharashtra

On-site
INR 1,000,000 - 1,500,000
Senior HPC Engineer
Senior HPC Engineer

Binaire Private Limited • New Delhi

On-site
INR 1,500,000 - 2,500,000
Opportunity to influence hardware selection
Ownership of high-performance compute infrastructure