HPC Engineer

PlusWealth Group

Gurugram District

On-site

INR 1,800,000 - 2,400,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical insurance
Catered meals
Generous paid time off
Performance-based bonuses

Job summary

PlusWealth Capital Management LLP is seeking an experienced HPC Engineer to design, deploy, and manage high-performance computing environments. The role emphasizes Slurm deployment, BeeGFS/Lustre file systems, Linux administration, and infrastructure automation.

You will build and support HPC clusters, monitor performance, and collaborate with cross-functional teams to optimize throughput and efficiency across compute, storage, and networking layers.

Qualifications

  • 5–8 years hands-on HPC administration experience.
  • Strong end-to-end Slurm deployment and administration.
  • Experience with BeeGFS or Lustre file systems.
  • Solid Linux system administration (RHEL/Ubuntu).
  • Solid HPC networking knowledge (Ethernet, InfiniBand/RDMA).
  • Experience with monitoring tools (Prometheus, Grafana, ELK/OpenSearch).
  • Scripting in Bash and/or Python.
  • Strong troubleshooting and problem-solving capabilities.

Responsibilities

  • Design, deploy, and administer HPC clusters from scratch.
  • Install, configure, and manage Slurm components.
  • Deploy and administer BeeGFS/Lustre file systems.
  • Monitor cluster health, performance, and resource usage.
  • Troubleshoot compute, storage, networking, and scheduling issues.
  • Plan capacity, patching, upgrades, and maintenance.
  • Automate routine tasks using Bash, Python, or Ansible.
  • Collaborate with infra, networking, and application teams to optimize HPC performance.

Skills

HPC administration
Slurm deployment
BeeGFS/Weka
Linux admin
HPC networking
Monitoring tools
Bash/Python
Troubleshooting

Tools

BeeGFS
Lustre
Ansible

Job description

About Us

PlusWealth Capital Management LLP is a proprietary high-frequency trading firm, active in multiple markets including equities, options, and futures. We thrive on building cutting edge, data-driven, and tech-based trading algorithms. As a dynamic, machine-learning oriented trading platform, we embody the ethos of THINK. TECH. TRADE. If you share our vision, we’d love to have you onboard.

About the Role

We are looking for a skilled HPC Engineer in designing, deploying, and managing High Performance Computing (HPC) environments. The ideal candidate should have strong hands‑on expertise in Slurm Workload Manager, Parallel File Systems (PFS), Linux administration, and HPC infrastructure operations. The role requires someone who can independently build and support HPC clusters.

Key Responsibilities
  • Design, deploy, and administer HPC clusters from scratch.
  • Install, configure, and manage Slurm (Controller, Compute Nodes, Munge, Accounting, Partitions, QOS, Scheduling, etc.).
  • Deploy and administer Parallel File Systems such as BeeGFS, Lustre
  • Monitor cluster health, performance, and resource utilization using enterprise monitoring tools.
  • Troubleshoot issues across compute, storage, networking, and job scheduling.
  • Perform capacity planning, patching, upgrades, and operational maintenance.
  • Automate routine administration tasks using Bash, Python, or Ansible.
  • Work closely with infrastructure, networking, and application teams to optimize HPC performance.
Required Skills
  • 5–8 years of hands‑on experience in HPC administration.
  • Strong expertise in Slurm Workload Manager with end‑to‑end deployment and administration.
  • Experience with at least one Parallel File System (BeeGFS or Weka).
  • Strong Linux administration skills (RHEL/Ubuntu).
  • Good understanding of HPC networking concepts, including Ethernet, InfiniBand/RDMA, VLANs, and storage networking.
  • Experience with monitoring tools such as Prometheus, Grafana, ELK/OpenSearch, or similar.
  • Good scripting skills in Bash and/or Python.
  • Strong troubleshooting, analytical, and problem‑solving skills.
Preferred Skills
  • Experience with GPU clusters, or cloud‑based HPC environments.
  • Knowledge of automation and configuration management tools such as Ansible.
Benefits & Perks:
  • Competitive compensation and performance‑based bonuses.
  • Flat organizational structure with high ownership and visibility.
  • Medical insurance – we've got you and your dependents covered.
  • Catered meals/snacks for 5 working days in the office.
  • Generous paid time off policies.

Pluswealth Capital Management is an equal opportunity employer

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Linux System Administrator
Linux System Administrator

SISL Global • Chennai District

On-site
INR 800,000 - 1,200,000
Senior HPC Engineer
Senior HPC Engineer

Netweb Technologies India Ltd. • Faridabad District

On-site
INR 3,500,000 - 5,500,000
HPC Senior System Integrator/System Administrator
HPC Senior System Integrator/System Administrator

GBB • Mumbai

On-site
INR 1,000,000 - 1,500,000
HPC Engineer
HPC Engineer

Clovertex • Hyderabad

On-site
INR 1,200,000 - 2,400,000
Senior HPC Engineer
Senior HPC Engineer

Binaire Private Limited • New Delhi

On-site
INR 1,500,000 - 2,500,000
Opportunity to influence hardware selection
Ownership of high-performance compute infrastructure
HPC Admin
HPC Admin

Randstad Digital • Bengaluru

On-site
INR 900,000 - 1,500,000
HPC Engineer
HPC Engineer

Whiteblue • Chennai

On-site
INR 1,500,000 - 2,500,000
Linux Administrator
Linux Administrator

Pacefin • Gurugram District

On-site
INR 1,800,000 - 3,000,000
Catered breakfast & lunch
Group health insurance
4 weeks of annual leave
+1
Lead HPC Engineer
Lead HPC Engineer

Clovertex • Hyderabad

On-site
INR 2,000,000 - 3,000,000
HPC ADMINISTRATOR
HPC ADMINISTRATOR

Vrinda International • Bengaluru

On-site
INR 1,260,000 - 1,540,000