Cognizant Hiring For Sr. HPC Engineer

Cognizant

Pune District, Chennai District, Bengaluru

On-site

INR 3,000,000 - 6,000,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Cognizant in Pune area seeks seasoned HPC infrastructure professionals to design, deploy and maintain HPC clusters, storage systems, and networking for demanding workloads. You will optimize compute, storage and interconnects, ensure security/compliance, and mentor a team across operations and automation.

The role focuses on hands-on leadership with Linux (CentOS/RHEL), Lustre/GPFS, InfiniBand and cluster management tools.

Qualifications

  • Strong knowledge of Linux OS (CentOS/RHEL) and HPC hardware platforms (HPE, NVIDIA DGX).
  • Hands-on experience with parallel file systems (Lustre, GPFS) and enterprise storage.
  • Proficiency in InfiniBand networking and high-speed interconnects.
  • Familiarity with job schedulers and cluster management tools (IBM LSF, Bright Cluster Manager, Altair Grid Manager).
  • Automation & scripting with Ansible, Chef, Cobbler; Bash and Python proficiency.
  • Experience with AWS ParallelCluster or similar cloud-based HPC solutions.
  • Monitoring with Zabbix, Grafana, ELK Stack for health and performance.

Responsibilities

  • HPC Infrastructure Management and operation of HPC clusters and storage systems.
  • Configure and maintain InfiniBand networking and compute/storage hardware.
  • Manage cluster scheduling and resource allocation using industry tools.
  • Automate provisioning and configuration with Cobbler, Chef, Ansible, AWS ParallelCluster.
  • Monitor systems with Zabbix, Grafana, and ELK Stack; tune performance and reliability.
  • Ensure security compliance and documentation across the HPC environment.

Skills

Linux OS expertise
HPC hardware expertise
Networking (InfiniBand)
Storage administration
Security & Compliance
Monitoring & logging
Automation & scripting
Cluster scheduling
Team leadership
Documentation

Education

Bachelor's or Master's degree in Computer Science or Engineering

Tools

Bright Cluster Manager
Altair Grid Manager
IBM LSF
Cobbler
Chef
Ansible
AWS ParallelCluster
Zabbix
Grafana
ELK Stack
Lustre
GPFS
InfiniBand
HPE
NVIDIA DGX
CentOS/RHEL

Job description

Role Overview

We are seeking seasoned professionals with deep expertise in operating and managing High-Performance Computing (HPC) platforms. The ideal candidate will have hands‑on experience in designing, deploying, and maintaining HPC clusters, storage systems, and networking infrastructure, leveraging industry-leading tools and technologies.

Key Responsibilities
  • HPC Infrastructure Management
  • Operate and maintain HPC clusters based on CentOS, RHEL, and hardware platforms like HPE and NVIDIA DGX.
  • Ensure optimal performance, scalability, and reliability of compute resources.
  • Storage Administration
  • Manage large‑scale storage systems including Dell Isilon, VAST Storage, Lustre, and GPFS.
  • Implement data lifecycle management and optimize storage performance for HPC workloads.
  • Networking
  • Configure and maintain InfiniBand‑based networking for low‑latency, high‑bandwidth communication.
  • Troubleshoot network performance issues and ensure secure connectivity.
  • Cluster and Job Scheduling
  • Administer cluster management tools such as Bright Cluster Manager, Altair Grid Manager, and IBM LSF.
  • Optimize job scheduling and resource allocation for diverse workloads.
  • Monitoring and Automation
  • Implement monitoring solutions using Zabbix, Grafana, and ELK Stack.
  • Automate provisioning and configuration using Cobbler, Chef, Ansible, and AWS ParallelCluster.
  • Performance Tuning & Troubleshooting
  • Conduct performance benchmarking and tuning for HPC workloads.
  • Diagnose and resolve hardware/software issues across compute, storage, and network layers.
  • Security & Compliance
  • Ensure HPC environment adheres to security best practices and compliance standards.
Required Skills & Qualifications
  • Technical Expertise
  • Strong knowledge of Linux OS (CentOS, RHEL) and HPC hardware platforms (HPE, NVIDIA DGX).
  • Hands‑on experience with parallel file systems (Lustre, GPFS) and enterprise storage solutions.
  • Proficiency in InfiniBand networking and high‑speed interconnects.
  • Familiarity with job schedulers and cluster management tools (IBM LSF, Bright Cluster Manager, Altair Grid Manager).
  • Automation & Scripting
  • Expertise in Ansible, Chef, Cobbler, and scripting languages (Bash, Python).
  • Experience with AWS ParallelCluster or similar cloud‑based HPC solutions.
  • Monitoring & Logging
  • Practical experience with Zabbix, Grafana, and ELK Stack for system health and performance monitoring.
  • Soft Skills
  • Strong problem‑solving and analytical skills.
  • Ability to work in a fast‑paced environment and lead technical teams.
  • Excellent communication and documentation skills.
Preferred Qualifications
  • Exposure to AI/ML workloads on HPC clusters.
  • Experience with containerization (Docker, Singularity) in HPC environments.
  • Knowledge of security hardening for HPC systems.
Education
  • Bachelors or Master’s degree in Computer Science, Engineering, or related field.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Sr. HPC ENGINEER
Sr. HPC ENGINEER

Cognizant • Hyderabad

On-site
INR 800,000 - 1,200,000
HighPerformance Computing ( HPC) Administrator
HighPerformance Computing ( HPC) Administrator

VIT • Chennai District

On-site
INR 900,000 - 1,300,000
HPC Engineer
HPC Engineer

Esconet Technologies • New Delhi

On-site
INR 1,200,000 - 1,700,000
Cognizant Hiring For HPC , Github Pipelines, AWS
Cognizant Hiring For HPC , Github Pipelines, AWS

Cognizant • Hyderabad, Chennai District, Bengaluru

Hybrid
INR 1,500,000 - 2,800,000
Linux System Administrator
Linux System Administrator

SISL Global • Chennai District

On-site
INR 800,000 - 1,200,000
Senior HPC Platform Architect
Senior HPC Platform Architect

NVIDIA Gruppe • Bengaluru

On-site
INR 400,000 - 900,000
Lead HPC Engineer
Lead HPC Engineer

Clovertex • Hyderabad

On-site
INR 2,000,000 - 3,000,000
Staff Data Engineer (HPC cluster software such as Slurm, NC, LSF or Grid Engine) experience wit[...]
Staff Data Engineer (HPC cluster software such as Slurm, NC, LSF or Grid Engine) experience wit[...]

SanDisk • Bengaluru

On-site
INR 3,000,000 - 5,000,000
HPC Admin
HPC Admin

5 Star Recruitment • Chennai District

On-site
INR 2,000,000 - 4,000,000
Senior HPC Engineer SME/Architect
Senior HPC Engineer SME/Architect

Tata Consultancy Services • Hyderabad, Chennai District, Bengaluru

On-site
INR 1,800,000 - 3,000,000