HPC Engineer

Tata Consultancy Services

Indianapolis (IN)

On-site

USD 75,000 - 80,000

Full time

2 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Discretionary annual incentive
Comprehensive medical coverage (M/D/V,
Family support leaves
Insurance options (auto/home, identity
Commuter benefits
Certification & training reimbursement
Vacation and holidays
401K plan and performance bonus
College fund and student loan refinanc

Job summary

Tata Consultancy Services is seeking an HPC Engineer to design, build, and maintain high-performance computing clusters for on-premises and cloud environments. The role emphasizes Open OnDemand integration, Slurm/PBS schedulers, and strong troubleshooting.

You will collaborate with researchers and engineers, maintain documentation, and implement security hardening while optimizing performance across diverse workloads. The position offers competitive compensation and benefits.

Qualifications

  • Experience building and deploying scientific applications and module environments in on-premises and cloud-based HPC environments.
  • Familiarity with Open OnDemand, AWS cloud-based infrastructure, and containerization of HPC applications (e.g., Singularity Registry HPC)
  • Knowledge of the PBS Grid Engine, Slurm job scheduler.
  • Excellent troubleshooting skills with the ability to resolve application-related issues.
  • Strong documentation and diagramming abilities.
  • Ability to work collaboratively within a team and communicate effectively.
  • Linux Administration
  • Linux OS internal knowledge
  • Linux Hardening
  • LDAP / AD intergrations
  • Bash Scripting

Responsibilities

  • System Design and Implementation: HPC Engineers Design, Build and configure HPC clusters including hardware and software components
  • System administration: Manage and maintain the HPC infrastructure including operating systems, storage, and networking
  • Performance Optimization: analyze system performance, identify bottlenecks, and implement solutions to optimize performance for various applications
  • Troubleshoot and support: Diagnose and resolve issues with the HPC system, providing support to researchers and users
  • Scripting and Automating: Develop scripts and automation tools to streamline routine tasks and improve efficiency
  • Collaboration: Work closely with researchers, data scientists, and other engineers to understand their needs and provide effective solutions
  • Documentation – Maintain clear and accurate documentation of system configurations, procedures and troubleshooting steps
  • Monitor and Maintenance – Monitor system health, perform maintenance tasks, and plan for upgrades and new technologies
  • Security: Ensure the secure and effective operation of HPC

Skills

HPC design
Open OnDemand
AWS
Containerization
PBS Grid Engine
Slurm
Troubleshooting
Documentation
Team collaboration
Linux Administration
Linux knowledge
Linux hardening
LDAP/AD integrations
Bash scripting

Education

Bachelor of Computer Science

Tools

Ansible
Splunk
Nagios
Base Command Manager
Jira

Job description

  • Experience in building and deploying scientific applications and module environments in on-premises and cloud-based HPC environments.
  • Familiarity with Open OnDemand, AWS cloud-based infrastructure, and containerization of HPC applications (e.g., Singularity Registry HPC)
  • Knowledge of the PBS Grid Engine, Slurm job scheduler.
  • Excellent troubleshooting skills with the ability to resolve application-related issues.
  • Strong documentation and diagramming abilities.
  • Ability to work collaboratively within a team and communicate effectively.
  • Linux OS internal knowledge
  • LDAP / AD intergrations
  • Bash Scripting
Job Description
Must Have Technical/Functional Skills
  • Experience in building and deploying scientific applications and module environments in on-premises and cloud-based HPC environments.
  • Familiarity with Open OnDemand, AWS cloud-based infrastructure, and containerization of HPC applications (e.g., Singularity Registry HPC)
  • Knowledge of the PBS Grid Engine, Slurm job scheduler.
  • Excellent troubleshooting skills with the ability to resolve application-related issues.
  • Strong documentation and diagramming abilities.
  • Ability to work collaboratively within a team and communicate effectively.
  • Linux Administration
  • Linux OS internal knowledge
  • Linux Hardening
  • LDAP / AD intergrations
  • Bash Scripting
Preferred Qualifications
  • Ansible Configuration Management
  • Splunk and Nagios monitoring
  • Base Command Manager (formerly called Bright Cluster Manager)
  • Jira
Roles & Responsibilities
  • System Design and Implementation: HPC Engineers Design, Build and configure HPC clusters including hardware and software components
  • System administration: Manage and maintain the HPC infrastructure including operating systems, storage, and networking
  • Performance Optimization: analyze system performance, identify bottlenecks, and implement solutions to optimize performance for various applications
  • Troubleshoot and support: Diagnose and resolve issues with the HPC system, providing support to researchers and users
  • Scripting and Automating: Develop scripts and automation tools to streamline routine tasks and improve efficiency
  • Collaboration: Work closely with researchers, data scientists, and other engineers to understand their needs and provide effective solutions
  • Documentation – Maintain clear and accurate documentation of system configurations, procedures and troubleshooting steps
  • Monitor and Maintenance – Monitor system health, perform maintenance tasks, and plan for upgrades and new technologies
  • Security: Ensure the secure and effective operation of HPC

Salary Range: $75,000 -$80,000 year

TCS Employee Benefits Summary
  • Discretionary Annual Incentive.
  • Comprehensive Medical Coverage: Medical & Health, Dental & Vision, Disability Planning & Insurance, Pet Insurance Plans.
  • Family Support: Maternal & Parental Leaves.
  • Insurance Options: Auto & Home Insurance, Identity Theft Protection.
  • Convenience & Professional Growth: Commuter Benefits & Certification & amp; Training Reimbursement.
  • Time Off: Vacation, Time Off, Sick Leave & Holidays.
  • Legal & Financial Assistance: Legal Assistance, 401K Plan, Performance Bonus, College Fund, Student Loan Refinancing.
Qualifications

BACHELOR OF COMPUTER SCIENCE

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

HPC Software Integration Engineer
HPC Software Integration Engineer

GIGATEC Engineering • Corridor North (MD)

On-site
USD 80,000 - 120,000
100% Paid Healthcare
10% 401k in every paycheck
100% Fully Vested!
HPC Software Integration Engineer
HPC Software Integration Engineer

GIGATEC Engineering • Maryland

On-site
USD 80,000 - 120,000
100% Paid Healthcare
10% 401k in every paycheck
100% Fully Vested!
Software Integration Engineer - Annapolis Junction, MD
Software Integration Engineer - Annapolis Junction, MD

AHU Technologies Inc • Washington

On-site
USD 246,000 - 253,000
100% Employer Paid Medical Premiums up to $25,000
401K with 10% employer contribution
20 Paid Time Off Days
+2
HPC Operations Engineer
HPC Operations Engineer

Career Techniques • New York (NY)

Hybrid
USD 175,000 - 225,000
HPC Orchestration Architect
HPC Orchestration Architect

NMC2 • Dallas (TX), Northern (KY)

Hybrid
USD 190,000 - 270,000
Lunch stipend
Company-Paid Benefits (medical, dental
401(k) match
HPC Linux Systems Engineer
HPC Linux Systems Engineer

Cadre5 • Knoxville (TN)

On-site
USD 120,000 - 160,000
Excellent medical insurance
Employer-paid benefits
HPC Software Engineer
HPC Software Engineer

Cornerstone Defense LLC • Colorado Springs (CO)

On-site
USD 140,000 - 230,000
Senior HPC DevOps Engineer | TS/SCI w/ MD POLY Security Clearance required
Senior HPC DevOps Engineer | TS/SCI w/ MD POLY Security Clearance required

Capstone Technology Partners • College Park (MD)

On-site
USD 222,000 - 257,000
Four weeks paid time off
Eleven paid holidays
401k with employer contributions and 3
+2
AWS/ HPC Consultant (TS/SCI Clearance Preferred) Remote, Remote
AWS/ HPC Consultant (TS/SCI Clearance Preferred) Remote, Remote

Strategic Business Systems, Inc (SBS) • Chantilly (VA)

Hybrid
USD 120,000 - 180,000
Flexible work arrangements
Software Integration Engineer HPC
Software Integration Engineer HPC

AHU Technologies Inc • Washington

On-site
USD 246,000 - 253,000
Medical Coverage (Multiple Plans, Employer Paid)
401(k): 10% company contribution (fully vested per pay period)
Performance-based Bonuses