Senior HPC Engineer

RCH Solutions

United States

Remote

USD 120,000 - 180,000

Full time

11 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Competitive salary + bonus
Health and wellness benefits
401(k) plan with match
Continuing education support
Team-focused culture

Job summary

RCH Solutions is seeking an HPC Engineer to deliver Compute at Scale, shaping HPC platforms for life sciences research. The role covers on-prem and cloud-based infrastructure, Linux administration, and solution architecture.

You will drive architecture, roadmaps, and best practices while mentoring junior staff and supporting customer workflows. The ideal candidate has 5+ years in HPC and cloud deployment, with strong scripting, automation, and collaboration skills.

Qualifications

  • Bachelor’s or Master’s degree in Computer Science or related field.
  • 5+ years administering HPC clusters and systems.
  • Experience with SLURM and Grid Engine scheduling software preferred.
  • 5+ years in Solution Architecture or Cloud Infrastructure deployment.
  • 5+ years developing compute solutions for Scientific/Research IT.

Responsibilities

  • Design and evolve HPC platforms and support scientific workflows.
  • Provide full-stack support for customer compute-at-scale initiatives.
  • Architect and implement on-prem and cloud infrastructure (AWS/GCP).
  • Document assets, provide mentoring, and ensure security/compliance.
  • Troubleshoot hardware, software, and networking issues across engagements.

Skills

HPC cluster admin
SLURM
Grid Engine
Solution architecture
Cloud infrastructure
Nextflow
Ansible
Terraform
CloudFormation
Python
R
Docker
Singularity
NVIDIA DGX

Education

Bachelor’s or Master’s in CS or related

Tools

POSIT products
Docker
Singularity
Nvidia DGX
Terraform
CloudFormation
Ansible

Job description

About Us

RCH Solutions is an established and rapidly growing global provider of computational, research, and data science expertise within Life Sciences and Healthcare. At RCH Solutions, our team rallies around a culture crafted for learning and achieving. We’re relentless in our pursuit for innovation and demanding of ourselves to deliver a ground-breaking computing experience for our clients, so that they can deliver life-saving science to humanity.

Core Values

At RCH, our Core Values are more than just words—they represent the threads that weave together the fabric of our culture. Used as a guide when interviewing new team members; as a barometer when evaluating our performance as individuals and teams, and even when deciding which customers to work with, RCH’s Values embody the behaviors upon which we measure our success and create a framework for our growth as people and professionals.

Our Core Values:

  • Embrace Excellence: We strive for best in class delivery of innovation and service.
  • Be Accountable: Integrity, ownership and accountability are non negotiables.
  • Adventure Together: We are committed to fostering a culture that embraces continuous improvement.
  • Succeed as a Team: We believe harnessing the power of a team drives outcomes not achievable by individuals.
  • Boundaries and Balance: Work-life balance is a core facet of our culture.

If you share in our core values, then we encourage you to continue reading this posting as you may have found a great home for your career.

Job Description

RCH Solutions is seeking an HPC Engineer to work closely with customer stakeholders, scientists, and IT professionals to deliver Compute at Scale and support our customer's scientific initiatives. The objectives for this role center on developing, evolving, and administering HPC platforms along with support for Scientific applications, workflows, and other related infrastructure both on-prem and Cloud hosted. Our ideal candidate also has hands on experience with Linux system administration as well as solution architecting and engineering (on-prem and cloud based) and will be instrumental in transforming how IT computing services are leveraged to support our client's growth. This role will involve driving architecture, roadmaps, and execution of projects to establish and operate IT infrastructure best practices for customers.

Responsibilities include full stack support - design and evolution of platforms, application administration, supporting customer workflows, profiling and performance tuning, monitoring and maintenance of scoped systems, platform and systems administration, troubleshooting hardware, software, and networking related issues, solution architecting and hands on engineering (on-prem + Cloud), as well as documentation. You will use your experience in these technologies to provide top of the line consulting services and recommendations to clients. These would be performed as part of customer Research or Analytics initiatives as well as in a consultative, advisory, or customer support manner.

Specific focuses and responsibilities include:
  • Collaborating with cross-discipline team members and customers to deliver HPC and peripheral Compute at Scale services.
  • Thorough understanding of related industry best practices.
  • Supporting internal and customer Architecture and Design efforts.
  • Supporting customers with their workflow pipelines (advisory and hands-on).
  • Comprehensively documenting new and existing computational assets.
  • Maintaining the flexibility to pivot as engagement scopes may evolve.
  • Support for AWS & GCP Cloud applications, migrations, and modernization.
  • CloudOps / IaC for on-going platform management.
  • Setup and configuration of AWS & GCP Cloud infrastructure for new platform builds.
  • Ensuring system compliance with company security standards and applicable regulatory requirements.
  • Transition support for modernized services to operational teams.
  • Provide engineering level troubleshooting and services restoration for operational issues as they arise on supported platforms.
  • Provide training/mentorship for junior level team members.
  • Escalation point on multiple engagements to ensure resolution.
Essential Qualifications
  • A bachelor’s degree or master’s degree in Computer Science or related field.
  • 5+ years of experience administering HPC clusters and systems.
  • Experience with SLURM and Grid Engine scheduling software preferred.
  • 5+ years of professional experience in Solution Architecture or Cloud Infrastructure Deployment and support.
  • 5+ years professional experience developing or administering compute solutions for Scientific / Research IT domains, Life Sciences being preferred.
  • Experience with POSIT products (Package Manager, Connect, Workbench) either in an end-user or administrator capacity.
  • Experience developing scientific workflows on HPC systems using Nextflow.
  • Extensive command-line system administration experience:
  • User and group management
  • Advanced knowledge of Active Directory, DNS, DHCP, LDAP, NFS, SMB
  • Building applications from source code, installing, maintaining, and troubleshooting application-level Linux and scientific software in line with industry best practices.
  • Installation of Linux operating system and fine tuning
  • Familiarity with leveraging and maintaining Linux package management systems
  • Intermediate OS level networking knowledge.
  • Experience using with scripting tools, automation tools, and configuration management tools
  • Ansible, Terraform and Cloud Formation experience preferred.
  • Experience administering and integrating Scientific / Research applications.
  • Strong time-management skills; able to complete projects in a timely manner, plan and prioritize tasks while keeping leadership and stakeholders updated regularly on status.
  • Excellent communication skills, including preparation of written documentation for IT colleagues and end users.
  • Proactive thinking skills to identify potential issues and solution options prior to incidents occurring.
  • Extreme attention to detail is needed to interface with multiple different clients simultaneously.
  • Ability to understand and analyze complex technical problems and situations.
  • Candidates must be a passionate engineer with a strong vision and a desire to stay on top of trends in the Scientific Computing sector.
  • Ability to work independently or with a team
  • Ability to take a project from start to finish with minimal supervision
Preferred Qualifications

RCH provides services and solutions for the unique challenges of Life Sciences advanced computing, and leverages teams with cross‑functional IT skills to meet these challenges. The ideal candidates for this role will have experience working with cross‑functional IT (Public Cloud skills being a plus) and sciences skillsets.

  • Experience with Python, R, or other related data science programming languages.
  • Experience working with databases and/or supporting.
  • Experience managing large amounts of data effectively.
  • Experience working with AI/ML technologies.
  • Experience with containerizing compute workload via Docker or Singularity.
  • Experience with Nvidia DGX systems.
Additional information

Great talent should benefit from a great work environment. If you join our team, you’ll have access to:

  • A competitive salary and bonus package based on experience
  • Comprehensive health and wellness benefits, including Medical, Dental, and Vision Insurance
  • Company-provided Life and Long-Term Disability Insurance
  • Company-sponsored 401(k) Plan
  • Company-provided continuing education benefit
  • Team-focused culture and unlimited opportunity for advancement
Notes

This is a fully remote position and the candidate will be required to work on an East Coast (US) schedule.

Role is only open to applicants not needing sponsorship now or in the future, no third parties please.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

GCP Cloud Engineer
GCP Cloud Engineer

RCH Solutions • Wayne (PA)

Remote
USD 110,000 - 150,000
Remote HPC Architect for Compute-at-Scale
Remote HPC Architect for Compute-at-Scale

RCH Solutions • United States

Remote
USD 120,000 - 180,000
Competitive salary + bonus
Health and wellness benefits
401(k) plan with match
+2
HPC Orchestration Architect
HPC Orchestration Architect

NMC2 • Dallas (TX), Northern (KY)

On-site
USD 190,000 - 270,000
Lunch stipend
Company-Paid Benefits (medical, dental
401(k) match
Senior HPC Engineer, Services
Senior HPC Engineer, Services

Quiet Capital • United States

On-site
USD 100,000 - 130,000
Senior HPC Systems Administrator
Senior HPC Systems Administrator

RedLine • Berkeley (CA)

Remote
USD 140,000 - 190,000
Paid time off
401k match
Health care benefits
Senior HPC Systems Administrator
Senior HPC Systems Administrator

RedLine Performance Solutions • Berkeley (CA)

Remote
USD 140,000 - 190,000
Paid time off
401k match
Health care benefits
HPC Cloud Systems Administrator
HPC Cloud Systems Administrator

RedLine Performance Solutions • College Park (MD)

Remote
USD 120,000 - 150,000
Health benefits
401(k) match
Paid time off
Customer Success Engineer - Remote in the US
Customer Success Engineer - Remote in the US

Rescale • North Township (IN)

On-site
USD 85,000 - 130,000
Senior HPC Engineer, Services
Senior HPC Engineer, Services

Rescale • United States

On-site
USD 100,000 - 150,000
Senior HPC Systems Administrator
Senior HPC Systems Administrator

RedLine Performance Solutions, LLC. • Berkeley (CA), Northern (KY)

Hybrid
USD 140,000 - 190,000
paid time off
401k match
health care benefits