Systems Engineer HPC&GPU Infrastructure - TS/SCI

Sunayu Llc

Bethesda, Northern (MD, KY)

Hybrid

USD 140,000 - 180,000

Full time

3 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Medical Plan Options
Dental and Vision
Life/AD&D Insurance
Disability Insurance
EAP
Training and Educational Assistance
PTO and Holidays
401k with match

Job summary

Sunayu, LLC in Bethesda, MD seeks a Senior Systems Engineer for HPC and GPU infrastructure. You will design, deploy, and optimize GPU clusters for IC community customers, with on-site work at the Intel Intelligence Community Campus.

The role emphasizes Linux-based hardware/software integration, HPC tooling, and performance tuning across Linux platforms. DoD/IC environment and TS/SCI clearance considerations apply.

Qualifications

  • Bachelor's degree in CS/EE or related field; higher degree considered.
  • 12+ years of systems engineering experience.
  • Expertise in Linux OS integration and hardware architecture.

Responsibilities

  • Install and maintain GPU/HPC hardware on-prem and cloud.Analyze performance and optimize across hardware and software.
  • Install and configure HPC/GPU job scheduling and workload management platforms (Slurm, PBS, Airflow, Kubernetes).
  • Work on power management and efficiency for GPU systems on Linux.
  • Design, execute tests to validate GPU performance and functionality; expand test suite.
  • Maintain technical documentation and Linux-specific best practices for GPU development.
  • Stay updated on GPU industry trends and Linux-based optimization approaches.

Skills

Linux
HPC
Kubernetes
NVIDIA GPUs
Python
Bash
Automation

Education

Bachelor's degree in Computer Science or Electrical Engineering

Tools

Slurm
PBS
Apache Airflow
Kubernetes
Docker
Prometheus Grafana

Job description

Location: Bethesda, MD

Category: Systems Engineering
Travel Required:No
Remote Type: No
Clearance: TS/SCI

Sunayu, LLC is hiring a Senior Systems Engineer – HPC & GPU Infrastructure with a deep understanding of operating systems, hardware, Kubernetes, and NVIDIA GPU products. As a senior Systems Engineer – HPC & GPU Infrastructure, you will play a pivotal role in designing, developing, and optimizing GPU clusters for the IC community customers.

This is a 100% on-site position. All work must be performed at the customer site in Bethesda at the Intelligence Community Campus.

Primary Responsibilities:
  • HPC and GPU environment engineering: Contribute to the installation and maintenance of GPU and HPC hardware on-prem and in the cloud, providing insights into hardware performance to ensure efficient interaction with software components.
  • Performance Optimization: Analyze HPC/GPU cluster performance, identify bottlenecks, and develop strategies to enhance performance across various applications in Linux, addressing both hardware and software considerations. Regularly monitor and improve performance.
  • HPC/GPU tooling: Install and configure HPC/GPU job scheduling and workload management platforms such as Slurm , PBS , Apache Airflow , Kubernetes
  • Power Efficiency: Work on power management techniques to optimize GPU power consumption, ensuring efficient operation on both mobile and desktop Linux platforms. Continuously assess and enhance power efficiency strategies.
  • Testing and Validation: Design and execute tests to validate GPU performance and functionality on Linux, including stress testing, benchmarking, and debugging to ensure robust operation. Maintain and expand the testing suite.
  • Documentation: Maintain comprehensive technical documentation, including architectural specifications, code documentation, and Linux-specific best practices for GPU development. Keep documentation up to date with changes and improvements.
  • Industry Insight: Stay updated on the latest trends, innovations, and competitive landscapes within the GPU industry, contributing to research efforts and proposing Linux-specific approaches to GPU design and optimization. Share regular updates and insights with the team.
Basic Requirements:
  • Bachelor's or higher degree in Computer Science, Electrical Engineering, or a related field. Additional years of experience may be considered in lieu of a degree.
  • 12+ years of relevant systems engineering experience
  • Expertise in operating system integration for Linux.
  • Strong understanding of computer hardware architecture, particularly as it relates to Linux systems.
  • Knowledge of parallel computing, graphics algorithms, and real-time rendering in Linux environments.
  • Excellent problem-solving skills and the ability to collaborate within a team.
  • Strong communication skills for conveying technical information in a Linux context.
  • Proficiency with scripting languages such as Python or BASH.
  • Proficiency with automation tools such Ansible, Puppet, Salt, Terraform, etc.
  • Candidate must, at a minimum, meet DoD 8570.11- IAT Level II certification requirements (currently Security+ CE, CCNA-Security, GICSP, GSEC, or SSCP along with an appropriate computing environment (CE) certification). An IAT Level III certification would also be acceptable (CASP+, CCNP Security, CISA, CISSP, GCED, GCIH, CCSP).
Clearance Information:
  • Active TS/SCI clearance with Polygraph required OR active TS/SCI and willingness to obtain and maintain a Poly.
  • US Citizenship is required due to the nature of the government contracts we support.
Preferred Qualifications
  • Knowledge of GPU virtualization, cloud computing, and emerging Linux-based technologies in the field.
  • Experience with container technologies (Docker, Kubernetes)
  • Experience with Prometheus/Grafana for monitoring
  • Knowledge of distributed resource scheduling systems
  • Understanding data center networking hardware and cabling concepts.
  • Understanding of networking technologies such as DHCP, DNS, TCP/IP, VLANs, HSRP, and SNMP.
  • Knowledge of data center networking security principles Firewall ACLs, IPS/IDS, and Policy Based Routing.

--------------------------------------------------------

Who We Are

Sunayu, LLC serves as a premier technology partner to the Defense and Intelligence communities, delivering mission-critical engineering solutions across the nation. Our operations are anchored in a commitment to trust, accountability, and ethical transparency, ensuring the high-performance outcomes necessary to protect our country's most vital interests.

Culture

Our strength lies in our community:Our team prioritizes collaboration, professional growth, and encourages open communication. At Sunayu, we don't just secure the mission—we grow together.

Career Development

We support and encourage our team members to continue their professional growth by providing company-reimbursed training and continuing education of up to $5,000 per year. We also participate in many industry conferences and events where we share our expertise and experiences.

Pay Rate

Salary range considers factors such as (but not limited to) scope and responsibilities of the position, candidate's work experience, education/ training, key skills, as well as market and business considerations when extending an offer.

Benefits
  • 3 Medical Plan Options
  • Dental and Vision
  • FSA, DCFSA, HSA
  • Life/AD&D Insurance
  • Short-Term & Long-Term Disability
  • Employee Assistance Program (EAP)
  • Training and Educational Assistance
  • Paid Time Off (PTO)
  • 11 Federal holidays
  • 401k plan with up to a 6% match (100% immediate vesting)
Equal Opportunity Employer

Sunayu, LLC is an Equal Opportunity/Affirmative Action Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, gender expression, national origin, age, protected veteran status, disability status, marital status, genetic information, medical condition, or any other characteristic protected by law.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Systems & Platform Engineer - TS/SCI
Systems & Platform Engineer - TS/SCI

Sunayu Llc • Bethesda (MD), Northern (KY)

Hybrid
USD 140,000 - 190,000
Medical plans
Dental and Vision
401k plan with match
+4
Linux Server Systems Engineer - TS/SCI
Linux Server Systems Engineer - TS/SCI

Sunayu Llc • Bethesda (MD), Northern (KY)

Hybrid
USD 120,000 - 170,000
Medical Plan Options
Dental and Vision
FSA/HSA
+7
Senior Systems Engineer - TS/SCI
Senior Systems Engineer - TS/SCI

Sunayu Llc • Bethesda (MD), Northern (KY)

Hybrid
USD 140,000 - 180,000
Medical Plan Options
Dental and Vision
FSA / DCFSA / HSA
+7
MD-Software Engineer 2 - TS/SCI w/ Polygraph
MD-Software Engineer 2 - TS/SCI w/ Polygraph

Sunayu • Maryland

On-site
USD 120,000 - 180,000
Medical plan
Dental and Vision
401k with match
+2
Senior Software Engineer - TS/SCI
Senior Software Engineer - TS/SCI

Sunayu Llc • Bethesda (MD), Northern (KY)

Hybrid
USD 120,000 - 190,000
Medical plans
Dental and Vision
FSA/HSA
+7
MD DevOps Engineer 3-TS/SCI w/ Polygraph
MD DevOps Engineer 3-TS/SCI w/ Polygraph

Sunayu Llc • Annapolis (MD)

On-site
USD 140,000 - 190,000
Medical plan options
Dental & Vision
401k match
Senior Systems Engineer - HPC & GPU Infrastructure
Senior Systems Engineer - HPC & GPU Infrastructure

RPMGlobal • Bethesda (MD), Northern (KY)

Hybrid
USD 140,000 - 200,000
Systems Engineer - HPC & GPU Infrastructure - TS/SCI
Systems Engineer - HPC & GPU Infrastructure - TS/SCI

Xcelerate Solutions • Bethesda (MD)

On-site
USD 120,000 - 170,000
Senior DevOps Engineer - TS/SCI
Senior DevOps Engineer - TS/SCI

Sunayu Llc • Bethesda (MD), Northern (KY)

Hybrid
USD 150,000 - 210,000
3 Medical Plan Options
Dental and Vision
FSA, DCFSA, HSA
+7
MD-Systems Administrator 3 - TS/SCI w/ Polygraph
MD-Systems Administrator 3 - TS/SCI w/ Polygraph

Sunayu • Aurora (CO)

On-site
USD 90,000 - 130,000
Medical Plan Options
Dental & Vision
FSA/DCFSA/HSA
+7