Linux Systems Engineer - TS/SCI

Jobless

Charlottesville, Northern (VA, KY)

Hybrid

USD 100,000 - 175,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Plus3 IT Systems is seeking a Linux Systems Engineer (HPC) in Charlottesville, VA to deploy and sustain multi-node Linux HPC clusters, optimize schedulers, and support GPU-enabled workloads. This onsite role requires TS/SCI clearance and strong Linux administration skills.

You will automate cluster operations with Bash/Python, work with SLURM-like scheduling, and collaborate with cross-functional teams to deliver high-performance computing solutions in secure DoD or research environments.

Qualifications

  • Active TS/SCI clearance required.
  • Active or ability to obtain DoD 8140 IAT Level II certification (e.g., Security+)
  • Bachelor's degree in Computer Science, Information Technology, Engineering or similar; an additional 4 years of experience will be considered in lieu of degree.
  • Minimum 6 years of Linux systems administration experience in enterprise, research computing, or distributed compute environments.
  • Demonstrated experience supporting HPC cluster platforms or distributed compute environments at scale.
  • Hands-on experience with workload schedulers, queue management and job troubleshooting.
  • Proficiency in Linux command-line administration, including server configuration and system troubleshooting in distributed environments.
  • Ability to work onsite (hybrid and remote options not available).

Responsibilities

  • Deploy, configure, and sustain multi-node Linux HPC cluster environments, including node provisioning, integration, and day-to-day operational support.
  • Administer and troubleshoot workload scheduling platforms, including queue configuration, job submission workflows, and scheduler performance optimization.
  • Support distributed and containerized compute workloads leveraging parallel frameworks and container technologies within the cluster environment.
  • Monitor and analyze performance across compute, storage, and network layers including high-performance networking technologies and drive resolution of cluster communication issues.
  • Support GPU-enabled compute environments and CUDA-based workloads, ensuring proper resource allocation and integration with the scheduling platform.
  • Develop and maintain operational scripts and automation tooling (Bash, Python) to improve cluster administration efficiency and reduce manual toil.

Skills

Linux systems administration
HPC cluster
Workload schedulers
GPU/CUDA
Linux CLI
Automation scripting
Distributed compute

Education

Bachelor's degree in Computer Science, Information Technology, Engineering or similar

Tools

Ansible
Puppet
CUDA

Job description

Job Title: Linux Systems Engineer (HPC)

Location: Charlottesville, VA (Onsite)

Clearance Requirement: Active TS/SCI

Compensation Range: $100K - $175K

Join Plus3 IT Systems! We are at the forefront of cloud computing, providing comprehensive and cutting‑edge solutions across a wide array of critical domains. But we don’t stop at implementing technology; we are trusted advisors, delivering expert analysis to fully understand our clients unique challenges and objectives. Our passion is all about empowering our customers to reach their strategic goals. This mission is fueled by our exceptional teams of innovative technology practitioners, who bring deep technical skills and an unwavering commitment to excellence. At Plus3 IT, we foster agile, collaborative processes, working hand‑in‑hand with our clients to ensure transparency, flexibility, and ultimately, their success in the cloud.

What You’ll Do
  • Deploy, configure, and sustain multi-node Linux HPC cluster environments, including node provisioning, integration, and day‑to‑day operational support.
  • Administer and troubleshoot workload scheduling platforms, including queue configuration, job submission workflows, and scheduler performance optimization.
  • Support distributed and containerized compute workloads leveraging parallel frameworks and container technologies within the cluster environment.
  • Monitor and analyze performance across compute, storage, and network layers including high‑performance networking technologies and drive resolution of cluster communication issues.
  • Support GPU‑enabled compute environments and CUDA‑based workloads, ensuring proper resource allocation and integration with the scheduling platform.
  • Develop and maintain operational scripts and automation tooling (Bash, Python) to improve cluster administration efficiency and reduce manual toil.
QUALIFICATIONS
Clearance and Certification Requirements
  • Active TS/SCI clearance required
  • Active or ability to obtain DoD 8140 IAT Level II certification (e.g., Security+)
  • Bachelor's degree in Computer Science, Information Technology, Engineering or similar; an additional 4 years of experience will be considered in lieu of degree
Minimum Requirements
  • Minimum 6 years of Linux systems administration experience in enterprise, research computing, or distributed compute environments
  • Demonstrated experience supporting HPC cluster platforms or distributed compute environments at scale
  • Hands‑on experience with workload schedulers, queue management and job troubleshooting
  • Proficiency in Linux command‑line administration, including server configuration and system troubleshooting in distributed environments
  • Ability to work onsite (hybrid and remote options not available)
Preferred Skills
  • Direct administration experience with multi‑node HPC cluster environments, including provisioning workflows and lifecycle management
  • Experience with parallel or distributed file systems in a cluster context
  • Familiarity supporting MPI or OpenMP parallel workloads and understanding of how they interact with schedulers and underlying hardware
  • Experience supporting GPU‑enabled compute environments and CUDA‑based workloads within an HPC cluster
  • Proficiency with configuration management tools such as Ansible or Puppet applied to cluster‑scale infrastructure
  • Prior experience supporting systems within DoD, IC, or research laboratory environments
What You’ll Love About Plus3
  • Agile & Collaborative: Work in a highly collaborative environment where your ideas are heard, and you can quickly adapt and innovate.
  • Invested in You: Your growth is our priority. We offer a culture of continuous learning and support designed to keep your skills sharp and help you advance your career.
  • Culture That Connects: Be part of a supportive team that values collaboration, quality, and a sense of belonging from the moment you join.
  • Cutting-Edge Work: Engage with advanced cloud and AI solutions that are at the forefront of technology.
What You’ll Bring to Plus3
  • A passion for working on cutting-edge, high‑profile projects and a drive for delivering solutions
  • An insatiable curiosity: you ask why, proactively exploring and sharing ideas
  • A love for learning new technologies and sharing them with your team
  • A keen interest in utilizing Cloud-based and Open Source tools for problem‑solving
  • A strong self‑starter that flourishes in a team environment; and love the ability to work on multiple projects simultaneously
  • Strong verbal and written communication skills for effective collaboration with customers, vendors, and engineering teams to solve complex business problems
EEO

At Plus3 IT Systems, we believe our success is driven by the contributions of every employee. As an Equal Opportunity Employer, we make employment decisions based solely on an individual’s qualifications, skills and merit, and without regard to race, color, religion, creed, sex, sexual orientation, gender identity, marital status, national origin, age, veteran status, disability, or any other protected class. We are committed to providing access, equal opportunity, and reasonable accommodation for individuals with disabilities in employment, its services, programs, and activities. To request reasonable accommodation, contact hr@plus3it.com.

*Compensation: Actual compensation will be determined based on the candidate's specific blend of experience, qualifications, and performance location. Salaries or compensation for part‑time roles may be prorated or adjusted based upon the agreed upon number of hours to be regularly worked.

Benefits

Plus3 IT Systems offers eligible employees a variety of benefits including Employer-paid health, dental, vision, life, short/long term disability, contribution to health savings account, 401(k) matching, parental leave, flexible paid vacation, and company paid holidays. A full listing of available benefits can be viewed at https://www.plus3it.com/work-here.

Application Duration

The anticipated duration for applications is 30 days from the date of posting. Please note that this timeframe is subject to change—it may be extended or shortened—based on business requirements and the availability of qualified applicants.

Position

This is a full-time position and work days are Monday through Friday. Specific work hours are established with Team Lead. Work location may vary, including onsite requirements which are outlined in the requirements and confirmed with Team Lead.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

TS/SCI Linux HPC Engineer (Onsite)
TS/SCI Linux HPC Engineer (Onsite)

Jobless • Charlottesville (VA), Northern (KY)

Hybrid
USD 100,000 - 175,000
Principal Federal HPC Technical Consultant, (Clearance Preferred TS/SCI with Poly) MD or Utah
Principal Federal HPC Technical Consultant, (Clearance Preferred TS/SCI with Poly) MD or Utah

Hewlett Packard Enterprise • Trappe (MD)

On-site
USD 119,000 - 275,000
Health & Wellbeing
Career development
Inclusion
Linux Systems Engineer (TS/SCI)
Linux Systems Engineer (TS/SCI)

Vantor Inc. • Herndon (VA)

On-site
USD 95,000 - 140,000
Robust 401(k) with company match
Mental health resources
Student loan repayment assistance
+2
HPC Administrator
HPC Administrator

Jahnel Group • United States

On-site
USD 120,000 - 180,000
HPC Platform Engineer
HPC Platform Engineer

Shield Consulting Solutions, Inc. • Maryland

On-site
USD 205,000 - 215,000
25 days PTO
11 paid holidays
Employer-paid healthcare for employees
+1
Senior HPC DevOps Engineer | TS/SCI w/ MD POLY Security Clearance required
Senior HPC DevOps Engineer | TS/SCI w/ MD POLY Security Clearance required

Capstone Technology Partners • College Park (MD)

On-site
USD 222,000 - 257,000
Four weeks paid time off
Eleven paid holidays
401k with employer contributions and 3
+7
Linux Systems Engineer (TS/SCI)
Linux Systems Engineer (TS/SCI)

maxar • Herndon (VA)

On-site
USD 95,000 - 140,000
401(k) with company match
Mental health resources
Student loan repayment assistance
+2
Linux Systems Engineer (TS/SCI)
Linux Systems Engineer (TS/SCI)

Vantor • Herndon (VA)

On-site
USD 95,000 - 140,000
401(k) matching
Mental health resources
Student loan repayment assistance
+2
Systems Administrator - HPC Linux (TS/SCI Clearance Required)
Systems Administrator - HPC Linux (TS/SCI Clearance Required)

North Point Technology • Herndon (VA)

On-site
USD 120,000 - 150,000
Systems Administrator - HPC Linux (TS/SCI Clearance Required)
Systems Administrator - HPC Linux (TS/SCI Clearance Required)

North Point Technology • Springfield (MO)

On-site
USD 90,000 - 130,000