Systems and Platform Engineer - TS/SCI

VMD Corp

Bethesda (MD)

On-site

USD 150,000 - 190,000

Full time

2 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Xcelerate Solutions in Bethesda, MD seeks a Systems and Platform Engineer with TS/SCI to design and optimize Kubernetes clusters for enterprise AI workloads. You will build secure, automated pipelines and maintain Linux systems across cloud and on-prem environments.

Ideal candidates will have 10+ years in platform engineering, strong Linux/Kubernetes expertise, and DoD security certification readiness. This role requires citizenship and a TS/SCI clearance, with potential for polygraph.

Qualifications

  • Bachelor's degree in a technical field or equivalent with extensive experience.
  • 10+ years of related experience for senior level.
  • DoD 8570 IAT Level II/III certifications or equivalent.
  • Strong Linux and Kubernetes administration skills.
  • Experience with IaC tools (Terraform, Ansible, etc.).

Responsibilities

  • Design, configure, and maintain enterprise Kubernetes clusters.
  • Develop and manage IaC pipelines (Terraform, Salt, Ansible, Bash, Python).
  • Collaborate with dev teams to build secure, automated CI/CD pipelines.
  • Troubleshoot complex issues across cloud, network, and platform layers.
  • Maintain documentation and support ATO processes and federal security standards.

Skills

Kubernetes
Docker
Linux
Terraform
Ansible
CI/CD
Python
Argo
Airflow
Kubeflow
RESTful APIs
IAT II
IAT III
GICSP
Security+ CE

Education

Bachelor's degree
Master's degree

Tools

NVIDIA Base Command Manager
NVIDIA AI Enterprise
DGX
HGX
Open MPI

Job description

Systems and Platform Engineer – TS/SCI Xcelerate Solutions is looking for a highly skilled platform engineer with deep expertise in operating systems, hardware, GPU, and high-speed networking. In this role, you will design, develop, and optimize Kubernetes clusters that power enterprise AI for the mission customers.

Location:

Bethesda, MD

Security Clearance:

TS/SCI

Responsibilities:
  • Kubernetes Cluster Engineering: Design, configure, and maintain enterprise Kubernetes platforms. Collaborate with a multidisciplinary team to define and optimize Kubernetes architecture, ensuring they meet performance, efficiency, and feature requirements.
  • Infrastructure as Code: Develop and manage Infrastructure as Code (IaC) using tools such as Terraform, Salt, Ansible, Bash, Python or similar frameworks.
  • Collaborate with development teams to design and implement secure, automated, and repeatable pipelines (e.g., GitLab CI/CD).
  • Troubleshot complex systems issues across cloud, network, and platform layers.
  • Compliance & Documentation: Maintain technical documentation, architectural specifications, and Linux best practices. Support ATO (Authority to Operate) and ensure compliance with federal security standards.
Minimum Requirements:
  • Due to the nature of the government contracts we support, US Citizenship is required.
  • TS/SCI with CI Poly is required for position or a TS/SCI and willingness to obtain a Poly.
  • Requires a Bachelor’s degree and 10+ years of relevant experience, or Masters degree with 8+ years of experience. Additional years of experience may be considered in lieu of a degree
  • 5+ years in Platform Engineering or System Engineering experience.
  • Strong expertise with Linux distributions. (RHEL, Ubuntu, Oracle Linux, and Rocky).
  • Experience administering Kubernetes clusters, including deploying, scaling, and maintaining containerized workloads.
  • Hands-on experience creating, managing, and troubleshooting Docker containers and container images throughout the software development lifecycle.
  • Experience with Kubernetes cluster management and AI/ML workflow orchestration (Argo, Airflow, and Kubeflow).
  • Strong track record with consuming, and troubleshooting RESTful APIs for platform integration and automation.
  • Excellent problem-solving skills and the ability to collaborate within a team.
  • Candidate must, at a minimum, meet DoD 8570.11- IAT Level II certification requirements (currently Security+ CE, CCNA-Security, GICSP, GSEC, or SSCP along with an appropriate computing environment (CE) certification). An IAT Level III certification would also be acceptable (CASP+, CCNP Security, CISA, CISSP, GCED, GCIH, CCSP).
Preferred Qualifications:
  • Experience in managing NVIDIA GPU data center platforms. (DGX, HGX, H200, H100, 200, B300, L40S).
  • Experience with NVIDIA enterprise tools such as Base Command Manager, Run:AI, Nvidia AI Enterprise.
  • Knowledge of enterprise server components (storage/network controllers, HBA, SSDs).
  • Familiarity with GPU virtualization and cloud computing.
  • Experience developing and deploying infrastructure in AWS.
  • Knowledge of distributed resource scheduling systems. (Slurm , LSF, Open MPI,etc.)
About Xcelerate Solutions:

Founded in 2009 and headquartered in McLean, VA, Xcelerate Solutions (www.xceleratesolutions.com) is one of America's fastest-growing companies. Xcelerate’s culture is defined by our diversified workforce of dynamic and versatile professionals, supported with growth and development opportunities that contribute to individual and company growth. This strong commitment to our employees has been recognized by our inclusion on the Washington Business Journal’s “50 Best Places to Work” list as well as being a “Great Place to Work” certified company with a 4.6 star, and a 99% CEO approval Glassdoor rating. Come find out why Xcelerate Solutions is one of the DC Metro top employers!

Xcelerate Solutions is an Equal Employment Opportunity/Affirmative Action Employer. We evaluate qualified applicants without regard to race, color, national origin, religion, age, equal pay, disability, veteran status, sex, sexual orientation, gender identity, genetic information, or expression of another protected characteristic. As part of this commitment to the full inclusion of all qualified individuals, Xcelerate provides reasonable accommodations if needed because of an applicant's or an employee's disability.

Pay Transparency Notice: Xcelerate Solutions will not discharge or in any other manner discriminate against employees or applicants because they have inquired about, discussed, or disclosed their own pay or the pay of another employee or applicant.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Systems and Platform Engineer - TS/SCI
Systems and Platform Engineer - TS/SCI

Xcelerate Solutions • Bethesda (MD)

Hybrid
USD 140,000 - 190,000
Competitive salary
Health benefits
Systems Engineer - HPC & GPU Infrastructure - TS/SCI
Systems Engineer - HPC & GPU Infrastructure - TS/SCI

VMD Corp • Bethesda (MD)

On-site
USD 140,000 - 180,000
Systems Engineer - HPC & GPU Infrastructure - TS/SCI
Systems Engineer - HPC & GPU Infrastructure - TS/SCI

Xcelerate Solutions • Bethesda (MD)

On-site
USD 120,000 - 170,000
Senior Systems and Platform Engineer
Senior Systems and Platform Engineer

Maxisiq • Bethesda (AR)

On-site
USD 180,000 - 260,000
Senior Systems and Platform Engineer
Senior Systems and Platform Engineer

MAXISIQ, Inc. • Bethesda (MD)

On-site
USD 170,000 - 210,000
DevOps Engineer - TS/SCI
DevOps Engineer - TS/SCI

VMD Corp • Bethesda (MD)

On-site
USD 120,000 - 190,000
Systems Engineers SME – TS
Systems Engineers SME – TS

VMD Corp • Quantico (VA)

Hybrid
USD 110,000 - 165,000
Senior Systems & Platform Engineer – Kubernetes & AI Infra
Senior Systems & Platform Engineer – Kubernetes & AI Infra

VMD Corp • Bethesda (MD)

On-site
USD 150,000 - 190,000
Senior Systems Engineer - TS/SCI
Senior Systems Engineer - TS/SCI

VMD Corp • Bethesda (MD)

Hybrid
USD 140,000 - 190,000
Flexible schedule
DevOps Engineer - TS/SCI
DevOps Engineer - TS/SCI

Xcelerate-Solutions-5 • Bethesda (MD)

Hybrid
USD 120,000 - 170,000