Systems and Platform Engineer - TS/SCI

Xcelerate-Solutions-5

Bethesda (MD)

On-site

USD 180,000 - 240,000

Full time

19 hours ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Xcelerate Solutions in Bethesda, MD seeks a Platform Engineer to design and optimize Kubernetes clusters powering enterprise AI for government customers. You will lead IaC efforts, automation pipelines, and cross-team collaboration to ensure secure, scalable deployments.

The role requires TS/SCI with CI Poly or willingness to obtain a poly, US citizenship, and a Bachelor’s or higher degree with 8+ years of related experience. On-site in Bethesda, MD.

Qualifications

  • Extensive experience designing and operating Kubernetes platforms for enterprise AI.
  • Strong Linux administration experience (RHEL/Ubuntu/Oracle Linux/Rocky).
  • Experience with IaC like Terraform, Salt, Ansible.

Responsibilities

  • Kubernetes Cluster Engineering: design, configure, and maintain enterprise Kubernetes platforms.
  • Infrastructure as Code: develop and manage IaC using Terraform, Salt, Ansible, Bash or Python.
  • Collaborate to design automated CI/CD pipelines (GitLab CI/CD).
  • Troubleshoot complex systems across cloud, network, and platform layers.
  • Maintain documentation and support ATO and federal security standards.

Skills

Kubernetes
Linux
Docker
CI/CD
REST APIs
Problem solving
Teamwork

Education

Bachelor’s degree
Master’s degree

Tools

Terraform
Salt
Ansible
Bash
Python
Argo
Kubeflow
Airflow

Job description

Xcelerate Solutions is looking for a highly skilled platform engineer with deep expertise in operating systems, hardware, GPU, and high-speed networking.In this role, you will design, develop, and optimize Kubernetes clusters that power enterprise AI for the mission customers.

Location:Bethesda, MD

Security Clearance:TS/SCI

Responsibilities:
  • Kubernetes Cluster Engineering: Design, configure, and maintain enterprise Kubernetes platforms. Collaborate with a multidisciplinary team to define and optimize Kubernetes architecture, ensuring they meet performance, efficiency, and feature requirements.
  • Infrastructure as Code: Develop and manage Infrastructure as Code (IaC) using tools such as Terraform, Salt, Ansible, Bash, Python or similar frameworks.
  • Collaborate with development teams to design and implement secure, automated, and repeatable pipelines (e.g., GitLab CI/CD).
  • Troubleshot complex systems issues across cloud, network, and platform layers.
  • Compliance & Documentation: Maintain technical documentation, architectural specifications, and Linux best practices. Support ATO (Authority to Operate) and ensure compliance with federal security standards.
  • Due to the nature of the government contracts we support, US Citizenship is required.
  • TS/SCI with CI Poly is required for position or a TS/SCI and willingness to obtain a Poly.
  • Requires a Bachelor’s degree and 10+ years of relevant experience, or Masters degree with 8+ years of experience.Additional years of experience may be considered in lieu of a degree
  • 5+ years in Platform Engineering or System Engineering experience.
  • Strong expertise with Linux distributions. (RHEL, Ubuntu, Oracle Linux, and Rocky).
  • Experience administering Kubernetes clusters, including deploying, scaling, and maintaining containerized workloads.
  • Hands-on experience creating, managing, and troubleshooting Docker containers and container images throughout the software development lifecycle.
  • Experience with Kubernetes cluster management and AI/ML workflow orchestration (Argo, Airflow, and Kubeflow).
  • Strong track record with consuming, and troubleshooting RESTful APIs for platform integration and automation.
  • Excellent problem-solving skills and the ability to collaborate within a team.
  • Candidate must, at a minimum, meet DoD 8570.11- IAT Level II certification requirements (currently Security+ CE, CCNA-Security, GICSP, GSEC, or SSCP along with an appropriate computing environment (CE) certification). An IAT Level III certification would also be acceptable (CASP+, CCNP Security, CISA, CISSP, GCED, GCIH, CCSP).
Preferred Qualifications:
  • Experience in managing NVIDIA GPU data center platforms. (DGX, HGX, H200, H100, 200, B300, L40S).
  • Experience with NVIDIA enterprise tools such as Base Command Manager, Run:AI, Nvidia AI Enterprise.
  • Knowledge of enterprise server components (storage/network controllers, HBA, SSDs).
  • Familiarity with GPU virtualization and cloud computing.
  • Experience developing and deploying infrastructure in AWS.
  • Knowledge of distributed resource scheduling systems. (Slurm , LSF, Open MPI,etc.)
About Xcelerate Solutions:

Founded in 2009 and headquartered in McLean, VA, Xcelerate Solutions (www.xceleratesolutions.com) is one of America's fastest-growing companies. Xcelerate’ culture is defined by our diversified workforce of dynamic and versatile professionals, supported with growth and development opportunities that contribute to individual and company growth. This strong commitment to our employees has been recognized by our inclusion on the Washington Business Journal's "50 Best Places to Work" list as well as being a "Great Place to Work" certified company with a 4.6 star, and a 99% CEO approval Glassdoor rating. Come find out why Xcelerate Solutions is one of the DC Metro top employers!

Xcelerate Solutions is an Equal Employment Opportunity/Affirmative Action Employer. We evaluate qualified applicants without regard to race, color, national origin, religion, age, equal pay, disability, veteran status, sex, sexual orientation, gender identity, genetic information, or expression of another protected characteristic. As part of this commitment to the full inclusion of all qualified individuals, Xcelerate provides reasonable accommodations if needed because of an applicant's or an employee's disability.

Pay Transparency Notice: Xcelerate Solutions will not discharge or in any other manner discriminate against employees or applicants because they have inquired about, discussed, or disclosed their own pay or the pay of another employee or applicant.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Systems Engineer - TS/SCI
Systems Engineer - TS/SCI

Xcelerate-Solutions-5 • Bethesda (MD)

On-site
USD 120,000 - 160,000
Systems Engineer - HPC & GPU Infrastructure - TS/SCI
Systems Engineer - HPC & GPU Infrastructure - TS/SCI

VMD Corp • Bethesda (MD)

On-site
USD 140,000 - 180,000
Systems Engineer - HPC & GPU Infrastructure - TS/SCI
Systems Engineer - HPC & GPU Infrastructure - TS/SCI

Xcelerate Solutions • Bethesda (MD)

On-site
USD 120,000 - 170,000
Senior Systems and Platform Engineer
Senior Systems and Platform Engineer

MAXISIQ, Inc. • Bethesda (MD)

On-site
USD 170,000 - 210,000
DevOps Engineer - TS/SCI
DevOps Engineer - TS/SCI

Xcelerate-Solutions-5 • Bethesda (MD)

Hybrid
USD 120,000 - 170,000
Sr. DevOps Engineer - TS/SCI
Sr. DevOps Engineer - TS/SCI

Xcelerate-Solutions-5 • Bethesda (MD)

Hybrid
USD 180,000 - 240,000
Senior Software Engineer - TS/SCI
Senior Software Engineer - TS/SCI

Xcelerate-Solutions-5 • Bethesda (MD)

Hybrid
USD 150,000 - 190,000
Sr. Systems Administrator – Secret
Sr. Systems Administrator – Secret

Xcelerate Solutions • McLean (VA)

Hybrid
USD 95,000 - 145,000
Sr. Systems Administrator – Secret
Sr. Systems Administrator – Secret

Xcelerate-Solutions-5 • McLean (VA)

Hybrid
USD 100,000 - 150,000
Remote work option
Senior Platform Engineer - Kubernetes & AI Infra
Senior Platform Engineer - Kubernetes & AI Infra

Xcelerate-Solutions-5 • Bethesda (MD)

On-site
USD 180,000 - 240,000