Senior Systems and Platform Engineer

MAXISIQ, Inc.

Bethesda (MD)

On-site

USD 170,000 - 210,000

Full time

2 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Vexterra Group seeks a platform engineer with deep Linux, GPU, and networking expertise to design, develop, and optimize Kubernetes clusters powering enterprise AI for mission customers. The role is based in Bethesda, MD at the Intelligence Community Campus.

You will implement IaC with Terraform, Salt, Ansible, Bash and Python, build secure CI/CD pipelines, troubleshoot across cloud, network, and platform layers, and ensure compliance with federal security standards and ATO.

Qualifications

  • Bachelor’s degree with 10+ years of relevant experience or Master’s with 8+ years.
  • 5+ years in Platform Engineering or System Engineering experience.
  • Strong Linux and Kubernetes cluster management experience.
  • DoD 8570.11-IAT Level II certification or higher required.
  • Experience with IaC and automation pipelines.

Responsibilities

  • Kubernetes Cluster Engineering: Design, configure, and maintain enterprise Kubernetes platforms.
  • Infrastructure as Code: Develop IaC using Terraform, Salt, Ansible, Bash, Python.
  • Design and implement secure, automated pipelines (GitLab CI/CD).
  • Troubleshoot complex systems across cloud, network, and platform layers.
  • Maintain documentation and support ATO/compliance with federal standards.

Skills

Kubernetes cluster administration
Linux administration
Containerization (Docker)
RESTful API integration
Problem solving
Team collaboration
IAT certification knowledge

Education

Bachelor’s degree + 10+ years experience, or Master’s + 8+ years

Tools

Terraform
Salt
Ansible
Bash
Python
GitLab CI/CD
Argo
Airflow
Kubeflow
Docker
Kubernetes tooling
Base Command Manager
Run:AI
NVIDIA AI Enterprise
Slurm
LSF
Open MPI

Job description

  • Compensation: USD 170,000 - USD 210,000 - yearly
Job Description

Our Partner company, Vexterra Group, is looking for a highly skilled platform engineer with deep expertise in operating systems, hardware, GPU, and high-speed networking. In this role, you will design, develop, and optimize Kubernetes clusters that power enterprise AI for the mission customers. The work location is in Bethesda at the Intelligence Community Campus. Primary Responsibilities Kubernetes Cluster Engineering: Design, configure, and maintain enterprise Kubernetes platforms. Collaborate with a multidisciplinary team to define and optimize Kubernetes architecture, ensuring they meet performance, efficiency, and feature requirements. Infrastructure as Code: Develop and manage Infrastructure as Code (IaC) using tools such as Terraform, Salt, Ansible, Bash, Python or similar frameworks. Collaborate with development teams to design and implement secure, automated, and repeatable pipelines (e.g., GitLab CI/CD). Troubleshot complex systems issues across cloud, network, and platform layers. Compliance & Documentation: Maintain technical documentation, architectural specifications, and Linux best practices. Support ATO (Authority to Operate) and ensure compliance with federal security standards.

Qualifications
Basic Qualifications
  • Requires a Bachelor’s degree and 10+ years of relevant experience, or Masters degree with 8+ years of experience. Additional years of experience may be considered in lieu of a degree
  • 5+ years in Platform Engineering or System Engineering experience.
  • Strong expertise with Linux distributions. (RHEL, Ubuntu, Oracle Linux, and Rocky).
  • Experience administering Kubernetes clusters, including deploying, scaling, and maintaining containerized workloads.
  • Hands-on experience creating, managing, and troubleshooting Docker containers and container images throughout the software development lifecycle.
  • Experience with Kubernetes cluster management and AI/ML workflow orchestration (Argo, Airflow, and Kubeflow).
  • Strong track record with consuming, and troubleshooting RESTful APIs for platform integration and automation.
  • Excellent problem-solving skills and the ability to collaborate within a team.
  • Candidate must, at a minimum, meet DoD 8570.11- IAT Level II certification requirements (currently Security+ CE, CCNA-Security, GICSP, GSEC, or SSCP along with an appropriate computing environment (CE) certification). An IAT Level III certification would also be acceptable (CASP+, CCNP Security, CISA, CISSP, GCED, GCIH, CCSP).
Security Clearance

Security Clearance: TS/SCI with CI Poly is required for position or a TS/SCI and willingness to obtain a Poly.

Preferred Qualifications
  • Experience in managing NVIDIA GPU data center platforms. (DGX, HGX, H200, H100, 200, B300, L40S).
  • Experience with NVIDIA enterprise tools such as Base Command Manager, Run:AI, Nvidia AI Enterprise.
  • Knowledge of enterprise server components (storage/network controllers, HBA, SSDs).
  • Familiarity with GPU virtualization and cloud computing.
  • Experience developing and deploying infrastructure in AWS.
  • Knowledge of distributed resource scheduling systems. (Slurm , LSF, Open MPI,etc)
Additional Information

All your information will be kept confidential according to EEO guidelines.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Systems and Platform Engineer - TS/SCI
Systems and Platform Engineer - TS/SCI

Xcelerate-Solutions-5 • Bethesda (MD)

On-site
USD 180,000 - 240,000
Systems and Platform Engineer
Systems and Platform Engineer

Quest International • Bethesda (MD)

Hybrid
USD 107,900 - 195,050
On-site at ICC Bethesda
Competitive pay and benefits
Health and wellness programs
Platform Engineer (GPU)
Platform Engineer (GPU)

Vero • United States

On-site
USD 136,000 - 160,000
Medical, dental, and vision insurance
Equity Scheme
401(k) with employer match
+3
Senior Kubernetes Platform Engineer
Senior Kubernetes Platform Engineer

The Brixton Group • Dallas (TX)

Hybrid
USD 138,000 - 165,000
Relocation assistance
Platform Systems Engineer
Platform Systems Engineer

Huntington Ingalls Industries • Springfield (VA)

On-site
USD 83,000 - 195,000
GPU Systems Engineer 3
GPU Systems Engineer 3

RPMGlobal • Bethesda (MD)

On-site
USD 120,000 - 180,000
GPU Systems Engineer 4
GPU Systems Engineer 4

RPMGlobal • Bethesda (MD), Northern (KY)

Hybrid
USD 180,000 - 280,000
AI Kernel / Cluster Engineer
AI Kernel / Cluster Engineer

Blue Signal Search • Santa Clara (CA)

On-site
USD 150,000 - 210,000
Platform Systems Engineer
Platform Systems Engineer

Mission Technologies, a division of HII • Springfield (VA)

On-site
USD 83,000 - 195,000
Senior Systems Software Engineer, Kubernetes Node Lifecycle - DGX Cloud
Senior Systems Software Engineer, Kubernetes Node Lifecycle - DGX Cloud

NVIDIA • Seattle (WA)

On-site
USD 184,000 - 288,000
Equity
Benefits