AI Platform Engineer

PGTEK

McLean (VA)

On-site

USD 125,000 - 185,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Comprehensive PPO medical coverage
401(k) matching plan
Generous PTO and Holidays
Education Assistance Program

Job summary

PGTEK, located in McLean, VA, is looking for a skilled AI Platform Engineer to enhance its AI workload infrastructure. You will oversee the deployment, security, and performance of scalable Kubernetes environments, collaborating closely with engineering and operations teams.

The ideal candidate has significant experience in Kubernetes and cloud-native technologies, employing DevOps practices to ensure platform reliability. This full-time position offers a competitive salary and requires onsite work five days a week, along with benefits including medical coverage and educational assistance.

Qualifications

  • 5+ years of experience in cloud-native technologies and Kubernetes administration.
  • Proven expertise in CI/CD methodologies and automation tools.
  • Strong Linux/Unix administration skills.

Responsibilities

  • Design and manage scalable Kubernetes environments.
  • Implement automated deployment pipelines for containerized applications.
  • Ensure platform reliability and security.

Skills

Kubernetes administration
DevOps practices
Automation
Cloud-native technologies
Problem-solving skills
Collaboration skills

Education

Current IAM Level II certification

Tools

Terraform
Docker
Kubernetes
GitOps tools
CI/CD platforms
Monitoring tools

Job description

AI Platform Engineer

Location: McLean, VA (Onsite 5 Days per Week)
Employment Type: Full-Time
Salary Range: $125,000 - $185,000
Clearance Required: Active TS/SCI with Counterintelligence Polygraph
Certification Requirement: Current IAM Level II certification meeting DoD 8570 IAT requirements

Position Overview

We are seeking an experienced AI Platform Engineer to play a critical role in building, maintaining, securing, and optimizing the infrastructure that supports advanced Artificial Intelligence (AI) workloads. This individual will be responsible for designing and managing scalable Kubernetes environments, implementing automated deployment pipelines, and ensuring platform reliability, security, and performance.

The ideal candidate combines deep expertise in cloud-native technologies, Kubernetes administration, DevOps practices, and automation with strong problem‑solving and collaboration skills. This role will work closely with engineering, operations, and security teams to deliver highly available AI platform solutions in a mission‑critical environment.

Kubernetes & Platform Engineering
  • Design, deploy, secure, maintain, and upgrade highly available Kubernetes clusters across cloud and on‑premises environments.
  • Manage Kubernetes control plane components, worker nodes, and supporting infrastructure.
  • Implement and maintain containerized workloads using Docker and Kubernetes best practices.
  • Configure and manage Kubernetes resources including Pods, Deployments, StatefulSets, Services, Ingress, ConfigMaps, Secrets, Persistent Volumes, and Namespaces.
  • Support advanced networking configurations, including CNI plugins, network policies, service meshes, and DNS services.
Security & Compliance
  • Implement security best practices across Kubernetes environments.
  • Manage RBAC, admission controllers, vulnerability scanning, secret management, and network security controls.
  • Ensure platform compliance with government and organizational security requirements.
  • Support secure deployment practices and infrastructure hardening initiatives.
DevOps & Automation
  • Design, implement, and maintain CI/CD pipelines for containerized applications.
  • Utilize GitOps methodologies and tools to automate application deployment and platform management.
  • Develop infrastructure as code (IaC) solutions using Terraform, Pulumi, CloudFormation, or similar tools.
  • Create automation scripts and tooling using Python, Go, Bash, or related languages.
Monitoring, Performance & Reliability
  • Implement monitoring, logging, alerting, and observability solutions across platform environments.
  • Diagnose and resolve complex performance issues affecting Kubernetes clusters and applications.
  • Optimize resource utilization and platform scalability.
  • Support distributed tracing, centralized logging, and operational analytics initiatives.
  • Apply DevOps and Site Reliability Engineering (SRE) principles to improve platform resilience and operational excellence.
Collaboration & Leadership
  • Collaborate with development, operations, security, and infrastructure teams.
  • Lead technical initiatives and mentor junior engineers.
  • Drive continuous improvement efforts across platform engineering and deployment practices.
  • Communicate effectively with technical and non‑technical stakeholders.
Experience
  • Extensive experience designing, deploying, and managing Kubernetes environments (EKS, AKS, GKE, OpenShift, or self‑managed clusters).
  • Advanced knowledge of Docker and containerization technologies.
  • Strong understanding of Kubernetes networking, service meshes, and cluster architecture.
  • Expertise in Kubernetes security, access controls, and secret management.
  • Experience with CI/CD platforms such as Jenkins, GitLab CI/CD, GitHub Actions, Tekton, Argo Workflows, or similar.
  • Proficiency with Infrastructure as Code tools including Terraform, Pulumi, or CloudFormation.
  • Strong scripting and automation experience using Python, Go, Bash, or similar languages.
  • Experience with GitOps tools such as Argo CD.
  • Hands‑on experience with monitoring and observability platforms including Prometheus, Grafana, ELK/OpenSearch, Datadog, or Splunk.
  • Strong Linux/Unix administration background.
  • Solid understanding of networking concepts including TCP/IP, DNS, HTTP, and load balancing.
  • Expert‑level Git and version control experience.
Professional Skills
  • Exceptional troubleshooting and analytical problem‑solving abilities.
  • Strong verbal and written communication skills.
  • Ability to work effectively in cross‑functional teams.
  • Experience mentoring engineers and leading technical efforts.
  • Strong sense of ownership and accountability.
  • Adaptability and commitment to continuous learning.
Preferred Qualifications
  • Certified Kubernetes Administrator (CKA)
  • Certified Kubernetes Application Developer (CKAD)
  • Certified Kubernetes Security Specialist (CKS)
  • AWS Certified DevOps Engineer
  • Azure DevOps Engineer Expert
  • Experience developing Kubernetes Operators and Custom Resource Definitions (CRDs)
  • Experience building Internal Developer Platforms (IDPs)
  • Familiarity with testing methodologies including unit, integration, and end‑to‑end testing
Travel Requirements
  • Up to 20% travel as required for on‑site installations, maintenance, and troubleshooting activities at customer locations or data centers.

Our comprehensive benefits package for full‑time salaried employees is effective immediately upon the start date. Benefits include comprehensive PPO medical coverage with access to a Health Savings Account (HSA) option, a vision plan, and dental insurance with the base dental plan option paid for by PGTEK. Life Insurance, Short and Long‑Term disability, and Critical Illness insurance have premiums covered. Additionally, PGTEK offers a matching 401(k) plan and a discount on pet insurance through ASPCA Pet Insurance. An Employee Assistance Program is available at no cost to all employees. PGTEK offers a generous amount of PTO and Holidays, and an Education Assistance Program is available after 12 months of employment.

EOE, including disability/veterans

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Platform Engineer - TS/SCI with CI Poly
AI Platform Engineer - TS/SCI with CI Poly

PGTEK • Illinois

On-site
USD 125,000 - 185,000
Comprehensive medical coverage
401(k) matching
Education assistance program
+1
AI Platform Engineer - TS/SCI with CI Poly
AI Platform Engineer - TS/SCI with CI Poly

PGTEK • North Dakota

On-site
USD 125,000 - 185,000
Comprehensive benefits
PPO + HSA
Life & Disability insurance
+5
AI Platform Engineer (TS/SCI) – Kubernetes & DevOps Lead
AI Platform Engineer (TS/SCI) – Kubernetes & DevOps Lead

PGTEK • Illinois

On-site
USD 125,000 - 185,000
Comprehensive medical coverage
401(k) matching
Education assistance program
+1
Senior AI Platform Engineer: Kubernetes & DevOps Lead
Senior AI Platform Engineer: Kubernetes & DevOps Lead

PGTEK • McLean (VA)

On-site
USD 125,000 - 185,000
Comprehensive PPO medical coverage
401(k) matching plan
Generous PTO and Holidays
+1
AI Platform Engineer (TS/SCI w/ Poly) - Kubernetes & DevOps
AI Platform Engineer (TS/SCI w/ Poly) - Kubernetes & DevOps

PGTEK • North Dakota

On-site
USD 125,000 - 185,000
Comprehensive benefits
PPO + HSA
Life & Disability insurance
+5
Associate Cloud AI Engineer - TS/SCI clearance - Travel
Associate Cloud AI Engineer - TS/SCI clearance - Travel

PGTEK • Washington

On-site
USD 80,000 - 100,000
Health insurance
Vision plan
Dental insurance
+4
Associate Cloud AI Engineer - TS/SCI clearance - Travel
Associate Cloud AI Engineer - TS/SCI clearance - Travel

PGTEK • Washington

On-site
USD 80,000 - 100,000
401(k) matching
Education Assistance
Paid time off
+1
Associate Cloud AI Engineer - TS/SCI clearance - Travel
Associate Cloud AI Engineer - TS/SCI clearance - Travel

PGTEK • New York (NY)

On-site
USD 80,000 - 100,000
Health insurance (PPO)
Vision plan
Dental insurance
+3
Cloud AI Engineer – TS/SCI clearance – travel role
Cloud AI Engineer – TS/SCI clearance – travel role

PGTEK • Washington

Hybrid
USD 105,000 - 115,000
PPO medical with HSA
Dental & Vision
Life Insurance
+2
Cloud AI Engineer - TS/SCI clearance - travel role
Cloud AI Engineer - TS/SCI clearance - travel role

PGTEK • New York (NY)

On-site
USD 105,000 - 115,000
Health Insurance
401(k) Matching
Paid Time Off & Holidays