Platform Engineer IV

Capgemini

St. Louis (MO)

On-site

USD 53,726 - 84,033

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Tundra Technical Solutions is seeking an experienced Infrastructure and Data Lake Engineer to design and manage AWS infrastructure for the IIA Data Lake, including S3 storage, Glue catalog, Athena queries, and EMR compute clusters. You will ensure secure cross-account connectivity and robust IAM policies.

You will build IaC with Terraform/CloudFormation, manage CI/CD pipelines with GitLab CI/CD, and collaborate with platform teams to deploy AI agent runtimes and graph databases like Neptune.

Qualifications

  • Bachelor's degree in Computer Science or related field; 5+ years infra engineering experience.
  • Experience with cross-account AWS architectures and IAM.
  • Strong scripting in Python or a similar language.

Responsibilities

  • Design and manage AWS infrastructure for Data Lake components (S3, Glue, Athena, EMR).
  • Implement IaC using Terraform or CloudFormation.
  • Operate CI/CD pipelines and agent deployments with GitLab CI/CD.
  • Maintain production stability with monitoring and incident response.

Skills

AWS
CI/CD
Python
Terraform
Git
Networking basics

Education

Bachelor's degree in Computer Science or related field

Tools

Docker
GitLab
Artifactory
Prometheus

Job description

Infrastructure and Data Lake:

Design and manage AWS infrastructure for the IIA Data Lake including S3 storage, Glue data catalog, Athena query engine, and EMR compute clusters.

Manage cross-account connectivity, VPC networking, security groups, and IAM roles/policies to enable secure data flow between IIA, upstream data providers, and downstream consumers.

Build and maintain infrastructure for AI agent runtime environments, including compute resources for LangGraph agents deployed via LangSmith Deployments.

Support deployment and operation of AWS Neptune for the network topology graph (digital twin), including capacity planning, schema design support, and performance tuning.

Implement and manage infrastructure-as-code (Terraform, CloudFormation) for repeatable, auditable environment provisioning.

Manage IAM access key rotations, secrets management (AWS Secrets Manager, Delinea), and security compliance for on-premises and cloud integrations (e.g., Splunk Edge Processor).

CI/CD and Agent Deployments:

Build and maintain CI/CD pipelines for AI agent deployments using GitLab CI/CD, Docker, and Artifactory.

Manage container lifecycle for agents deployed via LangSmith Deployments, including image builds, versioning, and rollback procedures.

Automate deployment workflows to enable rapid, reliable promotion of agents from development through production.

Coordinate with SpecGPT platform team on AI Gateway integration, cross-account deployment, and connectivity requirements.

Application Development and Tooling:

Develop and maintain utilitarian application code: scripts, CLIs, small services, and automation utilities (primarily Python) that support data ingestion, deployment, environment provisioning, and operational workflows.

Write integration code and glue services that connect IIA systems with upstream data providers, downstream consumers, and external platforms.

Production Operations

Ensure production environment stability through monitoring, alerting, and incident response. Maintain SLAs for data pipeline availability and agent uptime.

Implement production monitoring and alerting for deployed agents (health checks, error rates, latency, resource utilization).

Coordinate with upstream data teams and platform teams (SpecGPT, Splunk, Public Cloud) on connectivity, firewall requests, and integration requirements.

Support data engineering team with infrastructure needs for new data source onboarding (storage provisioning, access controls, pipeline compute).

Perform other duties as required.

Required Qualifications:

Skills/Abilities and Knowledge:

Ability to read, write, speak and understand English

Strong communication skills with ability to explain infrastructure decisions to non-infrastructure stakeholders

Expert-level experience with AWS services: EC2, S3, IAM, VPC, Glue, Athena, EMR, Secrets Manager, CloudWatch

Strong experience with infrastructure-as-code (Terraform preferred, CloudFormation acceptable)

Experience managing cross-account AWS architectures, VPC peering, PrivateLink, and transit gateway configurations

Experience with IAM policy design, least-privilege access patterns, and service account management

Experience with containerization (Docker) and container orchestration

Experience with CI/CD pipelines (GitLab CI preferred)

Proficiency with Linux-based operating systems and shell scripting

Experience with monitoring and alerting tools (CloudWatch, Prometheus, Grafana, or similar)

Understanding of networking fundamentals: DNS, CIDR, NAT, firewalls, security groups

Demonstrated ability to work across teams and coordinate with external platform owners on connectivity and access requirements

Proficiency in Python (or a comparable general-purpose language) for building automation, tooling, and applications

Solid software engineering fundamentals: Git-based workflows, code review, modular and reusable design, dependency management, and writing maintainable, documented code

Experience writing automated tests (unit/integration) for application and infrastructure code, and integrating those tests into CI/CD

Ability to write integration code against REST APIs and cloud SDKs (e.g., AWS SDK / boto3)

Preferred Qualifications:

Skills/Abilities and Knowledge:

Experience with graph databases (AWS Neptune, Neo4j) including deployment, scaling, and operational management

Experience with Apache Kafka or similar streaming platforms

Experience with Apache Spark (Scala preferred) for distributed data processing

Experience with Airflow or similar workflow orchestration platforms

Experience in the telecommunications industry or other large-scale network operations environments

Familiarity with AI/ML infrastructure requirements (model serving, GPU/CPU compute, artifact management via MLflow or similar)

Experience with Splunk integration, particularly Edge Processor and MCP connectivity

AWS certifications (Solutions Architect, DevOps Engineer, or similar)

Experience developing and operating small services or APIs (e.g., FastAPI/Flask) in a production environment

Education:

Bachelor's degree in Computer Science, Information Technology, Systems Engineering, or related field, or relevant experience

Related Experience:

Bachelor's degree: 5+ years of platform/infrastructure engineering experience

Master's degree: 3+ years of platform/infrastructure engineering experience

The pay range that the employer in good faith reasonably expects to pay for this position is $39.30/hour - $61.40/hour. Our benefits include medical, dental, vision and retirement benefits. Applications will be accepted on an ongoing basis.

Tundra Technical Solutions is among North America’s leading providers of Staffing and Consulting Services. Our success and our clients’ success are built on a foundation of service excellence. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. Qualified applicants with arrest or conviction records will be considered for employment in accordance with applicable law, including the Los Angeles County Fair Chance Ordinance for Employers and the California Fair Chance Act. Unincorporated LA County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: client provided property, including hardware (both of which may include data) entrusted to you from theft, loss or damage; return all portable client computer hardware in your possession (including the data contained therein) upon completion of the assignment, and; maintain the confidentiality of client proprietary, confidential, or non-public information. In addition, job duties require access to secure and protected client information technology systems and related data security obligations.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Cloud Platform Engineer (contract)
Senior Cloud Platform Engineer (contract)

Capgemini • Charlotte (NC)

On-site
Medical benefits
Dental benefits
Vision benefits
+1
Java AWS Developer
Java AWS Developer

Capgemini • Chicago (IL)

On-site
USD 37,000 - 58,000
Medical benefits
Dental benefits
Vision benefits
+1
Delivery Program Manager
Delivery Program Manager

Capgemini • Charlotte (NC)

On-site
USD 71,000 - 111,000
Java Full Stack Developer
Java Full Stack Developer

Capgemini • New York (NY)

On-site
USD 52,000 - 81,000
Medical, dental, vision benefits
Retirement benefits
Platform Engineer IV
Platform Engineer IV

Softcom Systems Inc • Denver (CO)

Hybrid
USD 120,000 - 180,000
Platform Engineer SOW229 Edge Compute
Platform Engineer SOW229 Edge Compute

Capgemini • Seattle (WA)

On-site
USD 51,000 - 80,000
Medical benefits
Dental benefits
Retirement benefits
GenAI Developer (contract)
GenAI Developer (contract)

Capgemini • Atlanta (GA)

Hybrid
Medical benefits
Dental benefits
Vision benefits
+1
Software Engineer
Software Engineer

Tundra Technical Solutions • North Reading (MA)

On-site
USD 91,000 - 104,000
optional medical
optional dental
optional vision
+2
Platform Engineer IV
Platform Engineer IV

Brooksource • Denver (CO)

Hybrid
401(k) match
Paid time off
Paid company holidays
+2
Principal - Platform Engineer
Principal - Platform Engineer

Capgemini • New Jersey

On-site
USD 70,050 - 109,464
Medical insurance
Dental insurance
Vision insurance
+1