AWS Cloud Engineer III

Remote Jobs

United States

Remote

USD 120,000 - 180,000

Full time

45 hours ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Rackspace Technology is seeking an experienced L3 DevOps/Cloud Engineer to lead complex infrastructure and DevOps activities, driving automation and platform reliability.

The role requires hands-on expertise in AWS, Terraform, AWX/Ansible, ArgoCD, GitOps, Kubernetes, CI/CD and Disaster Recovery, with a track record of delivering scalable, secure platforms. You will own technical standards and mentor engineers while leading DR tests and production deployments.

Qualifications

  • 8–12 years of hands-on cloud, DevOps, or infrastructure experience.
  • Strong expertise in AWS, EKS, Kubernetes, Terraform, and GitOps.
  • Proven ability to lead complex production incidents and RCA.
  • Experience with automation via AWX/Ansible and ArgoCD.
  • Design and implement IaC standards and reusable modules.

Responsibilities

  • Serve as L3 escalation point for Cloud, DevOps, and deployment issues.
  • Lead major technical activities across prod and non-prod environments.
  • Own complex infra changes, upgrades, migrations, and platform improvements.
  • Drive Disaster Recovery planning, testing, and runbooks.
  • Design infrastructure with Terraform and develop reusable modules.
  • Review Terraform code and enforce coding standards.
  • Develop Ansible roles, playbooks, and AWX automation for config, patching, deployments.
  • Implement GitOps with ArgoCD; manage ArgoCD and drift.
  • Maintain Kubernetes workloads on Amazon EKS and Helm charts.
  • Improve CI/CD processes with Jenkins and other tools.
  • MentOR L1/L2 engineers and produce SOPs/runbooks.

Skills

AWS
Amazon EKS
Kubernetes
Terraform
Terragrunt
AWX
Ansible
ArgoCD
GitOps
Helm
Docker
Jenkins
CI/CD
Git
Linux
Windows
Bash / Python
SSL/TLS
Infrastructure Automation
Disaster Recovery
Monitoring & Logging
Production Troubleshooting
Root Cause Analysis

Job description

Role Overview

The L3 DevOps / Cloud Engineer will act as a senior technical resource responsible for leading complex infrastructure and DevOps activities, driving automation, defining technical standards, and improving platform reliability.

The role requires strong hands-on expertise in AWS, Terraform, AWX/Ansible, ArgoCD, GitOps, Kubernetes, CI/CD, and Disaster Recovery.

Key Responsibilities
  • Act as the L3 escalation point for complex Cloud, DevOps, Kubernetes, infrastructure, and deployment issues.
  • Lead major technical activities across production and non-production environments.
  • Take ownership of complex infrastructure changes, upgrades, migrations, and platform improvements.
  • Lead and coordinate Disaster Recovery (DR) activities, including:
    • DR planning
    • Restore testing
    • Failover/failback activities
    • Application and infrastructure recovery
    • DR validation
    • Runbook preparation and improvement
  • Design and implement infrastructure using Terraform.
  • Develop and maintain reusable Terraform modules.
  • Review Terraform code and infrastructure changes implemented by L1/L2 engineers.
  • Define Terraform coding standards, repository structure, and implementation best practices.
  • Strong hands-on experience with AWX / Ansible.
  • Develop reusable Ansible roles, playbooks, and automation workflows.
  • Use AWX to automate:
    • Server configuration
    • Patching
    • Application deployment
    • Infrastructure operations
    • Repetitive BAU activities
  • Identify manual operational activities and convert them into automated workflows.
  • Design and maintain ArgoCD-based deployments.
  • Implement and support GitOps practices for Kubernetes and application deployments.
  • Define GitOps repository structures and deployment standards.
  • Manage and troubleshoot ArgoCD:
    • Applications
    • Sync issues
    • Configuration drift
    • Deployment failures
    • Environment promotion
    • Rollback activities
  • Design and maintain Kubernetes workloads running on Amazon EKS.
  • Develop and maintain reusable Helm Charts.
  • Define standards for Helm values, templates, and environment-specific configurations.
  • Design and improve CI/CD deployment processes.
  • Support and improve Jenkins and other CI/CD automation.
  • Define deployment strategies and operational standards.
  • Lead automation initiatives across Cloud and DevOps platforms.
  • Develop automation using:
    • Terraform
    • Ansible / AWX
    • Bash
    • Python
    • CI/CD pipelines
  • Define and enforce technical standards for:
    • Infrastructure as Code
    • Terraform
    • GitOps
    • Kubernetes
    • Helm
    • CI/CD
    • Automation
    • Cloud operations
    • Patching
  • Establish reusable templates, modules, pipelines, and automation frameworks.
  • Perform technical reviews for changes implemented by L1 and L2 engineers.
  • Lead complex production incidents and perform Root Cause Analysis.
  • Identify recurring issues and implement permanent automated solutions.
  • Improve monitoring, logging, alerting, and operational reliability.
  • Participate in architecture and technical design discussions.
  • Lead production deployments, infrastructure upgrades, patching, and maintenance activities.
  • Mentor and provide technical guidance to L1 and L2 engineers.
  • Create and maintain:
    • SOPs
    • Runbooks
    • Technical standards
    • Architecture documentation
    • DR documentation
    • Operational procedures
Experience

8–12 years of relevant experience in cloud engineering, infrastructure, DevOps, or a related field is required.

Core Technical Skills
  • AWS
  • Amazon EKS
  • Kubernetes
  • Terraform
  • Terragrunt
  • AWX
  • Ansible
  • ArgoCD
  • GitOps
  • Helm
  • Docker
  • Jenkins
  • CI/CD
  • Git
  • Linux
  • Windows
  • Bash / Python
  • SSL/TLS
  • Infrastructure Automation
  • Disaster Recovery
  • Monitoring & Logging
  • Production Troubleshooting
  • Root Cause Analysis
L3 Expectations
  • Leading technical activities independently
  • Owning complex Cloud and DevOps changes
  • Designing and implementing automation
  • Leading Disaster Recovery activities
  • Implementing GitOps using ArgoCD
  • Building automation using AWX/Ansible
  • Developing and reviewing Terraform code
  • Defining technical standards and best practices
  • Driving operational improvements
  • Mentoring L1/L2 engineers
  • Handling complex production incidents and RCA
About Rackspace Technology

We are the multicloud solutions experts. We combine our expertise with the world’s leading technologies — across applications, data and security — to deliver end-to-end solutions. We have a proven record of advising customers based on their business challenges, designing solutions that scale, building and managing those solutions, and optimizing returns into the future. Named a best place to work, year after year according to Fortune, Forbes and Glassdoor, we attract and develop world-class talent. Join us on our mission to embrace technology, empower customers and deliver the future.

More on Rackspace Technology

Though we’re all different, Rackers thrive through our connection to a central goal: to be a valued member of a winning team on an inspiring mission. We bring our whole selves to work every day. And we embrace the notion that unique perspectives fuel innovation and enable us to best serve our customers and communities around the globe.

We want you to know that we are committed to offering equal employment opportunity without regard to age, color, disability, gender reassignment or identity or expression, genetic information, marital or civil partner status, pregnancy or maternity status, military or veteran status, nationality, ethnic or national origin, race, religion or belief, sexual orientation, or any legally protected characteristic. If you have a disability or special need that requires accommodation, please let us know.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Technical Program Director – Healthcare
Technical Program Director – Healthcare

Webhosting • Northern (KY)

On-site
USD 166,000 - 243,000
Incentive compensation
Employee Stock Purchase Plan (ESPP)
Solutions Architect, Forward Deployed
Solutions Architect, Forward Deployed

Rackspace Technology • United States

On-site
USD 153,000 - 269,000
Senior Manager, Cyber Security – US
Senior Manager, Cyber Security – US

Webhosting • San Antonio (TX)

On-site
USD 148,011 - 217,082
Solutions Architect, Forward Deployed
Solutions Architect, Forward Deployed

Rackspace • United States

Hybrid
USD 153,000 - 269,000
Program Manager IV
Program Manager IV

Rackspace Technology, Inc. • United States

Remote
USD 108,000 - 159,000
Employee Stock Purchase Plan (ESPP)
Annual bonus opportunities
DevOps Engineer IV
DevOps Engineer IV

Rackspace Technology, Inc. • United States

Remote
USD 132,000 - 194,000
Solutions Architect, Forward Deployed
Solutions Architect, Forward Deployed

Webhosting • Northern (KY)

On-site
USD 153,000 - 269,000
DC Ops Technician III - US
DC Ops Technician III - US

Pace Industries, LLC • Richardson (TX)

On-site
USD 52,000 - 76,000
Network Engineer IV
Network Engineer IV

Pace Industries, LLC • San Antonio (TX)

On-site
USD 110,000 - 162,000
Senior AWS Cloud Engineer – DevOps & Kubernetes Architect
Senior AWS Cloud Engineer – DevOps & Kubernetes Architect

Remote Jobs • United States

Remote
USD 120,000 - 180,000