Senior Infrastructure Engineer

Jobtailor

Arizona

On-site

USD 120,000 - 190,000

Full time

36 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Jobtailor is seeking a Senior Platform/DevOps Engineer to build and scale high-availability AWS infrastructure, with a focus on GitOps, CI/CD, and IaC.

You will configure telemetry tooling (Prometheus, Grafana, OpenTelemetry), drive incident response, and mentor engineers while collaborating with product, security, and core engineering teams to improve platform adoption and performance.

Qualifications

  • 5+ years of experience contributing to complex, large-scale distributed systems in mission-critical environments.
  • Hands-on proficiency in AWS ecosystems and Terraform for reproducible environments.
  • Experience with Kubernetes (EKS), Docker, and GitOps workflows (Flux, Argo, GitHub Actions).
  • Proficient coding/scripting in Python, Go, or Bash for automation.
  • Experience with Prometheus, Grafana, or OpenTelemetry.
  • Knowledge of cloud security, API Gateways, load balancing, and network isolation.

Responsibilities

  • Build, maintain, and optimize high-availability AWS infrastructure.
  • Contribute to a standardized, globally scalable fleet managed entirely through code.
  • Implement automated infrastructure solutions using GitOps, modern CI/CD pipelines, and infrastructure as code.
  • Configure and maintain Prometheus, Grafana, and OpenTelemetry telemetry tooling.
  • Monitor performance and identify bottlenecks.
  • Participate in incident-response on-call rotations and lead blameless post-mortems.
  • Partner with Product, Security, and Core Engineering teams to support adoption of Golden Paths.
  • Mentor junior and mid-level engineers and foster strong technical practices.
  • Maintain and support high-performance, low-latency private connectivity for external customer traffic.
  • Partner with internal engineering teams to support platform adoption.

Skills

AWS
Terraform
Kubernetes
Docker
GitOps
Python
Go
Bash
Prometheus
Grafana

Tools

GitHub Actions
OpenTelemetry

Job description

  • Build, maintain, and optimize high-availability AWS infrastructure
  • Contribute to a standardized, globally scalable fleet managed entirely through code
  • Implement automated infrastructure solutions using GitOps, modern CI/CD pipelines, and infrastructure as code
  • Configure and maintain Prometheus, Grafana, and OpenTelemetry telemetry tooling
  • Monitor performance and identify bottlenecks
  • Participate in incident-response on-call rotations and lead blameless post-mortems
  • Partner with Product, Security, and Core Engineering teams to support adoption of Golden Paths
  • Mentor junior and mid-level engineers and foster strong technical practices
  • Maintain and support high-performance, low-latency private connectivity for external customer traffic
  • Partner with internal engineering teams to support platform adoption
Requirements
  • 5+ years of experience contributing to complex, large-scale distributed systems within mission-critical environments
  • Solid hands-on proficiency in AWS ecosystems and Terraform for building reproducible environments
  • Hands-on experience with Kubernetes (EKS), Docker, and GitOps workflows (Flux, Argo, GitHub Actions)
  • Proficient coding/scripting skills in Python, Go, or Bash for infrastructure automation and operational tooling
  • Hands-on experience working with Prometheus, Grafana, or OpenTelemetry
  • Good understanding of cloud security, API Gateways, load balancing, and network isolation principles

Demonstrates expertise in building and optimizing AWS infrastructure, implementing automated solutions with GitOps, and maintaining telemetry tools like Prometheus and Grafana. Proven ability to mentor engineers and collaborate with cross-functional teams to enhance platform adoption and performance.

Highest-signal resume keywords
  • AWS Infrastructure Management
  • Terraform Proficiency
  • Kubernetes (EKS) Experience
  • Infrastructure Automation with Python, Go, or Bash
  • Prometheus and Grafana Configuration
ATS Optimization Keywords
Hard Skills
  • AWS
  • Terraform
  • Kubernetes
  • Docker
  • GitOps
  • Python
  • Go
  • Bash
  • Prometheus
  • Grafana
Soft Skills
  • Mentoring
  • Collaboration
  • Incident Response
Industry Keywords
  • Distributed Systems
  • Cloud Security
  • API Gateways
  • Load Balancing
  • Network Isolation
Tools & Technologies
  • GitHub Actions
  • OpenTelemetry
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior DevOps Engineer
Senior DevOps Engineer

Jobtailor • Lehi (UT)

On-site
USD 120,000 - 180,000
Senior Forward Deployed Engineer
Senior Forward Deployed Engineer

Jobtailor • Washington

On-site
USD 140,000 - 185,000
Senior Site Reliability Engineer – AWS/Datacenter
Senior Site Reliability Engineer – AWS/Datacenter

Jobtailor • Bellevue (WA)

On-site
USD 160,000 - 220,000
Cloud Architect, Solution Design Engineer
Cloud Architect, Solution Design Engineer

Jobtailor • California (MO)

On-site
USD 150,000 - 230,000
AWS DevOps Engineer
AWS DevOps Engineer

Recru • Houston (TX)

On-site
USD 140,000 - 190,000
Forward Deployed Engineer – AI Training Data
Forward Deployed Engineer – AI Training Data

Jobtailor • Redwood City (CA), Northern (KY)

Hybrid
USD 140,000 - 170,000
Senior Software Engineer (Infrastructure)
Senior Software Engineer (Infrastructure)

Hayden AI • United States

On-site
USD 180,000 - 260,000
Tazapay - Staff DevOps Engineer
Tazapay - Staff DevOps Engineer

Tazapay • Bloomington (IN)

On-site
USD 150,000 - 190,000
DevOps Engineer DevOps Engineer
DevOps Engineer DevOps Engineer

Kurai • Seattle (WA)

On-site
USD 120,000 - 180,000
AWS DevOps Engineer
AWS DevOps Engineer

Data Pro Software • Lexington (MA)

On-site
USD 120,000 - 180,000