Director of Production Engineering

Coinscapture

Northern (KY)

Hybrid

USD 220,000 - 265,000

Full time

11 days ago
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Legion is seeking a Director of Production Engineering to lead reliability, automation, and security foundations for our AWS-based production environment. This is a hands-on leadership role with ~20-30% code and architecture work, while the remainder focuses on vision, roadmap, and cross-team execution.

You will hire and mentor a globally distributed DevOps/SRE team, own the infrastructure roadmap (EKS, RDS, AWS services), lead SecOps, define OKRs, and champion observability and IaC to scale our

Qualifications

  • 8–12 years in DevOps, SRE, or production infra with people mgmt experience.
  • Extensive hands-on AWS experience: EKS, RDS, VPC, IAM, S3.
  • 5+ years using observability tools (Datadog, Prometheus, Grafana).
  • 5+ years leading incident management and on-call practices.

Responsibilities

  • Hire, mentor, and lead a globally distributed DevOps/SRE team.
  • Own the reliability and infrastructure roadmap for the AWS-based production environment (EKS, RDS, AWS services).
  • Lead the organization's SecOps practice: vulnerability management, incident response, remediation.
  • Define and drive engineering OKRs for infra reliability, automation, and security.
  • Champion observability and alerting practices (Datadog) and automate triage and response.
  • Drive Infrastructure-as-Code, CI/CD, and automation to increase velocity.
  • Collaborate with engineering and IT to align standards and compliance.
  • Lead on-call rotation and incident management to meet availability goals.

Skills

DevOps Leadership
SRE
AWS
Kubernetes
SecurityOps
Observability
IaC/Automation
CI/CD
Go/Python/Bash

Education

Bachelor's degree in CS/Engineering
Master's degree preferred

Tools

Datadog
Prometheus
Grafana
Terraform
CloudFormation
Argo Workflows
Helm
Git

Job description

Director of Production Engineering

Remote, United States

About this Position

Are you passionate about building the reliability, automation, and security foundations that let engineering teams move fast with confidence? At Legion, we are seeking a Director of Engineering, DevOps & SRE to lead the teams responsible for the availability, scalability, and security of our production environment. Our production infrastructure runs on AWS, leveraging services such as EKS, RDS, and a broad set of AWS-native technologies. You will partner closely with engineering and IT to build resilient systems, drive operational excellence, and ensure our platform meets the highest standards of security and compliance.

This is a hands‑on leadership role where you'll spend ~20-30% of your time contributing directly to architecture, tooling, and incident response, and the rest driving vision, roadmap, and cross‑team execution.

Responsibilities
  • Hire and build a globally‑distributed DevOps/SRE engineering team. Recruit, mentor, and manage engineers, and foster a culture of ownership, collaboration, and continuous improvement.
  • Own the reliability and infrastructure roadmap for our AWS‑based production environment, including EKS, RDS, and related AWS services, ensuring scalability, high availability, and cost efficiency.
  • Lead the organization's security operations (SecOps) practice, including vulnerability management, threat detection, incident response, and remediation, to proactively identify and resolve security issues before they impact customers.
  • Define and drive engineering OKRs for infrastructure reliability, automation, and security, and track progress against measurable outcomes.
  • Champion observability and alerting best practices (e.g., Datadog), including automating alert triage and response to reduce mean‑time‑to‑resolution.
  • Solid understanding of agentic AI infrastructure and how AI agentic workflows apply to SDLC and DevOps processes (e.g., automated investigation, remediation, and PR‑generation pipelines).
  • Drive Infrastructure‑as‑Code, CI/CD, and automation practices to increase engineering velocity and reduce operational toil.
  • Work closely with engineering and IT teams to align on infrastructure standards, access controls, tooling, and compliance requirements across the organization.
  • Ensure the platform meets the highest standards of security, compliance, and data protection; implement and maintain robust security controls and audit‑readiness.
  • Lead and participate in the Incident Management on‑call rotation, working with SRE and development teams to meet and exceed availability goals.
  • Stay current on cloud, DevOps, and security best practices, and provide technical guidance and thought leadership to the broader engineering organization.
Required Qualifications
  • 8-12 years of experience in DevOps, Site Reliability Engineering, or production infrastructure roles, including people management experience.
  • Deep hands‑on experience running production workloads on AWS, including EKS (Kubernetes), RDS, and other core AWS services (e.g., VPC, IAM, Lambda, S3).
  • Demonstrated experience running security operations (SecOps) — vulnerability management, incident response, and remediation of production security issues.
  • 5+ years of experience leveraging observability platforms (e.g., Datadog, Prometheus, Grafana) to drive reliability, performance, and alerting improvements.
  • Strong experience with Infrastructure‑as‑Code (e.g., Terraform, CloudFormation) and CI/CD automation.
  • Proficiency in at least one of Go, Python, or Bash, with day‑to‑day use of Git and test automation pipelines.
  • Hands‑on experience operating Linux/Unix production platforms (Amazon Linux, Ubuntu, RHEL/CentOS).
  • Proven track record partnering cross‑functionally with engineering and IT teams to align on infrastructure, tooling, and security standards.
  • Demonstrated experience leading incident management and on‑call practices for high‑availability production systems.
  • Bachelor's degree in Computer Science, Engineering, or related field required; Master's degree preferred.
Preferred Qualifications
  • Experience with major cloud providers beyond AWS, such as Google Cloud Platform or Oracle Cloud Infrastructure (OCI).
  • Relevant security certifications (e.g., CISSP, AWS Security Specialty, CKS).
  • Experience with compliance frameworks such as SOC 2, ISO 27001, or HIPAA.
  • 3+ years of experience with Kubernetes or other containerization/orchestration platforms at scale.
  • Experience with Kubernetes‑native delivery tooling, including Argo Workflows and Helm.
  • 5+ years of experience leading teams in an agile/scrum environment.
  • Experience building or scaling automated investigation and remediation pipelines for production error classes.
COMPENSATION & BENEFITS

Salary Range

Base Salary

Range $220,000 - $265,000 + Bonus + Stock Equity

At Legion, we offer competitive compensation and benefits packages to…

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Director of Production Reliability & SRE
Director of Production Reliability & SRE

Coinscapture • Northern (KY)

Hybrid
USD 220,000 - 265,000
Director, DevOps & SRE — AWS Infra & SecOps Leader
Director, DevOps & SRE — AWS Infra & SecOps Leader

Legion • California City (CA)

On-site
USD 220,000 - 265,000
Free medical plans
401k plan
PTO & holidays
+3
Director of Production Engineering
Director of Production Engineering

Legion Technologies, Inc. • San Francisco (CA)

Hybrid
USD 220,000 - 265,000
Zero-premium health plans and other PP
401k plan
Discretionary PTO and paid holidays
+3
Head of Production DevOps, SRE & Security
Head of Production DevOps, SRE & Security

Legion Technologies, Inc. • San Francisco (CA)

Hybrid
USD 220,000 - 265,000
Zero-premium health plans and other PP
401k plan
Discretionary PTO and paid holidays
+3
Director of Production Engineering Legion Anywhere in the World $220,000 - $265,000/yr
Director of Production Engineering Legion Anywhere in the World $220,000 - $265,000/yr

Neura Market • Northern (KY)

Hybrid
USD 220,000 - 265,000
401k plan
Discretionary Paid Time Off and Paid H
Equity
+4
Remote Director of Production Engineering & SRE
Remote Director of Production Engineering & SRE

Neura Market • Northern (KY)

Hybrid
USD 220,000 - 265,000
401k plan
Discretionary Paid Time Off and Paid H
Equity
+4
Remote Cloud Reliability & DevOps Director
Remote Cloud Reliability & DevOps Director

Neura Market • Northern (KY)

Hybrid
USD 220,000 - 265,000
Health insurance
401k plan
Discretionary PTO & holidays
+4
DevSecOps Engineer
DevSecOps Engineer

Tari Labs, LLC. • United States

On-site
USD 135,000 - 220,000
Senior DevOps Engineer/Site Reliability Engineer-East Coast
Senior DevOps Engineer/Site Reliability Engineer-East Coast

Stellar Cyber • North Carolina

On-site
USD 165,000 - 215,000
Pre‑IPO Stock Options
Medical, Dental & Vision care
401(k)
+2
DevOps Engineer
DevOps Engineer

Engtal • Washington

On-site
USD 100,000 - 130,000
Equity in the form of stock options
Flexible time-off policy and company holidays
Health, dental, and vision insurance with company contributions
+3