DevOps / Site Reliability Engineer ID70127

AgileEngine

Miami (FL)

On-site

USD 150,000 - 225,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

AgileEngine seeks an SRE Manager to lead a global team focused on building scalable, reliable systems. The role blends deep technical leadership with strategic planning and operational excellence across production environments.

You will own SLAs/SLOs, drive automation, IaC and cost optimization, and collaborate with security, product, and infra teams to raise reliability and performance across cloud platforms.

Qualifications

  • Bachelor's or Master's degree in computer science, engineering, or related field.
  • 15+ years of experience in software engineering or infrastructure, with 5+ years in SRE/DevOps leadership.
  • Deep understanding of cloud platforms (AWS, GCP), containers (Docker, Kubernetes), monitoring (Prometheus, Grafana, Datadog, New Relic), and automation tools (Terraform, Ansible).
  • Experience with modern CI/CD tools (Jenkins, ArgoCD, GitHub Actions).
  • Strong leadership, communication, and team development skills.

Responsibilities

  • Build and lead a high-performing SRE/DevOps team across multiple segments.
  • Define and execute the SRE strategy aligned with business and engineering goals.
  • Own SLAs/SLOs/SLIs and drive incident management, RCA, and continuous improvement.
  • Oversee reliability tooling, runbooks, and automation frameworks.
  • Partner with Infrastructure, DevOps, and Cloud teams for scalable platform architecture.
  • Guide IaC adoption, CI/CD pipelines, and modern observability tools.
  • Drive cost optimization and efficient resource use in cloud environments.
  • Report reliability metrics to leadership and stakeholders.

Skills

Leadership
Cloud platforms
Observability
Automation
CI/CD

Education

Bachelor's or Master's in CS/Engineering

Tools

Terraform
Ansible
Jenkins
ArgoCD
GitHub Actions
Docker
Kubernetes
Prometheus
Grafana
Datadog
New Relic

Job description

Location: Miami, FL US (ONSITE).

Job Title: SRE Manager

Salary: $150,000.00–$225,000.00 yearly

Contract: Full-time

Salary: $85.00 hourly

Salary: $85,000.00–$125,000.00 yearly

Salary: $85,000 - 125,000 per year.

Salary: $128,000.00 yearly

Salary: $70.00 hourly

Salary: $204,400.00 yearly

Contract: Temporary

Contract: Full-time

Contract: Part-time

Location: Remote

Location: Onsite

Job Description

SRE - Manager to lead our global SRE team in building scalable, resilient, and highly available systems. This role combines deep technical expertise with strong leadership, strategic thinking, and a passion for delivering exceptional customer experiences through operational excellence. As SRE Manager, this individual will champion automation initiatives for SRE operations, aiming to enhance the performance and reliability of infrastructure and critical services. The role involves close collaboration with engineering, product, security, and operations teams to define and implement reliability best practices organization-wide.

Key Responsibilities
  • Build and lead a high-performing team of SREs across various Business segment.
  • Define and execute the SRE strategy aligned with business and engineering goals.
  • Foster a culture of reliability, observability, and performance.
  • Reliability Engineering
  • Own SLAs/SLOs/SLIs for key services and ensure they are met consistently.
  • Drive incident management practices, root cause analysis (RCA), and continuous improvement.
  • Oversee reliability tooling, runbooks, and automation frameworks.
  • Platform Infrastructure
  • Partner with Infrastructure, DevOps, and Cloud teams to ensure scalable platform architecture.
  • Guide the adoption of Infrastructure-as-Code (IaC), CI/CD pipelines, and modern observability tools.
  • Drive cost optimization and efficient resource utilization in cloud environments.
  • Act as a reliability evangelist across engineering teams, enabling them to own and improve their services.
  • Report reliability and performance metrics to leadership and stakeholders.
  • Collaborate closely with security, compliance, and governance teams to meet regulatory requirements.
Qualifications

Bachelor's or master's degree in computer science, Engineering, or related field.

15+ years of experience in software engineering or infrastructure roles, with at least 5+ years in SRE or DevOps leadership.

Deep understanding of cloud platforms (AWS GCP), containers (Docker, Kubernetes), monitoring (Prometheus, Grafana, Datadog, New Relic), and automation tools (Terraform, Ansible, etc.).

Experience with modern CI/CD tools (e.g., Jenkins, ArgoCD, GitHub Actions).

Strong leadership, communication, and team development skills.

Preferred Qualifications
  • Experience in regulated industries (e.g., Telecom, communications) and Global telco leaders.
  • Certifications in cloud platforms (AWS Certified DevOps Engineer, Google SRE Certificate, etc.).
  • Experience managing hybrid or multi-cloud environments.
  • Worked as senior role in Top 5 Consultancy companies.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior DevOps/SRE Engineer
Senior DevOps/SRE Engineer

VITG • Ellicott City (MD)

Hybrid
USD 90,000 - 120,000
401(k) with employer contribution
Medical/Dental/Vision insurance
Paid vacation (PTO)
Lead Site Reliability Engineer (SRE) / Principal Site Reliability Engineer (SRE)
Lead Site Reliability Engineer (SRE) / Principal Site Reliability Engineer (SRE)

Mindlance • Irving (TX)

Hybrid
USD 120,000 - 160,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Storm2 • Scottsdale (AZ)

Hybrid
USD 140,000 - 150,000
Competitive healthcare, dental, and vision coverage
401(k) with company match
Generous PTO and paid holidays
+1
Senior DevOps Engineer/Site Reliability Engineer-East Coast
Senior DevOps Engineer/Site Reliability Engineer-East Coast

Stellar Cyber • New Jersey

On-site
USD 165,000 - 215,000
Pre-IPO Stock Options
Medical, Dental & Vision care
401(k)
+2
Site Reliability Engineering (SRE)
Site Reliability Engineering (SRE)

Weekday (YC W21) • New York (NY)

On-site
USD 150,000 - 250,000
Health, dental, vision insurance
Generous PTO
Learning & development
+2
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

OutSolve • Mission (KS)

Remote
USD 90,000 - 130,000
100% remote work environment
Competitive compensation
Professional development opportunities
+1
Senior DevOps Engineer/Site Reliability Engineer-East Coast
Senior DevOps Engineer/Site Reliability Engineer-East Coast

Stellar Cyber • New York (NY)

Hybrid
USD 165,000 - 215,000
Pre-IPO Stock Options
Medical, Dental & Vision care
401(k)
+1
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

JPS Tech Solutions • Colorado

On-site
USD 160,000 - 230,000
Site Reliability Engineer
Site Reliability Engineer

TalentDome Staffing • United States

On-site
USD 140,000 - 210,000
Senior Vice President, Site Reliability Engineer
Senior Vice President, Site Reliability Engineer

BNY Mellon • Town of Florida (NY)

On-site
USD 170,000 - 230,000