Lead SRE — AI-Driven, Observability & Resilience

Next Frontier Capital

Dublin

On-site

EUR 120,000 - 150,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

JPMorgan Chase in Infrastructure Platforms is seeking a Lead Site Reliability Engineer who writes production software to strengthen reliability across enterprise-scale platforms. You will own systems end-to-end, define SLIs/SLOs, and lead on-call incident response.

You will work AI-native across the SDLC to accelerate delivery while maintaining accountability for correctness, security, reliability, and cost. Expect to drive durable fixes and reduce toil across teams.

Qualifications

  • Solid production coding experience in Python, Go, Java, C++, or Rust; software engineering for reliability.
  • Experience running production systems at scale, including on-call ownership and incident response.
  • Familiar with SLI/SLO and error-budget practices; capable of designing for reliability and operability.
  • Depth in observability, telemetry, and alerting with modern tools.
  • Experience with Unix-based systems and infrastructure automation tooling.

Responsibilities

  • Engineer reliability into enterprise-scale platforms by writing production software, automation, and tooling.
  • Build systems around declarative design and trusted data sources to reconcile reality with desired state.
  • Instrument and reason over telemetry to drive detection, diagnosis, and closed-loop remediation.
  • Define and operationalize SLIs and SLOs with stakeholders and implement SLO-based alerting.
  • Own services end-to-end with accountability for reliability, performance, security, and cost.

Skills

Python
Go
Java
C++
Rust
Grafana
Prometheus
Splunk
Datadog
Dynatrace
Kubernetes
Terraform
CI/CD
On-call ownership
AI-native tooling

Tools

Grafana
Prometheus
Splunk
Datadog
Dynatrace
Kubernetes
Terraform
CI/CD

Job description

JPMorgan Chase in Infrastructure Platforms is seeking a Lead Site Reliability Engineer who writes production software to strengthen reliability across enterprise-scale platforms. You will own systems end-to-end, define SLIs/SLOs, and lead on-call incident response.

You will work AI-native across the SDLC to accelerate delivery while maintaining accountability for correctness, security, reliability, and cost. Expect to drive durable fixes and reduce toil across teams.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Lead SRE: Scale Reliability, AI-Driven Ops
Senior Lead SRE: Scale Reliability, AI-Driven Ops

Next Frontier Capital • Dublin

On-site
EUR 120,000 - 180,000
Senior Lead SRE: AI-Driven Cloud Reliability
Senior Lead SRE: AI-Driven Cloud Reliability

JPMorgan Chase & Co. • Dublin

On-site
EUR 100,000 - 140,000
AI-Driven Lead Site Reliability Engineer
AI-Driven Lead Site Reliability Engineer

JPMorgan Chase & Co. • Dublin

On-site
EUR 120,000 - 180,000
AI-Driven Lead Site Reliability Engineer
AI-Driven Lead Site Reliability Engineer

JPMorganChase • Dublin

On-site
EUR 150,000 - 190,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

JPMorgan Chase & Co. • Dublin

On-site
EUR 120,000 - 180,000
Senior Lead SRE: Build Reliable Cloud Platforms & AI Ops
Senior Lead SRE: Build Reliable Cloud Platforms & AI Ops

JPMorganChase • Dublin

On-site
EUR 120,000 - 180,000
Senior SRE Lead: Reliability, Cloud & Automation
Senior SRE Lead: Reliability, Cloud & Automation

JPMorganChase • Dublin

On-site
EUR 90,000 - 130,000
Senior SRE: AI-Driven Reliability & Platform Leadership
Senior SRE: AI-Driven Reliability & Platform Leadership

JPMorganChase • Dublin

On-site
EUR 90,000 - 150,000
Senior Lead SRE (DSS)
Senior Lead SRE (DSS)

JPMorgan Chase & Co. • Dublin

On-site
EUR 100,000 - 140,000
Senior Site Reliability Engineer: AI-Driven Cloud
Senior Site Reliability Engineer: AI-Driven Cloud

Fairygodboss • Dublin

On-site
EUR 90,000 - 130,000