Lead SRE- Azure & GCP

Next Frontier Capital

Glasgow

On-site

GBP 90,000 - 140,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

JPMorgan Chase is seeking a Lead Site Reliability Engineer to drive SRE frameworks across global Google Cloud environments and promote a DevOps culture with automated, elastic services.

You will work within the Infrastructure Platform - Cloud Foundational Services SRE team, collaborating across regions to maintain high availability and strong SLAs while leveraging enterprise AI tools to optimize incident response and reliability.

Qualifications

  • Experience with Google Cloud and Azure in production environments.
  • Strong container tech: Docker, Kubernetes, GKE, HELM.
  • Proficiency in Python, shell scripting or Go with REST APIs knowledge.
  • Hands-on with cloud-based deployment, monitoring and operations tools (Prometheus, Grafana, DataDog, Splunk, Elasticsearch).
  • Experience using enterprise AI capabilities to improve SRE workflows with validation and data-sensitivity awareness.
  • Ability to evaluate AI-assisted recommendations, define guardrails, ensure resiliency and security.

Responsibilities

  • Lead and implement SRE frameworks to support global Google Cloud environments and achieve high SLOs.
  • Mastery of application, data, infrastructure, and AI disciplines.
  • Partner across teams on budgeting and governance for cost management.
  • Utilize AI capabilities to accelerate incident triage, troubleshooting, and post-incident analysis with proper data handling.
  • Develop and improve technical engineering documentation.
  • Provide technical supervision and problem resolution for engineering activities.
  • Champion a DevOps model to automate services and enable elasticity across platforms.

Skills

Google Cloud
Azure Cloud
Docker
Kubernetes
GKE
HELM
Python
Shell scripting
Go
REST APIs
DataDog
Prometheus
Grafana
Splunk
Elasticsearch
Git
CI/CD
Terraform
Jenkins

Tools

CI/CD Tools
Observability Tools

Job description

We have a Lead Site Reliability Engineer (SRE) opportunity within our Google Cloud Site Reliability Engineering team.

As a Lead Site Reliability Engineer at JPMorgan Chase within the Infrastructure Platform - Cloud Foundational Services SRE organization, you will join our Google Cloud Site Reliability Engineering team operating within a global follow-the-sun support model.

Job Responsibilities:
  • Lead and Implement SRE frameworks to support global google cloud environments and ensure the highest level of SLOs through operational excellence
  • Mastery of application, data, infrastructure, and Agentic AI disciplines
  • Keen understanding of financial control and budget management using expertise in working in partnership with colleagues throughout the firm, and in leading collaborative teams to achieve common goals
  • Uses enterprise-authorized AI capabilities within the work environment to accelerate major-incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirements.
  • Provide support to develop & improve the quality of technical engineering documentation
  • Provide technical supervision, oversight and problem resolution for engineering activities
  • Champion a DevOps model so that services are automated and elastic across all platforms
Required qualifications, capabilities, and skills:
  • Google & Azure cloud expertise in a mission critical production environment
  • Strong understanding about container technologies such as Docker, Kubernetes, GKE and HELM
  • Experience in programming in one of the following languages: Python, shell scripting or GO along with good understanding of REST APIs
  • Hands-on experience with cloud-based technologies and tools especially in deployment, monitoring and operations, such as Google Observability, Azure Monitor, Data Dog, Prometheus, Splunk, Elasticsearch and Grafana.
  • Demonstrated experience using enterprise-authorized AI capabilities within the work environment to improve SRE workflows (e.g., incident investigation support and knowledge capture) with strong validation habits and awareness of data sensitivity.
  • Ability to evaluate AI-assisted operational recommendations for correctness and risk, define appropriate guardrails for team usage, and ensure outcomes align to resiliency and security expectations.
  • Strong understanding about the Google Cloud governance and compliance and cost management
  • Strong working knowledge of modern development technologies and tools such Agile, CI/CD, Git, Infrastructure as Code, Terraform and Jenkins.
  • Google Cloud certification or equivalent technical experience in the Public Cloud.
  • Good understanding of Agentic AI SDKs and GitHub Copilot Skills.
Preferred qualifications, capabilities, and skills:
  • Good understanding of operating systems such as Windows, Linux (Redhat / Ubuntu)
  • Good understanding of LLM and other AI/ML frameworks which can be used in AIOPS

J.P. Morgan is a global leader in financial services, providing strategic advice and products to the world’s most prominent corporations, governments, wealthy individuals and institutional investors. Our first-class business in a first-class way approach to serving clients drives everything we do. We strive to build trusted, long-term partnerships to help our clients achieve their business objectives.

We recognize that our people are our strength and the diverse talents they bring to our global workforce are directly linked to our success. We are an equal opportunity employer and place a high value on diversity and inclusion at our company. We do not discriminate on the basis of any protected attribute, including race, religion, color, national origin, gender, sexual orientation, gender identity, gender expression, age, marital or veteran status, pregnancy or disability, or any other basis protected under applicable law. We also make reasonable accommodations for applicants’ and employees’ religious practices and beliefs, as well as mental health or physical disability needs. Visit our FAQs for more information about requesting an accommodation.

Our professionals in our Corporate Functions cover a diverse range of areas from finance and risk to human resources and marketing. Our corporate teams are an essential part of our company, ensuring that we’re setting our businesses, clients, customers and employees up for success. We have a Lead Site Reliability Engineer (SRE) opportunity within our JPMC Google & Azure Cloud Site Reliability Engineering team.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead SRE- Azure & GCP
Lead SRE- Azure & GCP

Fairygodboss • Glasgow

On-site
GBP 120,000 - 170,000
Lead SRE- Azure & GCP
Lead SRE- Azure & GCP

JPMorganChase • Glasgow

On-site
GBP 90,000 - 130,000
Lead SRE - Azure and GCP
Lead SRE - Azure and GCP

Hackajob Ltd • Glasgow

On-site
GBP 90,000 - 110,000
Lead SRE - AWS Platform
Lead SRE - AWS Platform

JPMorganChase • Glasgow

On-site
GBP 90,000 - 130,000
Senior Manager of SRE
Senior Manager of SRE

Fairygodboss • Glasgow

On-site
GBP 120,000 - 180,000
Lead Site Reliability Engineer - Chief Technology Office
Lead Site Reliability Engineer - Chief Technology Office

J.P. MORGAN • Glasgow

On-site
GBP 90,000 - 140,000
Senior Lead Software Engineer - Python / Go
Senior Lead Software Engineer - Python / Go

Next Frontier Capital • Glasgow

On-site
GBP 90,000 - 130,000
Senior Manager of SRE
Senior Manager of SRE

JPMorganChase • Glasgow

On-site
GBP 110,000 - 170,000
Lead SRE - AWS Platform
Lead SRE - AWS Platform

JPMorgan Chase & Co. • Glasgow

On-site
GBP 90,000 - 130,000
Lead SRE - Chase UK
Lead SRE - Chase UK

JPMorganChase • Greater London

On-site
GBP 90,000 - 150,000