Lead Site Reliability Engineering - Network

JPMorganChase

Palo Alto (CA)

On-site

USD 150,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

JPMorgan Chase in Palo Alto is seeking a Lead Site Reliability Engineer to play a critical role in enhancing network reliability. The position involves leading complex technical projects and collaborating with multiple teams to achieve operational excellence.

The successful candidate will possess advanced knowledge in network engineering and cloud technologies, and have over 10 years of experience managing technical teams. A commitment to innovation and effective communication is essential for driving success in this role.

Qualifications

  • 5+ years of applied experience in network engineering concepts.
  • 10+ years of experience leading technologists in complex technical domains.
  • Ability to influence team culture through innovation and change.

Responsibilities

  • Apply network reliability principles to ensure balance in delivery and stability.
  • Partner with network engineering domains to align goals.
  • Drive adoption of best practices and observability.

Skills

Network reliability engineering
SD-WAN
Cloud platforms (AWS, Azure)
Observability tools (Grafana, Splunk)

Education

Formal training or certification in network engineering

Tools

Jenkins
GitLab
Terraform

Job description

Assume a critical role in defining the future of a globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Network Product, you hold a leadership role in your team, demonstrate strong knowledge across multiple technical domains, and advise others on the technical and business issues facing them. Take lead and conduct resiliency design reviews, break up complex problems into digestible work for other engineers, act as a technical lead for medium to large-sized products, and provide advice and mentoring to other engineers.

Job responsibilities
  • Applies network reliability principles (Permit to Operate, FMEA, operational readiness), balancing feature delivery, efficiency, and stability.
  • Partners with network engineering domains (Datacenter, Firewall, Proxies, DMZ, Load Balancing, etc.) and Lines of Business to align goals and outcomes.
  • Drives adoption of reliability best practices and observability, demonstrating impact through stability/reliability metrics.
  • Bridges Engineering, Operations, DevOps, and customers to build resilient, scalable, and secure network services.
  • Provides Tier-3 network support, leading major incident response, rapid restoration, RCA, and follow-through on corrective actions.
  • Leads reliability and stability initiatives using data-driven analysis to improve service levels and reduce recurring failure modes.
  • Defines SLI/SLOs and error budgets with stakeholders and customers, ensuring measurable performance targets and trade-off clarity.
  • Identifies and removes technical bottlenecks within core domains of expertise, proactively preventing reliability and capacity risks.
  • Runs blameless, data-driven post-mortems and debriefs, converting learnings (successes and failures) into actionable improvements.
  • Fosters continuous improvement and strong knowledge sharing, soliciting real-time feedback, avoiding duplicated work, and promoting innovation via internal communities.
  • Produces and packages thought leadership with specialists/product/engineering teams-documenting best practices and lessons learned for internal assets and industry forums/conferences.
Required qualifications, capabilities, and skills
  • Formal training or certification in network engineering concepts and 5+ years of applied experience.
  • 10+ years of experience leading technologists to manage and solve complex technical items within your domain of expertise.
  • Advanced proficiency in network reliability engineering, including Permit to Operate, FMEA, and operational readiness processes.
  • Experience leading technologists to manage and solve complex network issues at a firmwide level.
  • Ability to influence team culture by championing innovation and change for success.
  • Proficiency in SD-WAN, cloud platforms (AWS, Azure, etc.), and major network technologies (Palo Alto, Juniper, F5, Broadcom, Arista, Cisco, etc.).
  • Proficiency in observability and monitoring tools such as Grafana, SevOne, Prometheus, Kibana, ThousandEyes, and Splunk.
Preferred qualifications, capabilities, and skills
  • CCIE, Load-balancing, SD-WAN, Observability tools, eBPF, Cloud certs
  • Demonstrated proficiency in troubleshooting and supporting complex networking environments, including Tier-3 operational support for major incidents.
  • Experience with continuous integration and delivery tools (e.g., Jenkins, GitLab, Terraform, etc.).
  • Experience in scalable networking design, including high availability, redundancy, failover, and load balancing.
  • Experience troubleshooting networking protocols such as TCP/IP, HTTPS, and BGP.
  • Experience in customer-facing migration, including service discovery, assessment, planning, execution, and operations.

This position is subject to Section 19 of the Federal Deposit Insurance Act. As such, an employment offer for this position is contingent on JPMorganChase's review of criminal conviction history, including pretrial diversions or program entries.

We are an equal opportunity employer and place a high value on diversity and inclusion at our company. We do not discriminate on the basis of any protected attribute, including race, religion, color, national origin, gender, sexual orientation, gender identity, gender expression, age, marital or veteran status, pregnancy or disability, or any other basis protected under applicable law. We also make reasonable accommodations for applicants' and employees' religious practices and beliefs, as well as mental health or physical disability needs. Visit our FAQs for more information about requesting an accommodation. JPMorgan Chase & Co. is an Equal Opportunity Employer, including Disability/Veterans

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead Site Reliability Engineering - Network
Lead Site Reliability Engineering - Network

Next Frontier Capital • Palo Alto (CA)

On-site
USD 180,000 - 270,000
Health care coverage
On-site wellness centers
Retirement savings plan
+3
Lead Site Reliability Engineering - Network
Lead Site Reliability Engineering - Network

JPMorgan Chase & Co. • Palo Alto (CA)

On-site
USD 140,000 - 180,000
Lead Site Reliability Engineering - Network
Lead Site Reliability Engineering - Network

JPMorgan Chase • Palo Alto (CA)

On-site
USD 152,000 - 215,000
Comprehensive health care coverage
Tuition reimbursement
Retirement savings plan
Lead Site Reliability Engineer Market Risk
Lead Site Reliability Engineer Market Risk

Next Frontier Capital • Houston (TX)

On-site
USD 120,000 - 150,000
Lead Site Reliability Engineer Market Risk
Lead Site Reliability Engineer Market Risk

JPMorganChase • Houston (TX)

On-site
USD 120,000 - 150,000
Comprehensive health care coverage
Retirement savings plan
Tuition reimbursement
+1
Lead Site Reliability Engineer
Lead Site Reliability Engineer

JPMorganChase • Jersey City (NJ)

On-site
USD 170,000 - 230,000
AWS Certifications
Lead Site Reliability Engineer
Lead Site Reliability Engineer

JPMorganChase • Columbus (OH)

On-site
USD 150,000 - 190,000
Senior Lead Site Reliability Engineer
Senior Lead Site Reliability Engineer

Fairygodboss • Plano (TX)

On-site
USD 150,000 - 190,000
Comprehensive health care coverage
On-site health and wellness centers
Retirement savings plan
+4
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Next Frontier Capital • Wilmington (DE)

On-site
USD 180,000 - 230,000
Health insurance
Retirement savings plan
Lead Site Reliability Engineer
Lead Site Reliability Engineer

JPMorganChase • Jersey City (NJ)

On-site
USD 170,000 - 250,000