Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Chubb Ltd. in India is seeking a CI Production Services and Site Reliability Engineering (SRE) Lead to run the NA Commercial Insurance Production Services and SRE function from the engineering center in India.
The role focuses on reliability, availability, and performance of critical production systems, bridging software engineering and operations. The successful candidate will mentor junior engineers, collaborate with development, testing and business teams, and drive a culture of reliability
Chubb is a world leader in insurance. With operations in 54 countries and territories, Chubb provides commercial and personal property and casualty insurance, personal accident and supplemental health insurance, reinsurance and life insurance to a diverse group of clients. The company is defined by its extensive product and service offerings, broad distribution capabilities, exceptional financial strength and local operations globally. Parent company Chubb Limited is listed on the New York Stock Exchange (NYSE: CB) and is a component of the S&P 500 index. Chubb employs approximately 40,000 people worldwide. Additional information can be found at: http://www.chubb.com/
At Chubb India, we are on an exciting journey of digital transformation driven by a commitment to engineering excellence and analytics. With a team of over 2500 talented professionals, we foster a startup mindset that promotes collaboration, diverse perspectives, and a solution driven attitude. We are dedicated to building expertise in engineering, analytics, and automation, empowering our teams to excel in a dynamic digital landscape. We offer an environment where you will be part of an organization that is dedicated to solving real-world challenges in the insurance industry. Together, we will work to shape the future through innovation and continuous learning.
We are seeking a leadership role to run NA Commercial Insurance Production Services and Site Reliability Engineering (SRE) function out of engineering center in India. This role will be responsible for the reliability, availability, and performance of critical production systems. In this role, you will bridge software engineering and operations, driving the cultural and technical transformation toward engineering-based operations. You will collaborate closely with SREs, development, testing and business teams to ensure our systems are robust and scalable effectively providing 99.9% availability. While the focus is on technical engineering, you will also mentor junior team members and share best practices with regional teams.
Lead a team of SREs; mentor engineers on reliability practices
Team building and performance evaluation of production services/SRE staff in CECI centers
Partner with development, testing and business teams to embed reliability requirements into the SDLC
Partner with development, infrastructure teams to ensure lower environment stability across CI footprint
Act as the escalation point for major incidents and production crises
Communicate effectively with business and operations partners, especially during critical system outages, client escalations
Partner with technology and business partners to define and own SLAs, SLOs, SLIs, and error budgets across CI portfolio
Lead incident response, blameless post-mortems, and remediation follow-through
Drive reduction in MTTR and MTTD across the production estate
Enforce change management, release gating, and production readiness reviews
Establish on-call practices, runbooks, and operational playbooks
Report on production health metrics to senior stakeholders
Partner with India leadership team in providing oversight into AMS MSM vendor engagement. This may include weekly/monthly governance calls, site visits, performance evaluation etc.
Design and implement observability frameworks (metrics, logs, traces)
Build self-healing automation to reduce toil and manual intervention
Own capacity planning, performance baselining, and scalability initiatives