Site Reliability Engineer

Guidewire Software

Bengaluru

On-site

INR 1,200,000 - 1,800,000

Full time

10 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Guidewire Software in Bengaluru seeks an experienced SRE/DevOps engineer to design, build, and operate highly reliable, scalable production systems for a multi-tenant SaaS platform.

The role emphasizes automation, internal tool development, and collaboration with development teams to meet availability and performance targets. Candidates should have deep AWS/Kubernetes experience and strong coding skills in Python or Go.

Qualifications

  • 8-12 years of hands-on experience in SRE/DevOps, Cloud Infrastructure, or related platform engineering.
  • Strong programming skills in Python or Go (Java/Spring Boot is a plus).
  • Deep experience with AWS and building/operating production systems at scale.
  • Hands-on expertise with Kubernetes (EKS), Docker, Helm, CNI, and Ingress networking.
  • Strong understanding of Kubernetes primitives and patterns (deployments, services, operators, etc.).
  • Experience with Infrastructure as Code (Terraform, Terragrunt, or similar).
  • Solid understanding of Linux systems and networking fundamentals.

Responsibilities

  • Design, build, and operate highly reliable, scalable infrastructure for a multi-tenant SaaS platform.
  • Automate deployment, provisioning, and operational workflows across cloud infrastructure and applications.
  • Develop internal tools, services, and frameworks to improve efficiency and reduce manual effort.
  • Participate in a 24x7 follow-the-sun on-call rotation to support production systems.
  • Collaborate with engineering teams to meet availability and performance targets.
  • Define and track SLOs and reliability metrics; lead incident response and blameless postmortems.

Skills

Python
Go
AWS
Kubernetes
Linux
Networking
Observability
CI/CD

Education

Bachelor's degree in Computer Science or related field

Tools

Datadog
Prometheus
OpenTelemetry
CloudWatch
Kafka

Job description

Drive Reliability, Automation & Scale:
  • Design, build, and operate highly reliable, scalable infrastructure for a multi-tenant SaaS platform.
  • Automate deployment, provisioning, and operational workflows across cloud infrastructure and applications.
  • Develop internal tools, services, and frameworks to improve efficiency and reduce manual effort.
  • Participate in a 24x7 follow-the-sun on-call rotation to support critical production systems.
Improve Platform & Infrastructure:
  • Contribute to core platform systems by building features, resolving issues, and enhancing reliability.
  • Partner with development teams to ensure systems meet availability, performance, and scalability requirements.
  • Proactively identify risks, bottlenecks, and failure modes, and implement solutions before they impact customers.
Observability, Incident Management & Resilience:
  • Build and maintain observability systems (metrics, logging, tracing, dashboards).
  • Define and track Service Level Objectives (SLOs) and reliability metrics.
  • Lead or contribute to incident response, root cause analysis, and blameless postmortems.
  • Drive improvements toward self-healing systems and reduced operational toil.
Security & Identity:
  • Design and support secure access patterns, including SSO, SAML, and OAuth-based authentication systems.
  • Ensure platform services meet security and compliance standards.
Enablement & Collaboration:
  • Collaborate across engineering teams, providing guidance, feedback, and hands‑on contributions.
  • Create and maintain documentation, runbooks, and training materials.
  • Mentor engineers and promote best practices in reliability engineering and automation.
Who You Are
Infrastructure Development:
  • 812 years of hands‑on experience in Site Reliability Engineering (SRE), DevOps, Cloud Infrastructure, or a related platform engineering role, with a proven track record of designing, building, and operating highly available, scalable, and reliable production systems.
  • Strong programming skills in Python or Go (Java/Spring Boot is a plus).
  • Deep experience with AWS and building/operating production systems at scale.
  • Hands‑on expertise with Kubernetes (EKS), Docker, Helm, CNI, and Ingress networking.
  • Strong understanding of Kubernetes primitives and patterns (deployments, services, operators, etc.).
  • Experience with Infrastructure as Code (Terraform, Terragrunt, or similar).
  • Solid understanding of Linux systems and networking fundamentals.
Observability & Operations:
  • Experience with observability platforms such as Datadog, Prometheus, OpenTelemetry, or CloudWatch.
  • Familiarity with incident management practices and production support in a microservices environment.
  • Experience with messaging/streaming systems (e.g., Kafka, SQS) and relational databases (e.g., Aurora, RDS) is a plus.
Security & Identity:
  • Working knowledge of SSO, SAML, OAuth, and identity providers (Okta is a plus).
  • Experience with AWS IAM (roles, policies, IRSA), VPC security groups, and Kubernetes security primitives (RBAC, network policies, pod security standards, secrets management).
DevOps & Delivery:
  • Experience with CI/CD and GitOps tools such as GitHub Actions, TeamCity, Jenkins, FluxCD, or Bitbucket.
  • Comfortable working in agile environments (Scrum, Kanban).
Mindset & Collaboration:
  • Strong troubleshooting and problem‑solving skills with a proactive, systems‑thinking mindset.
  • Passion for automation: “If you have to do it more than once, automate it.”
  • Excellent communication skills and ability to work across distributed teams.
  • A collaborative team player who can influence, mentor, and lead through technical expertise.
  • Demonstrated ability to leverage AI and data‑driven insights to improve productivity and outcomes.
Preferred Qualifications
  • Bachelor's degree in Computer Science or related field, or equivalent experience
  • Experience supporting large‑scale SaaS platforms
  • AWS or Kubernetes certifications
  • Exposure to modern platform frameworks such as KubeVela (OAM) or Crossplane
  • Contributions to open‑source projects
Why Guidewire?
  • Work on a mission‑critical global platform used by leading P&C insurers worldwide
  • Solve complex, real‑world infrastructure problems at genuine scale
  • Be part of a collaborative, high‑impact engineering culture grounded in integrity, rationality, and collegiality
  • Opportunity to shape the future of a rapidly evolving cloud platform
  • A culture of curiosity and innovation where engineers are empowered to leverage AI and emerging technologies.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Platform Engineer III (EKS and Python/Go)
Platform Engineer III (EKS and Python/Go)

Guidewire Software • Bengaluru

On-site
INR 2,400,000 - 4,200,000
Senior Platform Site Reliability Engineer (EKS and Python/Golang)
Senior Platform Site Reliability Engineer (EKS and Python/Golang)

FHLB Des Moines • Bengaluru

On-site
INR 2,400,000 - 3,600,000
Staff Software Engineer - Cloud Platform
Staff Software Engineer - Cloud Platform

FHLB Des Moines • Bengaluru

On-site
INR 4,500,000 - 6,500,000
Staff Software Engineer - Platform Engineering
Staff Software Engineer - Platform Engineering

Guidewire Software • Bengaluru

On-site
INR 3,000,000 - 6,000,000
Software Engineer II - Cloud Platform
Software Engineer II - Cloud Platform

Guidewire Software • Bengaluru

On-site
INR 1,200,000 - 2,000,000
Site Reliability Engineer
Site Reliability Engineer

United States Digital Space LLC • Karnataka

On-site
INR 900,000 - 1,200,000
Significant equity in a venture-backed company
Opportunity to work with modern tech stack
Senior Staff Software Engineer - Cloud Platform Engineering
Senior Staff Software Engineer - Cloud Platform Engineering

Guidewire Software Solutions India Private Limited • Bengaluru

On-site
INR 3,500,000 - 5,500,000
Senior Software Engineer - Full Stack Application Development
Senior Software Engineer - Full Stack Application Development

Guidewire Software • Bengaluru

On-site
INR 3,600,000 - 6,000,000
Software Engineer - Cloud Platform
Software Engineer - Cloud Platform

Guidewire Software Solutions India Private Limited • Bengaluru

On-site
INR 1,200,000 - 2,200,000
Platform Specialist
Platform Specialist

Ascendion • Bengaluru

On-site
INR 1,500,000 - 2,000,000
Innovation opportunities
Work with modern technology
Influence engineering strategy