Senior Site Reliability Engineer

Clearwater Analytics

Mumbai

On-site

INR 4,000,000 - 7,000,000

Full time

9 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Beacon by Clearwater is seeking a seasoned software/DevOps engineer to build internal tools and automation for a fleet of client deployments across AWS and Azure. You will standardize environments, improve observability, and automate remediation and provisioning tasks.

The role emphasizes writing production‑grade Python, Terraform IaC, and close collaboration with client-facing teams. The position requires 7–10 years of experience in software engineering or DevOps, strong Linux skills, and a

Qualifications

  • 7–10 years of experience in software engineering, site reliability engineering, DevOps, or platform engineering.
  • Strong production‑quality Python coding with tests.
  • Hands‑on experience with AWS or Azure cloud platforms: networking, IAM, storage, compute.
  • Experience with infrastructure‑as‑code (Terraform) and managing multiple environments.
  • Solid Linux fundamentals: log reading, debugging, automation to avoid manual toil.
  • Automation mindset: build tools when a task is repeated.
  • Collaborative, service‑oriented mindset supporting internal teams.

Responsibilities

  • Build internal tools and automation in Python to monitor and support client deployments across AWS and Azure.
  • Drive standardization of client environments and drift remediation.
  • Improve fleet observability with monitoring, alerts, and dashboards across deployments.
  • Convert runbooks into automated checks, self‑healing jobs, and one‑click tools.
  • Extend provisioning and deployment pipelines to accelerate onboarding of new clients.
  • Collaborate with onboarding, support, and client success teams to reduce toil.

Skills

Python
Cloud AWS/Azure
Linux fundamentals
Automation
Observability

Tools

Terraform
Git

Job description

About The Team

Beacon by Clearwater is the AI‑powered risk analytics and modeling arm of the Clearwater platform, giving institutional investors the tools to test scenarios and evaluate portfolio exposures in real time.

As Clearwater brings Beacon to more clients, the number of client environments we provision, monitor, and support grows with it and the only way that works is through standardization and automation. This team builds the tooling that keeps a growing fleet of client deployments consistent, observable, and supportable: automating away repetitive operational work, turning incident learnings into permanent platform fixes, and giving client‑facing teams the self‑service tools they need to onboard and support clients without engineering escalations.

What You’ll Do
  • Build internal tools and automation primarily in Python to monitor, diagnose, and support a fleet of client deployments across AWS and Azure.
  • Drive standardization across client environments: detect and remediate configuration and infrastructure drift, converge legacy deployments onto golden paths, and make “the standard way” the easy way.
  • Improve fleet‑wide observability: build monitoring, alerting, and dashboards that surface problems across all client deployments before clients notice them.
  • Turn runbooks into code; converting the manual diagnostic and remediation steps support engineers perform today into automated checks, self‑healing jobs, and one‑click tools.
  • Extend the client provisioning and deployment pipeline (Terraform, configuration generation) to make onboarding new clients faster and more repeatable.
  • Work directly with client‑facing teams (onboarding, support, client success) to find where operational toil lives.
What We’re Looking For
  • 7-10 years of experience in software engineering, site reliability engineering, DevOps, or platform engineering.
  • Strong programming skills in Python (our platform core and tooling language); comfort writing production‑quality code with tests, not just scripts.
  • Hands‑on experience with at least one major cloud provider (AWS or Azure): networking (VPCs/VNets, subnets, security groups, load balancers, VPN), IAM/RBAC, storage, and compute.
  • Working knowledge of infrastructure‑as‑code, ideally Terraform, and what it means to manage many environments from shared modules and per‑environment configuration.
  • Solid Linux fundamentals: you can read logs, trace a process, debug a service that won’t start, and automate what you did, so no one must do it by hand again.
  • An automation reflex: when you solve a problem twice, your instinct is to build a tool.
  • A collaborative, service‑oriented mindset: your customers are internal teams, and your success is measured by how much easier you make their jobs.
Nice to Have
  • Experience operating multi‑tenant or fleet‑style environments (many similar deployments managed as one).
  • Observability stack experience (metrics, log aggregation, alerting, dashboards).
  • Formal incident management experience (on‑call, post‑mortems, blameless RCA culture).
  • Exposure to financial services, fintech, or other regulated environments.
Why This Role
  • Direct, visible impact: every tool you ship makes onboarding the next client faster and supporting every existing client cheaper. This team is a force multiplier for the entire Beacon business.
  • Breadth: you’ll touch cloud infrastructure, a large Python platform codebase, deployment pipelines, and the human workflows of support and onboarding teams.
  • Growth: you’ll work across nearly every layer of a sophisticated financial‑engineering platform, alongside experts in cloud infrastructure, quantitative finance, and large‑scale SaaS operations.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

United States Digital Space LLC • Maharashtra

On-site
INR 1,000,000 - 1,500,000
Site Reliability Engineer
Site Reliability Engineer

Clearwater Analytics India Private Limited • Mumbai

On-site
INR 800,000 - 1,200,000
Site Reliability Engineer
Site Reliability Engineer

Clearwater Analytics, Ltd • Mumbai

On-site
INR 1,000,000 - 1,500,000
Director of Engineering
Director of Engineering

Beacon.li • Hyderabad

On-site
INR 2,000,000 - 3,000,000
Competitive compensation and benefits
Mentorship opportunities
Work on complex problems in AI
Site Reliability Engineer
Site Reliability Engineer

United States Digital Space LLC • Karnataka

On-site
INR 900,000 - 1,200,000
Significant equity in a venture-backed company
Opportunity to work with modern tech stack
Sr. Software Development Engineer
Sr. Software Development Engineer

Clearwater Analytics, Ltd • Mumbai, Dadri

On-site
INR 4,000,000 - 6,000,000
Software Development Engineer III - DevOps Engineer
Software Development Engineer III - DevOps Engineer

Safe Security • Bengaluru

On-site
INR 4,000,000 - 6,000,000
Site Reliability Engineer
Site Reliability Engineer

HighRadius • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Platform Engineer
Platform Engineer

United States Digital Space LLC • Maharashtra

On-site
INR 1,500,000 - 2,800,000
Lead DevOps Engineer
Lead DevOps Engineer

Caizin, LLC • Pune District

On-site
INR 1,400,000 - 1,800,000