Delivery Lead-SRE

Acuity Analytics

Bengaluru

On-site

INR 2,400,000 - 3,800,000

Full time

10 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Acuity Analytics seeks a Site Reliability Engineer to safeguard reliability and resilience of critical platforms. You will build observability, automate toil, and partner with engineering for production readiness across batch cycles and live trading services.

Ideal for experienced engineers with strong Linux, cloud, and container expertise, focusing on end-to-end observability and incident response in a capital markets environment.

Qualifications

  • Experience with end-to-end reliability engineering for critical systems.
  • Strong observability, monitoring, alerting and incident management skills.
  • Exposure to cloud platforms (AWS/Azure) and Linux fundamentals.
  • Hands-on scripting and automation for CI/CD and IaC.

Responsibilities

  • Define service-health metrics, monitoring, alerting and dashboards.
  • Support incident response, problem management and post-incident reviews.
  • Identify and remediate reliability, performance and resilience risks.
  • Monitor batch/EOD cycles and time-critical processes for failures.
  • Engineer automation to reduce toil and improve MTTR.
  • Collaborate with engineering to improve production readiness before release.
  • Contribute to SLI/SLOs and error-budget management.
  • Support Kubernetes-based platforms and platform reliability.

Skills

Observability
Prometheus
Grafana
ELK
Datadog
Dynatrace
AppDynamics
SRE practices
Incident management
Automation scripting
Python
Bash/PowerShell
CI/CD
Infrastructure as code
AWS
Azure
Linux
Kubernetes
Docker
Batch processing
Disaster recovery

Education

Bachelor’s / Master’s degree

Tools

Docker
Kubernetes
Autosys
PagerDuty
Kafka
MQ
Datadog

Job description

We are seeking a Site Reliability Engineer to safeguard the reliability, observability and resilience of critical platforms supporting a Global Capital Markets platform-modernization program. The role ensures that Front

Office Trading systems and the emerging target-state services - spanning event-driven integration, data-lake /

Databricks platforms, reporting, migration and the front-to-back trade lifecycle for FX Cash, Cleared IRS and future products - run reliably across the trading day and its critical batch cycles. The ideal candidate blends strong engineering with an operations mindset, automating toil, engineering for resilience, and improving supportability before services reach production.

Key responsibilities
  • Define and implement service-health metrics, monitoring, alerting and dashboards to provide end-to-end observability.
  • Support incident response, problem management and operational readiness, including root-cause analysis and post-incident reviews.
  • Identify and remediate reliability, performance and resilience risks across critical trading and post trade services.
  • Monitor batch / EOD cycles and time-critical processes, detecting failures, delays and job / queue issues early.
  • Engineer automation to reduce operational toil (self-healing, runbooks, scheduling) and support capacity planning and performance tuning & and improve MTTR.
  • Partner with engineering teams to improve supportability, deployment safety and production readiness before release.
  • Contribute to SLIs / SLOs, error-budget management and continuous improvement of operational resilience and BCP.
  • Support Kubernetes-based platforms including monitoring cluster health, workload performance and platform reliability.
  • Bachelor’s / Master’s degree with 8+ years in SRE / production-engineering / DevOps roles supporting business-critical, high-availability systems.
  • Strong observability skills — monitoring, alerting and dashboards using tools such as Prometheus/Grafana, ELK, Splunk, Datadog, Dynatrace, AppDynamics or Cloud-native monitoring platforms.
  • Show more lines
  • Incident and problem management experience, including on-call, escalation and RCA; scheduling tools (e.g., Autosys) and alerting (e.g., PagerDuty).
  • Automation and scripting (Python / Bash / PowerShell), with CI/CD and infrastructure-as-code familiarity.
  • Cloud experience (AWS and/or Azure), strong Linux fundamentals, and understanding of resilience, performance and capacity engineering.
  • Strong analytical, communication and cross-team collaboration skills, with attention to detail.
  • Understanding of disaster recovery (DR), high availability (HA), failover mechanisms and resilience engineering practices.
Preferred qualifications
  • Experience defining and operating against SLIs / SLOs and error budgets.
  • Exposure to containers / orchestration (Docker / Kubernetes) and streaming / event-driven platforms (Kafka / MQ).
  • Prior experience in Capital Markets or financial services, including awareness of settlement and stress-test batch processing.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

SRE Trade Support
SRE Trade Support

Gemini Solutions Pvt. ltd. • Gurugram District

On-site
INR 1,200,000 - 1,500,000
Site Reliability Engineer
Site Reliability Engineer

AssetMark Global Wealth • Hyderabad

On-site
INR 1,800,000 - 3,600,000
Platform Engineer-SRE
Platform Engineer-SRE

INDmoney • Gurugram District

On-site
INR 3,000,000 - 6,000,000
Associate Site Reliability Engineer
Associate Site Reliability Engineer

AssetMark Global Wealth • Hyderabad

Hybrid
INR 1,200,000 - 1,800,000
Senior SRE/Support Engineer
Senior SRE/Support Engineer

Gemini Solutions Pvt. ltd. • Panchkula

On-site
INR 1,200,000 - 1,800,000
Site Reliability Engineer I
Site Reliability Engineer I

CME Group • Bengaluru

On-site
INR 600,000 - 900,000
SRE Engineer @ Investment Banking | Mumbai
SRE Engineer @ Investment Banking | Mumbai

Net Connect Global • Bengaluru, Mumbai

Hybrid
INR 1,800,000 - 2,400,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Sierra Ventures • Bengaluru

On-site
INR 3,500,000 - 5,500,000
Senior SRE Engineer
Senior SRE Engineer

Epam Systems • Bengaluru

On-site
INR 2,500,000 - 4,200,000
Site Reliability Engineer - Vice President
Site Reliability Engineer - Vice President

Citi • Pune District

On-site
INR 4,000,000 - 5,500,000