Site Reliability Engineer - SRE Intermediate

FinThrive, Inc.

India

On-site

INR 1,200,000 - 1,800,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Term life insurance
Meal and Transport arrangements

Job summary

FinThrive, Inc. is seeking a Site Reliability Engineer at an intermediate level to design, operate, and optimize cloud-native platforms with a strong Azure focus. You will drive automation-first initiatives, implement IaC, and contribute to reliable, scalable services across distributed systems.

The role emphasizes incident response, RCA leadership, and collaboration with cross-functional teams to reduce toil and improve availability and performance.

Qualifications

  • Bachelor’s degree in Computer Science or Engineering or related field.
  • Experience designing and operating cloud-native platforms, with Azure focus.

Responsibilities

  • Lead incident response and root-cause analysis (RCA).
  • Improve reliability with monitoring, tuning, and capacity planning.
  • Design resilient architectures using Azure services (App Services, ASEv3, AGW, Front Door).
  • Develop IaC and automation to reduce toil (Terraform/Bicep/ARM).
  • Collaborate with SRE, CloudOps, and development teams to improve pipelines.

Skills

Azure Cloud
SRE
Incident Management
Automation
Infrastructure as Code
Observability
CI/CD
AI-assisted tooling

Education

Bachelor's Degree in CS/Engineering

Tools

Terraform
Bicep
ARM templates
Azure Functions
Azure Monitor
Grafana
App Services
ASEv3
Application Gateway
Front Door
GitHub Copilot

Job description

Site Reliability Engineer - SRE Intermediate
Job Description

Posted Sunday, August 16, 2026 at 6:30 PM

Professional Summary

Site Reliability Engineer with 3–5+ years of experience in designing, operating , and optimizing cloud-native platforms with a strong focus on Azure environments . Proven expertise in building highly available , scalable, and secure systems using Infrastructure as Code ( IaC ) and automation-first practices.

Experienced in managing application hosting architectures including Azure App Services, ASEv3, Application Gateway (AGW) and Azure Front Door , ensuring high performance and resilience across distributed systems.

Demonstrates an automation mindset by leveraging modern engineering tools and AI-assisted development platforms (e.g., GitHub Copilot, Microsoft Copilot) to accelerate delivery, reduce operational toil, and improve reliability standards — with careful validation of outputs for security and production readiness.

Core Competencies

Understanding and experience in developing Azure function Apps, Azure logic Apps

Understanding of event triggers, event hub, service bus.

Cloud Architecture: High Availability, Fault Tolerance, Scalability Patterns

Incident Management and RCA

Incident Management, P1 troubleshooting, Change Management

Experienced in leading RCA and representing on the weekly call

SLA / SLO / Error Budget concepts

System Performance Optimization & Capacity Planning

Toil Reduction through Automation

Infrastructure as Code & Automation

API-based automation and orchestration

Observability & Monitoring

Azure Monitor, Log Analytics Workspace, Grafana, Site 24x7 (or similar SaaS based synthetic monitoring tool)

Application Insights

Alert tuning and signal-to-noise optimization

AI-Enabled Productivity (Not as Skill)

Code acceleration and script generation

Troubleshooting and log analysis assistance

Proven track record of workforce optimization leveraging AI tools.

Applying validation frameworks to ensure secure, accurate, and production-grade outputs

DevOps & Integration

Deep understanding on version control

API integrations (REST, Postman, SoapUI)

Source control and release management

Professional Experience

SRE & Reliability Engineering

Managed production environments ensuring high availability and reliability of cloud-hosted applications

Led incident response, performed deep root cause analysis , and implemented preventive measures to reduce recurrence

Improved system resilience through proactive monitoring and performance tuning strategies

Designed and supported application architectures using:

Azure App Services and App Service Plans

Azure App Service Environment v3 (ASEv3) for isolated, high-scale workloads

Azure Application Gateway (WAF-enabled) for L7 traffic management

Azure Front Door for global traffic routing and failover

Implemented secure and scalable cloud networking patterns , optimizing latency and throughput

Automation & Toil Reduction

Identified repetitive operational tasks and reduced manual effort through automation-first solutions

Developed automation using:

Terraform / Bicep / ARM templates

Azure Functions for event-driven workflows

Leveraged AI-assisted tools (GitHub Copilot, Copilot) to accelerate scripting and automation development, while ensuring strict validation for enterprise use

Observability & Monitoring

Built and enhanced observability using:

Created KQL-based queries and dashboards for proactive issue detection

Reduced false alerts by optimizing alert thresholds and improving signal quality

Performance & System Optimization

Analyzed application performance across distributed systems to identify bottlenecks

Implemented improvements through:

Scaling strategies (horizontal & vertical)

Network optimization (AGW / Front Door tuning)

Backend service improvements

Collaboration & Engineering Enablement

Partnered with SRE, CloudOps, and development teams to design resilient systems

Contributed to runbooks, documentation, and operational standards

Enabled engineering teams by improving platform reliability and deployment pipelines

Key Achievements
  • Reduced manual operational effort by X% through automation initiatives
  • Improved system availability to 99.X% by strengthening monitoring and failure handling mechanisms
  • Decreased incident resolution time by X% via enhanced observability and streamlined runbooks
  • Optimized application performance using Front Door and AGW tuning, reducing latency by X%
Education

Bachelor’s Degree in Computer Science / Engineering or related field

Preferred/Additional Experience

Experience with microservices and distributed architectures

Working knowledge of AWS cloud services

Preferred/Additional Certifications

AZ-700 Designing and Implementing Microsoft Azure Networking Solutions

About FinThrive

FinThrive is advancing the healthcare economy. For the most recent information on FinThrive’s vision for healthcare revenue management visit finthrive.com/why-finthrive

Award-winning Culture of Customer-centricity and Reliability

At FinThrive we’re proud of our agile and committed culture, which makes FinThrive an exceptional place to work. Explore our latest workplace recognitions at https://finthrive.com/careers#culture

Our Perks and Benefits
  • Term life, Accidental & Medical Insurance
  • Meal and Transport arrangements
FinThrive’s Core Values and Expectations
  • Demonstrate integrity and ethics in day-to-day tasks and decision-making, adhere to FinThrive’s core values of being Customer-Centric, Agile, Reliable, and Engaged, operate effectively in the FinThrive environment and the environment of the workgroup, maintain a focus on self-development and seek out continuous feedback and learning opportunities
  • Support FinThrive’s Compliance Program by adhering to policies and procedures about HIPAA, GLBA, FCRA, and other laws applicable to FinThrive’s business practices; this includes becoming familiar with FinThrive’s Code of Ethics, attending training as required, notifying management or FinThrive’s Helpline when there is a compliance concern or incident, HIPAA-compliant handling of patient information, and demonstrable awareness of confidentiality obligations.

FinThrive is an Equal Opportunity Employer and ensures its employment decisions comply with principles embodied in Title VII, the Age Discrimination in Employment Act, the Rehabilitation Act of 1973, the Vietnam Veterans Readjustment Assistance Act of 1974, Executive Order 11246, Revised Order Number 4, and applicable state regulations.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

33365-Senior Software Engineer
33365-Senior Software Engineer

FinThrive, Inc. • Pune District

On-site
INR 1,800,000 - 2,400,000
Term life insurance
Medical insurance
Meal and transport arrangements
33365-Associate Software Engineer
33365-Associate Software Engineer

FinThrive, Inc. • Pune District

On-site
INR 420,000 - 640,000
Term life, Accidental & Medical Insur
Meal and Transport arrangements
Software Engineer
Software Engineer

FinThrive Revenue Systems, LLC • Pune District

On-site
INR 600,000 - 1,200,000
Term life, Accidental & Medical Insurance
Meal and Transport arrangements
33365-Senior Software Engineer II
33365-Senior Software Engineer II

FinThrive, Inc. • India

On-site
INR 1,800,000 - 3,000,000
Term life insurance
Meal and transport arrangements
33365-Software Engineer_Dev Engineering
33365-Software Engineer_Dev Engineering

FinThrive, Inc. • Pune District

On-site
INR 800,000 - 1,200,000
Term life insurance
Medical Insurance
Meal and Transport arrangements
Cloud Network Engineer
Cloud Network Engineer

FinThrive Revenue Systems, LLC • India

On-site
INR 800,000 - 1,200,000
Term life insurance
Accidental & Medical Insurance
Meal and Transport arrangements
33365-Software Engineer
33365-Software Engineer

FinThrive, Inc. • India

On-site
INR 1,200,000 - 1,800,000
Term life, Accidental & Medical Insur
Meal and Transport arrangements
33365-Senior Software Quality Engineer
33365-Senior Software Quality Engineer

FinThrive, Inc. • Pune District

On-site
INR 900,000 - 1,400,000
Meal and Transport arrangements
Term life, Accidental & Medical Ins
Senior Software Engineer
Senior Software Engineer

FinThrive, Inc. • Pune District

On-site
INR 1,200,000 - 1,800,000
Term life Insurance
Medical Insurance
Meal and Transport arrangements
33365-Senior Software Quality Engineer
33365-Senior Software Quality Engineer

FinThrive • Gurugram District

On-site
INR 900,000 - 1,200,000
Professional development opportunities
Meal and Transport arrangements
Insurance