Senior Site Reliability Engineer

ACG World

Mumbai

On-site

INR 2,200,000 - 3,200,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

ACG World in Mumbai, India, seeks an Azure Cloud Infrastructure & Site Reliability Engineer to own the Azure cloud environment, drive SRE practices, and enforce a Code-First delivery approach across software, QA, and hardware teams.

You will architect and manage scalable Azure environments, configure IaC with Terraform or Bicep, implement monitoring with Azure Monitor, and lead incident management with blameless post-mortems while reducing toil through automation.

Qualifications

  • Total 6+ years of experience in Azure infrastructure management.
  • Hands-on experience with Azure services (VMs, AKS, App Service, Azure SQL, Functions).
  • Strong knowledge of CI/CD tools and DevOps practices.
  • Familiarity with project management tools like Jira and Azure Boards.

Responsibilities

  • Architect and manage scalable Azure environments including App Service, AKS, Cosmos DB, and SQL.
  • Lead IaC adoption with Terraform or Bicep and strict state management.
  • Own the Code-First lifecycle: enforce zero-manual-change policies in production.
  • Coordinate hardware-software integration with Mechanical/Operations teams.
  • Define and maintain SLIs/SLOs for all customer-facing services; publish monthly error budgets.
  • Own end-to-end incident management: detection, triage, war room, post-mortems.
  • Track MTTR/MTTD and change failure rates; present trends to leadership.
  • Drive toil reduction: automate at least 30% per quarter.

Skills

Azure
CI/CD
Terraform
Bicep
Python
PowerShell
Docker
Kubernetes
Azure DevOps
Git
Azure Monitor
Security
RBAC
Ansible
Networking

Tools

Terraform
Bicep
Ansible
Git
GitHub Actions
Azure Pipelines
Jenkins
Azure DevOps
Jira
Azure Boards
PowerShell

Job description

ROLE TITLE: Azure Cloud Infrastructure & Site Reliability Engineer

This is a cloud infrastructure and operations-first role , the candidate must be a technical bridge between software development, QA, and mechanical hardware teams. You will own the Azure cloud environment, drive SRE practices (SLIs/SLOs), and enforce "Code-First" delivery standards to reduce toil and ensure hardware-software integration stability.

Primary responsibilities
  • Architect and manage scalable Azure environments (App Service, APIM, Cosmos DB, AKS, SQL).
  • Manage hybrid connectivity and hardware integration: VNets, NSGs, Private Endpoints, and ExpressRoute/VPN.
  • Lead IaC adoption: Maintain production-grade Terraform or Bicep modules with strict state management.
  • Own the "Code-First" lifecycle: Enforce zero-manual-change (Click-Ops) policies in production.
  • Coordinate hardware-software integration: Validate compatibility through simulations and real-time testing in collaboration with Mechanical/Operations teams.
  • Define and maintain SLIs/SLOs for all customer-facing services; publish error budgets monthly.
  • Own end-to-end incident management: detection, triage, war room coordination, and blameless post-mortems.
  • Track MTTR, MTTD, and change failure rates; present performance trends to engineering/CTO leadership.
  • Drive toil reduction: Identify manual operational bottlenecks and automate at least 30% per quarter.
Monitoring, Observability & Security
  • Own the monitoring stack: Azure Monitor, Log Analytics, Application Insights, Prometheus, and Grafana.
  • Implement distributed tracing and robust alerting; eliminate "noisy" false-positive alerts.
  • Enforce security/compliance (ISO, GxP, GDPR): Implement RBAC, Managed Identities, and Azure Key Vault policies.
  • Use Microsoft Defender for Cloud and Sentinel to conduct vulnerability assessments and risk mitigation.
DevOps, Automation & Agile Delivery
  • Maintain Azure DevOps (YAML) pipelines for infrastructure, environment refresh, and config drift correction.
  • Translate business requirements into technical deliverables and work items in Azure Boards.
  • Act as the infrastructure approver for all UAT and Production deployments.
  • Scripting & Automation: Develop production-grade automation in Python, Bash, or PowerShell for backup verification, cert rotation, and log archival.
Collaboration and Support:
  • Work closely with development, QA, and operations teams to ensure smooth delivery of applications.
  • Provide guidance and support to development teams on best practices for cloud architecture and DevOps.
  • Conduct training and knowledge-sharing sessions for team members.
Documentation and Reporting:
  • Maintain comprehensive documentation of infrastructure, configurations, and procedures.
  • Generate reports and dashboards to provide visibility into system performance, deployments, and issues.
  • Knowledge and understanding of ITIL
Operating System:
  • Windows and Linux Proficiency
Certifications
Experience:
  • Total 6 + Years
  • Proven experience in Azure infrastructure management and DevOps practices.
  • Hands-on experience with Azure services (e.g., VMs, AKS, App Services, Azure SQL, Functions, etc.).
  • Strong knowledge of CI/CD tools and practices.
  • Familiarity with project management tools like Jira, Azure Boards.
Technical Skills:
  • CI/CD and Automation Tools: Strong hands-on experience with Git, Jenkins, GitHub Actions, Azure Pipelines, and automation tools like Ansible.
  • PaaS/SaaS Depth: Hands-on experience with Azure PaaS (APIM, Cosmos DB, Functions, AKS) and SaaS service integration
  • IaC & Automation: Production-grade Terraform or Bicep; experience with Ansible for configuration management.
  • Scripting and Automation: PowerShell, Python, Bash, YAML for task automation and infrastructure management.
  • Containerization and Orchestration: Docker, Kubernetes (AKS) for container management and orchestration.
  • Monitoring and Logging: Experience with Azure Monitor, Application Insights, Log Analytics, Sentinel for performance monitoring and alerting.
  • Security and Compliance: Familiarity with Microsoft Defender for Cloud, Azure Security Center, OWASP, Static Code Analysis tools, and implementing RBAC, secure networking, and identity management.
  • Version Control: Git, GitHub, Azure Repos for version control and collaboration.
  • Databases: Hands-on experience with Azure SQL, Cosmos DB, MySQL, MSSql.
  • Operating Systems: Windows Server and Linux (Ubuntu, Red Hat).
  • Networking: configuring Virtual Networks, Load Balancers, VPN Gateways, NSG Rules, and implementing cloud networking best practices.
  • DevOps and Agile Methodologies: Strong understanding of agile development, cloud automation, and DevOps practices.
Soft Skills:
  • Excellent problem-solving and troubleshooting skills.
  • Strong communication and collaboration abilities.
  • Ability to work in a fast-paced, dynamic environment.
  • Keep good communication and coordination with cross functional teams.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer (Azure) - S
Senior Site Reliability Engineer (Azure) - S

Tata Consultancy Services • Kolkata District, Chennai District, Bengaluru

On-site
INR 2,400,000 - 4,200,000
Site Reliability Engineer (Azure) - S
Site Reliability Engineer (Azure) - S

Tata Consultancy Services • Bengaluru

On-site
INR 2,500,000 - 4,500,000
Azure Senior Site Reliability Engineer
Azure Senior Site Reliability Engineer

Elabs Infotech • Bengaluru

Hybrid
INR 2,500,000 - 4,000,000
Senior Site Reliability Engineering (Azure Cloud)
Senior Site Reliability Engineering (Azure Cloud)

Cvent • Bengaluru

On-site
INR 3,500,000 - 6,000,000
SRE / Infrastructure DevOps Engineer
SRE / Infrastructure DevOps Engineer

Exiga Solutions • Chennai District, Bengaluru, Gurugram District, Hyderabad

On-site
INR 1,500,000 - 2,100,000
Site Reliability Engineer
Site Reliability Engineer

PwC India • Bengaluru

On-site
INR 1,500,000 - 2,500,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Cvent • Bengaluru

On-site
INR 3,500,000 - 6,500,000
Senior Devops And Cloud Infrastructure Engineer
Senior Devops And Cloud Infrastructure Engineer

Bigsun Technologies • Navi Mumbai

On-site
INR 3,000,000 - 6,000,000
Site Reliability Engineer
Site Reliability Engineer

MishiPay • Bengaluru

On-site
INR 2,500,000 - 3,800,000
Technical Lead
Technical Lead

Infinite Computer Solutions • Chennai District

On-site
INR 1,800,000 - 3,200,000