Site Reliability Engineer

Zorba AI

Hyderabad

On-site

INR 4,000,000 - 7,500,000

Full time

9 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Zorba AI in Hyderabad, India seeks a Site Reliability Engineer to drive reliability practices across e-commerce platforms. You will join a cross-functional team, implement CI/CD, and manage cloud infrastructure using Azure/GCP with Terraform/Ansible.

The role requires 8–10 years of SRE experience and hands-on expertise in Kubernetes, monitoring tools, and incident response. The ideal candidate will contribute to production stability, participate in on-call rotations, and collaborate with product

Qualifications

  • Must-Have: cross-functional reliability expertise in a product area
  • Apply reliability engineering practices with SRE governance support
  • 5+ years in Site Reliability Engineering, maintenance & operations or development
  • Strong experience in eCommerce environments
  • DevOps practices including automated testing and CI/CD
  • Experience in solution architecture and rapid issue pinpointing
  • Experience with API-based frameworks (Commerce tools/Fabric)
  • Experience building CI/CD pipelines with GitHub Actions
  • Experience maintaining/developing desktop and mobile apps
  • Cloud experience with Azure and GCP
  • Infra as Code with Terraform/Ansible
  • ITSM/ITIL and ServiceNow in a microservices context
  • Monitoring with Splunk, Grafana, etc.
  • SRE metrics (SLI/SLO/Error Budget) knowledge
  • Managed cloud Kubernetes services (AKS/GKE)
  • Deliver quality and KPI reporting within product teams

Responsibilities

  • Should analyze critical issues and resolve them promptly
  • Willingness to take on development challenges and resolve them
  • Be a good team player and able to lead the team
  • Willingness to work in shifts
  • Focus on on-time delivery

Skills

Azure/GCP
.NET
DevOps Practices
ITIL
ReactJS
React Native
NodeJS
Terraform/Ansible
Python
C#
Kubernetes
Splunk
Dynatrace
GitHub Actions
Microservices

Job description

Job Description
  • mandatory

SN

Required Information

Details

1

Role**

Site Reliability Engineer

2

Required Technical Skill Set

Azure/GCP, .NET, DevOps Practices, ITIL, ReactJS, React Native, NodeJS, Terraform/Ansible, Python, C#, Kubernetes, Splunk, Dynatrace.

3

No of Requirements

15

4

Desired Experience Range

8years – 10 years

5

Location of Requirement

Chennai/Hyderabad/Kochi/Banglore

Desired Competencies (Technical/Behavioral Competency)

Must-Have

  • Work in a cross functional team working with Reliability as Expertise in a product or a product area.
  • Apply Reliability engineering practices with support from SRE governance teams.
  • 5+ years of experience in Site Reliability Engineering, maintenance & operations and/or development.
  • Strong working experience eCommerce.
  • Strong working experience in DevOps practices (automated testing, CI/CD etc.).
  • Experience within solutions architecture and how to fast pinpoint causes of issues.
  • Experience from working with API-based frameworks (e.g., Commerce tools or Fabric is ideal).
  • Experience in building CI/CD workflows using GitHub Actions.
  • Experience of maintaining/supporting and/or developing desktop and mobile applications.
  • Experience working on cloud-based infrastructure e.g., Azure and GCP.
  • Experience in provisioning Infra resources leveraging Infra as Code (Terraform / Ansible).
  • Experience from ITIL support processes and ITSM tools (e.g., ServiceNow) in a microservices context.
  • Experience in monitoring tools (Splunk, Grafana etc.).
  • Experience working through SRE Metrics such as SLI, SLO and Error Budget.
  • Experience with managed cloud Kubernetes services (e.g. AKS, GKE).
  • Ensure delivery quality and supply KPI reporting.
  • Collaborate closely within product teams to ensure predictable operations and minimal disruptions to Production.
  • Collaborate closely within your Capability, share best practices as well as discuss and improve on operations ways of working.
  • Technical analysis, troubleshooting of complex issues/Incidents in production.
  • Improve monitoring performance by focusing on preventive measures.
  • Product Improvements (code & log analysis).
  • Continuous improvement on proactive monitoring, housekeeping automation to proactively detect and avoid incidents.
  • Automate processes impacting development and production leveraging tools and building scripted solutions.
  • Participate in On-Call technical support to resolve business critical incidents.

Good to have

  • Familiarity with common tech stacks in Headless Ecommerce is a nice to have.
  • Knowledge of design principles and fundamentals of solutions architecture is a plus.
  • Understanding of performance engineering (Application Reliability).
  • Knowledge of multiple front-end languages and libraries (ReactJS, React Native, NodeJS).
  • Knowledge of Azure DevOps and/or other cloud environments is nice to have.
  • A passion for problem solving with strong analytical capabilities.
  • Stay current on technical trends to suggest innovative tools and approaches to interesting problems.
  • Knowledge on at least one of {Python, Ruby, Java, C#, Go} at an intermediate level.

SN

Responsibility of / Expectations from the Role

  1. Should be able to analyze the critical issue and resolve in time
  2. Willingness to accept development challenges and work towards resolving them
  3. Should be a good team player and should be able to lead the team
  4. Should be willing to work in shifts
  5. Should be focused in on-time delivery

Type

Details of The Role (For Candidate Briefing)

Reporting To Which Role

Site Reliability Engineer

Size of the Team, if any Reporting to this Role

Not Applicable

On-site Opportunity

No Visibility

Unique Selling Proposition (USP) of The Role

Opportunity to work in multiple technologies & directly with Customer.

Details of The Project (A short Briefing on the Project may be attached with this document for candidate- briefing). It may be shared with external stakeholders like job-agencies etc.

Providing technical leadership to teams and integrate technical expertise and business understanding to collaborate with Customer involved in Retail Business in all over the world

Skills: reliability,devops,azure,ansible

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Zorba AI • Chennai District

On-site
INR 4,000,000 - 7,500,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Hucon Solutions • Chennai District, Bengaluru, Hyderabad

On-site
INR 1,000,000 - 1,500,000
Tech & Digital-Site Reliability Engineer
Tech & Digital-Site Reliability Engineer

Hdfc Bank • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Site Reliability Engineer
Site Reliability Engineer

S P A Enterprise Info Services • Chennai District

On-site
INR 1,800,000 - 2,800,000
Lead SRE
Lead SRE

United States Digital Space LLC • Karnataka

On-site
INR 900,000 - 1,400,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

VMC Soft Technologies, Inc • Hyderabad

Hybrid
INR 1,500,000 - 2,000,000
Lead SRE
Lead SRE

UST • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Site Reliability Engineer
Site Reliability Engineer

Spot Your Leaders & Consulting • Pune District

On-site
INR 2,500,000 - 4,000,000
Site Reliability Engineer (SRE) L2
Site Reliability Engineer (SRE) L2

DigitalXNode • Hyderabad

Hybrid
INR 1,500,000 - 2,500,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

AcquireX • Maharashtra

On-site
INR 1,800,000 - 2,500,000