Site Reliability Engineer (SRE)

Bahwan CyberTek

Hyderabad

On-site

INR 1,200,000 - 1,800,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Bahwan CyberTek is seeking an SRE to improve reliability and operability of PAH-supported systems. You will implement observability, drive incident management, and reduce toil through automation.

The role requires strong DevOps practices and collaboration with engineering, architecture, and product teams to enforce production readiness. Responsibilities include capacity planning, performance analysis, and mature release controls.

Qualifications

  • Bachelor's degree in Computer Science, Engineering, or equivalent practical experience.
  • 5+ years experience in reliability engineering, DevOps, production operations, or software engineering with on-call responsibilities.
  • Demonstrated experience improving production reliability through automation, monitoring, and incident/problem management.

Responsibilities

  • Define, implement, and maintain observability (monitoring, logging, tracing) and actionable insights aligned to service health.
  • Drive incident management: on-call readiness, triage, incident command support, communications, and post-incident reviews (RCA).
  • Reduce operational toil through automation (runbooks-to-automation, self-healing, deployment/rollback automation).
  • Establish reliability standards: SLOs/SLIs, error budgets, production readiness reviews, and release risk controls.
  • Performance and reliability engineering: capacity planning, load/performance analysis, resilience testing, and failure-mode mitigation.
  • Partner with engineering teams to improve operational hygiene (deployability, rollback strategy, configuration, secrets, dependency management).

Skills

SRE/DevOps
Automation
Monitoring
Incident management
Scripting
Cloud platforms
ETL/ELT
CI/CD
Capacity planning
On-call

Education

Bachelor's degree in CS/Engineering

Tools

Azure DataBricks
Unity Catalog
AWS S3
AWS RDS

Job description

The Site Reliability Engineer (SRE) is responsible for improving the reliability, availability, performance, and operability of PAH-supported software systems. This role combines software engineering and IT operations to automate operational work, monitor system performance, and reduce toil. The SRE establishes and manages monitoring, ing, incident response, and problem management practices to ensure applications remain available and performant during updates and failures. The role partners with engineering, architecture, and product teams to define reliability standards and production readiness requirements. SRE is a practical implementation of DevOps focused on maintaining software quality in fast-paced development environments.

Define, implement, and maintain observability (monitoring, logging, tracing) and actionable ing aligned to service health.

Drive incident management: on-call readiness, triage, incident command support, communications, and post-incident reviews (RCA)
Reduce operational toil through automation (runbooks-to-automation, self-healing, deployment/rollback automation).

Establish reliability standards: SLOs/SLIs, error budgets, production readiness reviews, and release risk controls

Performance and reliability engineering: capacity planning, load/performance analysis, resilience testing, and failure-mode mitigation

Partner with engineering teams to improve operational hygiene (deployability, rollback strategy, configuration, secrets, dependency management)

Required Skills
  • Bachelor's degree in Computer Science, Engineering, or equivalent practical experience (preferred).
  • 5+ years experience in reliability engineering, DevOps, production operations, or software engineering with on-call responsibilities (preferred).
  • Demonstrated experience improving production reliability through automation, monitoring, and incident/problem management.
Required
  • Strong grounding in SRE/DevOps practices: incident management, blameless postmortems, SLOs/SLIs, error budgets, production readiness.
  • Experience building/operating monitoring and ing, and using logs/metrics to diagnose issues.
  • Automation/scripting skills (e.g., Python, PowerShell, Bash) and ability to reduce manual operational work.
  • Strong understanding of cloud-based platforms such as Azure DataBricks + Unity Catalog, AWS S3 and RDS.
  • Strong experience in ETL / ELT work.
  • Understanding of CI/CD concepts, safe deployment patterns, rollback strategies, and change risk controls.
Preferred
  • Experience with cloud environments and infrastructure-as-code.
  • Experience with large datasets (Multi-million row datasets).
  • Experience with container orchestration and modern runtime platforms (where applicable).
  • Experience building dashboards and reliability reporting for executives and delivery teams.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Bahwan CyberTek • Chennai District, Coimbatore District, Bengaluru

Hybrid
INR 1,800,000 - 2,400,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

C1X • Chennai District

On-site
INR 1,800,000 - 3,200,000
Site Reliability Engineer
Site Reliability Engineer

Spot Your Leaders & Consulting • Pune District

On-site
INR 2,500,000 - 4,000,000
Site Reliability Engineer (SRE) – DevOps Infrastructure
Site Reliability Engineer (SRE) – DevOps Infrastructure

PQAngels Technologies Pvt. Ltd. • Bengaluru

On-site
INR 1,200,000 - 2,100,000
Site Reliability Engineer
Site Reliability Engineer

Snapmint • Gurugram District

On-site
INR 800,000 - 1,200,000
SRE - AWS, GCP & Azure
SRE - AWS, GCP & Azure

PibyThree • Thane

On-site
INR 1,200,000 - 1,500,000
Site Reliability Engineer
Site Reliability Engineer

Recro • Bengaluru

On-site
INR 3,500,000 - 6,500,000
SRE - AWS, GCP & Azure
SRE - AWS, GCP & Azure

PibyThree • Navi Mumbai

On-site
INR 1,200,000 - 1,800,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Embarkgcc Services • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Infosys • Bengaluru

On-site
INR 900,000 - 1,500,000