SRE Engineer II/III

Rackspace Technology

Gurugram District

On-site

INR 1,500,000 - 2,500,000

Full time

8 days ago
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Rackspace Technology in Hyderabad is seeking a Site Reliability Engineer – L2/L3 to maintain the reliability, availability, and performance of production environments while partnering with engineering and operations teams.

You will own incident response, perform root cause analysis, and reduce toil through automation, scripting, and tooling. This role requires strong Linux, cloud, networking, and observability skills, plus a bias toward continuous improvement.

Qualifications

  • Strong Linux administration and file system management.
  • Cloud platforms experience (AWS/Azure/GCP) with scripting basics.
  • Proficient in Python and Bash for automation.
  • Hands-on networking knowledge, TCP/IP, DNS, TLS, and firewalls.
  • Experience with observability tools and incident response.

Responsibilities

  • Own L2/L3 incident response, RCA, and post-mortems for production issues.
  • Monitor system health and ensure SLA/SLO adherence.
  • Automate operational tasks to reduce toil and manual work.
  • Collaborate with development teams on deployment reliability and capacity planning.
  • Participate in on-call rotation and maintain runbooks.

Skills

Linux
Windows Server
Cloud
Terraform
Python
Bash
Ansible
TCP/IP
HTTP
Observability
Kubernetes
Docker
Kafka/RabbitMQ
MySQL/PostgreSQL/Redis
ITIL/ITSM

Tools

Terraform
Ansible
Nginx
HAProxy
tcpdump
Wireshark
Prometheus
Grafana
Datadog
ELK
Kubernetes
Docker
Jira SM
ServiceNow
Postman

Job description

Hiring: Site Reliability Engineer

Location: Hyderabad
Work Mode: Work from Office
Work Shift: 24/7
Experience: 3–8 Years
Level: L2/L3 Engineer

About the Role

We are looking for a Site Reliability Engineer – L2/L3 to join our team in Hyderabad. The ideal candidate will be responsible for maintaining the reliability, availability, and performance of production environments while working closely with engineering and operations teams. You serve as an escalation point, drive root cause analysis, and reduce toil through scripting and tooling.

Key Responsibilities
  • Own L2/L3 incident response, RCA, and post-mortems for production issues.

  • Monitor system health and maintain SLA/SLO adherence.

  • Automate operational tasks to eliminate repetitive toil.

  • Collaborate with dev teams on deployment reliability and capacity planning.

  • Participate in on-call rotation and maintain runbooks.

Required Skills:
Operating Systems

Hands-on with Linux (RHEL/Ubuntu) — system, process management, file systems, performance tuning. Working knowledge of Windows Server and event log analysis.

Cloud

Practical experience on AWS / Azure / GCP — compute, storage, IAM, networking, and managed services. Familiarity with Terraform or equivalent IaC tools.

Scripting & Automation

Proficiency in Python and Bash for automation, API interaction, and operational tooling. Exposure to Ansible or similar config management is a plus.

Network Troubleshooting (In-Depth)

Strong command of TCP/IP internals — handshake lifecycle, connection states (TIME_WAIT, CLOSE_WAIT, SYN_FLOOD), packet flow, and socket behavior. Hands-on with tools like tcpdump, Wireshark, netstat/ss, traceroute, mtr, and dig. Solid understanding of DNS resolution, TLS/SSL negotiation, NAT, firewalls, and routing. Able to diagnose latency, packet loss, port exhaustion, and network-level bottlenecks at the OS and infrastructure layer.

Application & HTTP Troubleshooting

Deep understanding of HTTP/HTTPS methods, status codes, headers, and request lifecycle. Comfortable debugging through curl, Postman, access logs, and reverse proxy configs (Nginx / HAProxy).

Observability

Experience with Prometheus, Grafana, Datadog, or ELK. Ability to build dashboards, configure meaningful alerts, and trace issues end-to-end.

Good to Have
  • Kubernetes / Docker experience.

  • Familiarity with message queues (Kafka, RabbitMQ).

  • Basic database troubleshooting (MySQL / PostgreSQL / Redis).

  • ITIL fundamentals and ITSM tools (Jira SM / ServiceNow).

Soft Skills

Strong analytical thinking, clear communication under pressure, and a bias toward automation and continuous improvement.

About Rackspace Technology

We are the multicloud solutions experts. We combine our expertise with the world’s leading technologies — across applications, data and security — to deliver end-to-end solutions. We have a proven record of advising customers based on their business challenges, designing solutions that scale, building and managing those solutions, and optimizing returns into the future. Named a best place to work, year after year according to Fortune, Forbes and Glassdoor, we attract and develop world-class talent. Join us on our mission to embrace technology, empower customers and deliver the future.

More on Rackspace Technology

and want you to know that we are committed to offering equal employment opportunity without regard to age, color, disability, gender reassignment or identity or expression, genetic information, marital or civil partner status, pregnancy or maternity status, military or veteran status, nationality, ethnic or national origin, race, religion or belief, sexual orientation, or any legally protected characteristic. If you have a disability or special need that requires accommodation, please let us know.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

SRE Engineer II/III
SRE Engineer II/III

Rackspace Technology • Hyderabad

On-site
INR 1,000,000 - 1,800,000
SRE Engineer II
SRE Engineer II

Webhosting • Hyderabad

On-site
INR 1,400,000 - 2,000,000
SRE Engineer II/III
SRE Engineer II/III

Webhosting • Hyderabad

On-site
INR 1,200,000 - 2,200,000
SRE Engineer II
SRE Engineer II

Rackspace Technology • Hyderabad

On-site
INR 1,800,000 - 2,800,000
SRE Engineer II/III
SRE Engineer II/III

Pace Industries, LLC • Hyderabad

On-site
INR 1,500,000 - 2,300,000
Site Reliability Engineer
Site Reliability Engineer

Arch Systems • Hyderabad

On-site
INR 2,800,000 - 4,200,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

VMC Soft Technologies, Inc • Hyderabad

Hybrid
INR 1,500,000 - 2,000,000
Site Reliability Engineer
Site Reliability Engineer

Acesoft Labs • Hyderabad

Hybrid
INR 1,800,000 - 3,000,000
Site Reliability Engineer
Site Reliability Engineer

Acesoft Labs • Ahmedabad District

Hybrid
INR 400,000 - 700,000
SRE Engineer
SRE Engineer

Mphasis • Hyderabad

On-site
INR 1,200,000 - 2,400,000