Site Reliability Engineer - AWS (4 to 8 Years)

PHONEPE LIMITED

Bengaluru

On-site

INR 1,500,000 - 3,000,000

Full time

4 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Medical Insurance
Critical Illness Insurance
Accidental Insurance
Life Insurance
Wellness Program
Maternity Benefit
Paternity Benefit
Adoption Assistance
Day-care Support
Relocation Benefits
Transfer Support
Travel Policy
PF Contribution
Gratuity
NPS
Leave Encashment
Higher Education Assistance
Car Lease
Salary Advance Policy

Job summary

PhonePe Limited is hiring a Site Reliability Engineer (AWS) in Bengaluru. You will manage Linux systems, EC2 instances, and cloud components to ensure high availability. The role focuses on automation, security guardrails, and scalable infrastructure, with on-call rotation and incident RCA duties.

The ideal candidate has 4–8 years in SRE/DevOps, deep AWS experience, containerization skills (Docker/Podman), and strong networking knowledge. Excellent English communication is required.

Qualifications

  • 4 to 8 years in an SRE, DevOps, or Systems Engineering role.
  • Strong hands-on Linux/RHEL administration and troubleshooting.
  • Deep, hands-on experience exclusively within the AWS ecosystem.
  • Mandatory hands-on experience with containerization using Docker, Podman, or equivalent.
  • Solid grasp of the TCP/IP stack, DNS, routing concepts.
  • Understanding of SSL/TLS certificate lifecycle management.
  • Ability to communicate effectively in English, both written and verbally.

Responsibilities

  • Perform daily administration, configuration, and troubleshooting of Linux systems (RHEL).
  • Provision, configure, and maintain AWS EC2 instances and cloud-native components (IAM, Load Balancers, S3, CloudWatch).
  • Proactively monitor for vulnerabilities and implement security guardrails.
  • Configure and maintain AWS VPCs, Route Tables, and Security Groups; troubleshoot networking issues.
  • Collaborate on infrastructure migrations and runbook-driven upgrades.
  • Write Ansible playbooks and SaltStack states; automate tasks using Python/Go/Bash.
  • Manage MySQL and Aerospike data stores, including backups and scaling.
  • Build and extend platform observability with dashboards and alerts.
  • Create runbooks and SOPs for operational consistency.
  • Participate in on-call rotation and draft RCA reports.

Skills

AWS
Linux RHEL
Docker
Networking
Python/Go/Bash
Ansible

Tools

SaltStack
tcpdump/netstat
tcptraceroute/mtr
MySQL
Aerospike

Job description

Site Reliability Engineer - AWS (4 to 8 Years)
  • Full-time

PhonePe Limited (Formerly PhonePe Private Limited) is a technology company that builds digital platforms for Payments, Digital Distribution Services and Financial Services. Headquartered in India, the PhonePe digital payments app was launched in 2016. As of April 2026, PhonePe has over 70 Crore life-till-date registered users and a digital payments acceptance network spread across over 5 Crore merchants. PhonePe’s products and services include Consumer Payments (including Digital Distribution Services), Merchant Payments, Lending and Insurance Distribution services, and New Platforms, which comprise Share.Market (stock broking and mutual funds distribution platform), and Indus Appstore (Android-based mobile app marketplace). Culture:

At PhonePe, we go the extra mile to make sure you can bring your best self to work, Everyday!. And that starts with creating the right environment for you. We empower people and trust them to do the right thing. Here, you own your work from start to finish, right from day one. PhonePe-rs solve complex problems and execute quickly; often building frameworks from scratch. If you’re excited by the idea of building platforms that touch millions, ideating with some of the best minds in the country and executing on your dreams with purpose and speed, join us!

We are seeking a highly motivated Site Reliability Engineer (SRE) with 4 to 8 years of experience to manage, scale, and ensure the high availability of our core infrastructure. This role is designed for experts specialized in AWS. You will leverage a profound background in Linux (specifically RHEL) to drive deep-level cloud architecture, automation, complex networking, and security compliance, supporting a high-volume, mission-critical environment that demands exceptional uptime and resilience.

Roles and Responsibilities

Perform daily administration, configuration, and troubleshooting of Linux systems (specifically RHEL). Manage system resources, file systems, and package installations to ensure optimal OS-level health and performance.

Provision, configure, and maintain AWS EC2 instances and cloud-native components (IAM, Load Balancers, S3, CloudWatch, etc), specializing in RHEL environments.

Proactively monitor for, identify, and remediate infrastructure vulnerabilities to maintain a secure environment. Implement security guardrails defined by organizational policies.

Configure and maintain AWS VPCs, Route Tables, and Security Groups. Perform network-level troubleshooting for routing issues using tcptraceroute, mtr, tcpdump, netstat etc.

Collaborate with stakeholders to execute infrastructure migration plans and component upgrades with strict adherence to established runbooks.

Write and maintain Ansible playbooks and SaltStack states for configuration management. Automate routine operational tasks using Python, Go, or Bash.

Manage MySQL and Aerospike data stores, including routine backups, maintenance, upgrades and scaling.

Build, maintain, and extend platform observability by configuring monitoring dashboards and actionable alerts.

Create, update, and maintain comprehensive runbooks, standard operating procedures (SOPs), and architecture documentation to ensure operational consistency and seamless knowledge sharing.

Participate in the on-call rotation to handle active incidents, mitigate downtime, and draft initial Root Cause Analysis (RCA) reports.

Experience: 4 to 8years in an SRE, DevOps, or Systems Engineering role.

Strong proficiency in hands-on Linux/RHEL administration and troubleshooting.

Deep, hands-on experience exclusively within the AWS ecosystem

Mandatory hands-on experience with containerization using Docker, Podman, or equivalent technologies.

Solid grasp of the TCP/IP stack, DNS, routing concepts

Understanding of industry best practices for maintaining a highly secure infrastructure

Working knowledge of SSL/TLS certificate and their lifecycle management.

Ability to communicate effectively in English, both written and verbally

Preferred Qualifications (A Plus)

Working experience with Nginx and HAProxy proxies

Understanding of MySQL/MariaDB/Percona or any other relational database concepts and administration

Knowledge of NoSQL databases like Aerospike

Familiarity with Kafka or similar distributed event streaming platforms

PhonePe Full Time Employee Benefits (Not applicable for Intern or Contract Roles)
  • Insurance Benefits - Medical Insurance, Critical Illness Insurance, Accidental Insurance, Life Insurance
  • Wellness Program - Employee Assistance Program, Onsite Medical Center, Emergency Support System
  • Parental Support - Maternity Benefit, Paternity Benefit Program, Adoption Assistance Program, Day-care Support Program
  • Mobility Benefits - Relocation benefits, Transfer Support Policy, Travel Policy
  • Retirement Benefits - Employee PF Contribution, Flexible PF Contribution, Gratuity, NPS, Leave Encashment
  • Other Benefits - Higher Education Assistance, Car Lease, Salary Advance Policy

Our inclusive culture promotes individual expression, creativity, innovation, and achievement and in turn helps us better understand and serve our customers. We see ourselves as a place for intellectual curiosity, ideas and debates, where diverse perspectives lead to deeper understanding and better quality results. PhonePe is an equal opportunity employer and is committed to treating all its employees and job applicants equally; regardless of gender, sexual preference, religion, race, color or disability. If you have a disability or special need that requires assistance or reasonable accommodation, during the application and hiring process, including support for the interview or onboarding process, please fill out this form.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer - AWS (7 to 12 Years)
Site Reliability Engineer - AWS (7 to 12 Years)

PHONEPE LIMITED • Bengaluru

On-site
INR 2,500,000 - 4,000,000
Insurance Benefits - Medical Insurance
Critical Illness Insurance
Accidental Insurance
+5
Site Reliability Engineer 2 Years
Site Reliability Engineer 2 Years

PhonePe • Bengaluru

On-site
INR 900,000 - 1,500,000
Insurance Benefits
Wellness Program
Parental Support
+3
Site Reliability Engineer (4 to 8 Years)
Site Reliability Engineer (4 to 8 Years)

PhonePe • Bengaluru

On-site
INR 2,500,000 - 4,200,000
Medical Insurance
Critical Illness Insurance
Accidental Insurance
+6
Site Reliability Engineer - On-Prem (4 to 8 Years)
Site Reliability Engineer - On-Prem (4 to 8 Years)

PHONEPE LIMITED • Bengaluru

On-site
INR 2,600,000 - 4,200,000
Medical Insurance
Critical Illness Insurance
Onsite Medical Center
+1
Site Reliability Engineer (2+ Years)
Site Reliability Engineer (2+ Years)

PhonePe • Bengaluru

On-site
INR 1,200,000 - 2,400,000
Medical Insurance
Critical Illness Insurance
Accidental Insurance
+16
Site Reliability Engineer - On-Prem (4 to 12 Years)
Site Reliability Engineer - On-Prem (4 to 12 Years)

PHONEPE LIMITED • Bengaluru

On-site
INR 2,800,000 - 5,500,000
Medical Insurance
Onsite Medical Center
PF Contribution
+1
Site Reliability Engineer - On-Prem (1 to 4 Years)
Site Reliability Engineer - On-Prem (1 to 4 Years)

PHONEPE LIMITED • Bengaluru

On-site
INR 1,000,000 - 1,800,000
Medical Insurance
Critical Illness Insurance
Accidental Insurance
+5
Site Reliability Engineer (4+ YOE)
Site Reliability Engineer (4+ YOE)

PhonePe • Bengaluru

On-site
INR 3,500,000 - 6,000,000
Medical Insurance
Critical Illness Insurance
Accidental Insurance
+19
Site Reliability Engineer 2
Site Reliability Engineer 2

PhonePe • Bengaluru

On-site
INR 1,000,000 - 2,000,000
Medical Insurance
Wellness Program
Maternity Benefit
+2
Software Engineer - SRE (Rust)
Software Engineer - SRE (Rust)

PHONEPE LIMITED • Bengaluru

On-site
INR 900,000 - 1,400,000
Insurance Benefits
Wellness Program
Parental Support
+3