Module Lead - Systems

Mphasis

Hyderabad

On-site

INR 2,500,000 - 4,200,000

Full time

3 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Mphasis is seeking a seasoned Site Reliability Engineer to support large-scale distributed systems. The role emphasizes quick learning, strong communication, and hands-on skills in monitoring, automation, and incident management.

You will manage production incidents and drive remediation across Java, .NET, and Batch applications deployed on GCP, PCF, and on-premise environments. The ideal candidate has 6+ years of SRE experience, deep knowledge of CI/CD, and a solid background in Linux/Unix,

Qualifications

  • 6+ years of experience as Site Reliability Engineer in large-scale distributed systems.
  • Proven ability to manage production failures, perform root cause analysis and remediation.
  • 24/7 support experience as part of an SRE team; strong in monitoring, release management, and automation.
  • Hands-on with CI/CD deployments and cloud platforms (GCP, PCF, AWS, Azure) and on‑prem environments.

Responsibilities

  • Provide 24/7 SRE support for mission-critical Java, .NET, and Batch applications across GCP, PCF and on‑prem environments.
  • Investigate incidents, perform RCA, and drive remediation to prevent recurrence.
  • Collaborate with engineering teams during outages and major incidents; manage release and change processes.
  • Develop and maintain monitoring, automation, and alerting to improve reliability.

Skills

SRE expertise
Incident management
Root cause analysis
CI/CD automation
Linux/Unix
Cloud platforms
Networking
Python/Shell scripting
Communication
Automation

Tools

Splunk
AppDynamics
ThousandEyes
ITRS
AppMetrics
Moogsoft
Kafka

Job description

Job Summary –

Seasoned Site Reliability Engineer (SRE) with 6+ years of experience in supporting complex, large-scale distributed systems. Highly skilled in managing production failures, conducting root cause analysis, and driving effective remediation. Strong communicator with expertise in ing, monitoring, and release management, complemented by automation proficiency and a keen ability to learn quickly.

Job Summary –

Seasoned Site Reliability Engineer (SRE) with 6+ years of experience in supporting complex, large-scale distributed systems. Highly skilled in managing production failures, conducting root cause analysis, and driving effective remediation. Strong communicator with expertise in ing, monitoring, and release management, complemented by automation proficiency and a keen ability to learn quickly. This role involves providing 24/7 support as part of the SRE team, ensuring the reliability and performance of mission-critical Java, .NET, and Batch applications deployed across GCP, PCF, and on-premise environments.

Years of experience needed –

Candidate experience – 6+ Years

Technical Skills
  • Expertise in understanding large scale production systems and technologies, for example load balancing, monitoring, distributed systems, microservices, and configuration management.
  • Should have solid hands‑on experience in troubleshooting and fixing application failures, application Performance degradation, Code issues, cloud platform issues, Batch Failures, Infra failures, DB failures, Network failures.
  • Hands‑on experience in performing Production deployments using CI/CD and exposure to deployment strategies.
  • Experience in troubleshooting of Linux/Unix.
  • Monitor the application/Services/batch availability.
  • Act quickly on the application s(Performance, Availability) and Batch Job failures
  • Perform the required analysis (Code/Log) and escalation to the Engineering team as required.
  • Initiate and drive the Techlines in case of outages/major incidents/Batch abends and ensure Service Restoration in the least time possible.
  • Effectively handle the Incident, Problem, Release and Change management.
  • Own and deliver the user stories assigned as part of the sprint.
  • The user stories range from application code Debugging, Issue analysis, Code fix, Knowledge base creation, documentation of SOP’s, Production Deployments, Pre & Post Patching/Maintenance activities, Service Requests.
  • Build monitoring solutions using APM tools like Splunk, Appdynamics, Thousand Eyes, ITRS, AppMetrics, MoogSoft, Kafka etc.
  • Automate of day‑day operational tasks.
  • Be part of the Exit reviews to ensure the best practices are followed to have the right code deployed to Production systems
  • Provide feedback/recommend improvements to the system which would enable highly stable systems.
  • Strong understanding of Networking Concepts (TCP/IP, SSL/TLS, IPSec, VPN etc), Firewall and Load Balancers.
  • Experience in Scripting – Shell/Powershell/Python
  • Strong Experience in working with any Cloud‑based infrastructure (PCF, GCP, AWS, Azure Cloud or others)
Certifications Needed

As per industry standards

About Mphasis

Mphasis applies next‑generation technology to help enterprises transform businesses globally. Customer centricity is foundational to Mphasis and is reflected in the Mphasis’ Front2Back™ Transformation approach. Front2Back™ uses the exponential power of cloud and cognitive to provide hyper‑personalized (C=X2C2TM=1) digital experience to clients and their end customers. Mphasis’ Service Transformation approach helps ‘shrink the core’ through the application of digital technologies across legacy environments within an enterprise, enabling businesses to stay ahead in a changing world. Mphasis’ core reference architectures and tools, speed and innovation with domain expertise and specialization are key to building strong relationships with marquee clients.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

SRE Engineer
SRE Engineer

Mphasis • Hyderabad

On-site
INR 1,200,000 - 2,400,000
Senior Tech Lead/Architect
Senior Tech Lead/Architect

Cloudxtreme • Bengaluru

On-site
INR 2,500,000 - 4,000,000
Senior Team Lead | Site Reliability Engineering | Bengaluru | Engineering
Senior Team Lead | Site Reliability Engineering | Bengaluru | Engineering

Deloitte & Touche GmbH Wirtschaftsprüfungsgesellschaft • Bengaluru

On-site
INR 4,000,000 - 6,000,000
Technical Lead
Technical Lead

Mphasis • Hyderabad

On-site
INR 1,800,000 - 3,200,000
Production Support Lead
Production Support Lead

Cloudxtreme • Hyderabad

On-site
INR 3,200,000 - 6,000,000
SRE
SRE

Genpact • Karnataka

On-site
INR 1,200,000 - 1,600,000
Associate, Specialist, SRE Engineer, Site Reliability Engineering
Associate, Specialist, SRE Engineer, Site Reliability Engineering

DBS Bank • Hyderabad

On-site
INR 1,200,000 - 1,800,000
Site Reliability Engineer Lead Position
Site Reliability Engineer Lead Position

Cloudxtreme • Hyderabad

On-site
INR 2,600,000 - 5,000,000
Software Engineer-DevOps
Software Engineer-DevOps

SMC Squared India • Bengaluru

On-site
INR 1,000,000 - 1,500,000
Tech & Digital-Lead Site Reliability Engineer
Tech & Digital-Lead Site Reliability Engineer

Hdfc Bank • Bengaluru

On-site
INR 3,500,000 - 5,500,000