SRE Engineer

Mphasis

Hyderabad

On-site

INR 1,200,000 - 2,400,000

Full time

3 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Mphasis is hiring a Site Reliability Engineer to join our Hyderabad team. The candidate will manage production incidents, perform root cause analysis, and drive remediation across Java, .NET, and Batch apps deployed on GCP, PCF, and on-prem environments.

24/7 support experience is preferred. Ideal candidates excel in monitoring, automation, and release management, with strong communication and quick learning abilities.

Qualifications

  • Experience supporting large-scale distributed systems and production incidents.
  • Proficient in Linux/Unix troubleshooting and performance optimization.
  • Strong networking knowledge (TCP/IP, TLS, VPNs, load balancers).

Responsibilities

  • Monitor services and batch jobs; respond to performance and availability issues.
  • Perform root cause analysis on failures using logs and code, escalate when needed.
  • Lead incident management, postmortems, and change control for deployments.
  • Own user stories in sprints including debugging and documenting SOPs.
  • Conduct deployments via CI/CD and refine deployment strategies.
  • Build and tune monitoring with APM tools and dashboards.
  • Automate routine ops tasks to improve reliability and velocity.
  • Participate in post-incident reviews to drive improvements.

Skills

Large-scale systems
Linux/Unix
Networking
Shell/Python
Cloud platforms
Incident management
CI/CD

Education

Bachelor's degree in CS/IT

Tools

Splunk
AppDynamics
Kafka
APM tools

Job description

Job Title: Site Reliability Engineer (SRE) Designation: SRE Engineer

Location: Hyderabad_Phoenix

Job Summary

We are seeking a seasoned Site Reliability Engineer (SRE) to join our dynamic team. The ideal candidate will have extensive experience in supporting complex, large-scale distributed systems. You will be responsible for managing production failures, conducting root cause analysis, and driving effective remediation. As a strong communicator, you will leverage your expertise in ing, monitoring, and release management, complemented by automation proficiency and a keen ability to learn quickly. This role involves providing 24/7 support as part of the SRE team, ensuring the reliability and performance of mission-critical Java, .NET, and Batch applications deployed across GCP, PCF, and on-premise environments.

Responsibilities
  • Monitor application, services, and batch availability, acting swiftly on s related to performance and availability.
  • Perform thorough analysis (code/log) and escalates issues to the engineering team as necessary.
  • Initiate and drive tech lines during outages, major incidents, or batch abends, ensuring service restoration in the least time possible.
  • Effectively manage incidents, problems, releases, and changes.
  • Own and deliver user stories assigned as part of the sprint, including application code debugging, issue analysis, code fixes, knowledge base creation, and documentation of SOPs.
  • Conduct production deployments using CI/CD and exposure to deployment strategies.
  • Build monitoring solutions using APM tools such as Splunk, AppDynamics, Thousand Eyes, ITRS, AppMetrics, MoogSoft, and Kafka.
  • Automate day-to-day operational tasks to enhance efficiency.
  • Participate in exit reviews to ensure best practices are followed for code deployment to production systems.
  • Provide feedback and recommend improvements to enhance system stability.
Mandatory Skills
  • Expertise in large-scale production systems and technologies, including load balancing, monitoring, distributed systems, microservices, and configuration management.
  • Solid hands-on experience in troubleshooting and resolving application failures, performance degradation, code issues, cloud platform issues, batch failures, infrastructure failures, database failures, and network failures.
  • Proficiency in troubleshooting Linux/Unix environments.
  • Strong understanding of networking concepts (TCP/IP, SSL/TLS, IPSec, VPN, etc.), firewalls, and load balancers.
  • Experience in scripting languages such as Shell, PowerShell, or Python.
  • Strong experience working with cloud-based infrastructure (PCF, GCP, AWS, Azure, or others).
  • Proven ability to handle application production support effectively.
Preferred Skills
  • Familiarity with CI/CD tools and deployment strategies.
  • Experience with APM tools and monitoring solutions.
  • Knowledge of automation frameworks and tools.
Qualifications

Certifications as per industry standards are preferred. A strong educational background in Computer Science, Information Technology, or a related field is advantageous.

About Mphasis

Mphasis applies next-generation technology to help enterprises transform businesses globally. Customer centricity is foundational to Mphasis and is reflected in the Mphasis Front2Back Transformation approach. Front2Back uses the exponential power of cloud and cognitive to provide hyper-personalized (C=X2C2TM=1) digital experience to clients and their end customers. Mphasis Service Transformation approach helps shrink the core through the application of digital technologies across legacy environments within an enterprise, enabling businesses to stay ahead in a changing world. Mphasis core reference architectures and tools, speed and innovation with domain expertise and specialization are key to building strong relationships with marquee clients.

Equal Opportunity Employer

Mphasis is an equal opportunity/affirmative action employer. We provide equal employment opportunities to applicants and existing associates and evaluate qualified candidates without regard to race, gender, national origin, ancestry, age, color, religious creed, marital status, genetic information, sexual orientation, gender identity, gender expression, sex (including pregnancy, breast feeding and related medical conditions), mental or physical disability, medical conditions military and veteran status or any other status or condition protected by applicable federal, state, or local laws, governmental regulations and executive orders. View the EEO in the law poster here, view the EEO in the law supplement here. To view the pay transparency nondiscrimination provision please click here and to view the E-Verify posting click here.

Mphasis is committed to providing reasonable accommodations to individuals with disabilities. If you need a reasonable accommodation because of disability to search and apply for a career opportunity, please send an email to accomodationrequest@mphasis.com and let us know your contact information and the nature of your request.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Module Lead - Systems
Module Lead - Systems

Mphasis • Hyderabad

On-site
INR 2,500,000 - 4,200,000
SRE Engineer II/III
SRE Engineer II/III

Rackspace Technology • Gurugram District

On-site
INR 1,500,000 - 2,500,000
SRE Engineer II/III
SRE Engineer II/III

Rackspace Technology • Hyderabad

On-site
INR 1,000,000 - 1,800,000
Site Reliability Engineer
Site Reliability Engineer

Acesoft Labs • Ahmedabad District

Hybrid
INR 400,000 - 700,000
Site Reliability Engineer
Site Reliability Engineer

Acesoft Labs • Hyderabad

Hybrid
INR 1,800,000 - 3,000,000
Site Reliability Engineer
Site Reliability Engineer

Arch Systems • Hyderabad

On-site
INR 2,800,000 - 4,200,000
Java Developer with SRE experience
Java Developer with SRE experience

Indihire Consultants • Hyderabad

Hybrid
INR 1,800,000 - 2,800,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

VMC Soft Technologies, Inc • Hyderabad

Hybrid
INR 1,500,000 - 2,000,000
SRE
SRE

Metlife • Hyderabad

Hybrid
INR 1,500,000 - 2,100,000
SRE Engineer II
SRE Engineer II

Webhosting • Hyderabad

On-site
INR 1,400,000 - 2,000,000