Site Reliability Engineer

ISA

Maharashtra

On-site

INR 1,200,000 - 1,800,000

Full time

2 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

ISA is seeking a Site Reliability engineer to monitor, troubleshoot, and improve reliability of enterprise and airline systems. You will resolve incidents, optimize performance, and collaborate with Software Engineering and SRE teams to uphold regulatory and policy standards.

The role requires hands-on experience with Java, Spring Boot, Docker, and Kubernetes, plus strong SQL skills and Linux familiarity. Excellent English communication is essential, with 2–4 years in Java development or support.

Qualifications

  • Bachelor’s degree in computer engineering, computer science, or information technology.
  • Fluent in English.
  • Strong knowledge of Java, Spring Boot, and enterprise application development fundamentals.
  • Understanding of microservices and monolithic architectures and their deployment models.
  • Basic debugging and troubleshooting skills for production issues.
  • Familiarity with CI/CD tools (e.g., Jenkins), Git, and automated deployments.
  • Strong SQL skills with database querying; Oracle DB is a plus.
  • Familiarity with Linux OS, CLI, and networking basics.
  • MS Office proficiency.

Responsibilities

  • Investigate day-to-day production incidents by analyzing logs, metrics, and monitoring data to identify root causes.
  • Troubleshoot, debug, and support Java-based applications to ensure system stability and availability.
  • Develop, implement, and validate bug fixes and enhancements following reliability standards.
  • Operate and maintain monitoring, alerting, and observability tools; flag gaps in coverage.
  • Execute CI/CD pipelines, deployments, and release tasks per runbooks.

Skills

Java
Spring Boot
Microservices
Linux
SQL
CI/CD
Docker
Kubernetes
Debugging
English communication

Education

Bachelor's degree in Computer Engineering/CS/IT

Tools

Jenkins
Git
Docker
Kubernetes
Oracle Database

Job description

Job Purpose

To support the reliability, availability, and performance of enterprise applications and mission-critical airline systems by monitoring production environments, resolving incidents, troubleshooting application issues, and implementing reliability improvements. Contributes to maintaining stable and resilient software services through collaboration with Software Engineering and Site Reliability Engineering teams while ensuring compliance with organizational policies, industry standards, and applicable regulatory requirements.

Job Purpose

To support the reliability, availability, and performance of enterprise applications and mission-critical airline systems by monitoring production environments, resolving incidents, troubleshooting application issues, and implementing reliability improvements. Contributes to maintaining stable and resilient software services through collaboration with Software Engineering and Site Reliability Engineering teams while ensuring compliance with organizational policies, industry standards, and applicable regulatory requirements.

Key Result Responsibilities
  • Investigate day-to-day production incidents by analyzing application logs, system metrics, and monitoring data to identify root causes and escalated complex issues appropriately.
  • Troubleshoot, debug, and support Java-based applications and services under the guidance of senior engineers to ensure day-to-day system stability and availability.
  • Develop, implement, and validate bug fixes and application enhancements under the direction of senior engineers, following established coding and reliability standards.
  • Operate and maintain existing application monitoring, alerting, and observability tools as configured by senior team members, flagging gaps in coverage.
  • Execute day-to-day CI/CD pipeline runs, deployment activities, and release tasks according to established processes and runbooks.
Key Result Responsibilities-Continued
  • Operate and support existing containerized applications on Docker and Kubernetes within pre-defined production configurations, escalating architectural changes to senior engineers.
  • Participate in the on-call rotation, responding to production alerts within defined SLAs and escalating unresolved issues to senior engineers or the Lead SRE.
  • Document incident resolutions, troubleshooting steps, and known issues in runbooks and the team knowledge base to support faster future resolution.
  • Perform routine system health checks, log reviews, and capacity/performance monitoring to proactively flag potential issues before they impact production.
  • Execute scheduled maintenance activities, including patching, backups, and routine configuration changes, following approved change procedures.
Qualifications (Academic, Training, Languages)
  • Bachelor’s degree in computer engineering/computer science/information technology.
  • Fluent in English Language.
  • Good knowledge of Java, Spring Boot, and enterprise application development principles.
  • Understanding of microservices and monolithic application architectures and their deployment models.
  • Basic debugging and troubleshooting skills with the ability to diagnose and resolve application and production issues.
  • Familiarity with CI/CD tools (e.g., Jenkins), Git-based version control, and automated deployment practices.
  • Strong knowledge of SQL with experience in database querying and troubleshooting; experience with Oracle Database is an advantage.
  • Familiarity with Linux operating systems, command-line tools, and networking fundamentals.
  • Proficient in MS Office.
  • Exposure to Node.js and Next.js is an advantage
Work Experience
  • With 2–4 years of experience in Java development or support.
  • Experience with JBoss Application Server or similar enterprise Java application servers is an added advantage.
  • Hands-on experience with Docker and Kubernetes for deploying and supporting containerized applications.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

ISA • Maharashtra

On-site
INR 1,500,000 - 2,300,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

ISA • Maharashtra

On-site
INR 1,400,000 - 2,200,000
Senior Associate Site Reliability Engineer
Senior Associate Site Reliability Engineer

NTT DATA BUSINESS SOLUTIONS • Hyderabad

On-site
INR 1,400,000 - 2,000,000
Site Reliability Engineer
Site Reliability Engineer

Hackajob • Ahmedabad District, Gurugram District, Mumbai

Hybrid
INR 1,800,000 - 2,600,000
Site Reliability Engineer
Site Reliability Engineer

Snapmint • Gurugram District

On-site
INR 800,000 - 1,200,000
Site Reliability Engineer
Site Reliability Engineer

Peoplefy • Pune District

On-site
INR 1,200,000 - 1,800,000
Lead Site Reliability Engineer/ Expert
Lead Site Reliability Engineer/ Expert

SITA • Bengaluru

Hybrid
INR 3,500,000 - 6,000,000
Flex Week: Hybrid
Flex Location: Up to 30 days remote
Employee Wellbeing programs (EAP)
+2
Application Support Engineer
Application Support Engineer

Accenture in India • Indore District

On-site
INR 1,200,000 - 2,000,000
AWS Site Reliability Engineer
AWS Site Reliability Engineer

Infosys • Bengaluru

Hybrid
INR 900,000 - 1,500,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Infosys • Hyderabad

On-site
INR 1,400,000 - 2,200,000