Tech S and T -Resilience And Reliability Architect-Manager-GDSF02

EY

Pune District

On-site

INR 3,500,000 - 7,000,000

Full time

5 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

EY in India seeks an experienced Site Reliability Engineering (SRE) Architect/Consultant to design and implement resilient, scalable IT operations across enterprise environments. You will define NFRs, SLIs/SLA, and roadmaps to raise maturity and efficiency, partnering with product teams to reduce toil and optimize costs.

The role requires hands-on expertise in Java, CI/CD, cloud platforms, observability tools, and Linux systems, with a track record of delivering large-scale distributed systems.

Qualifications

  • 12+ years of experience in software product engineering principles, processes and systems.
  • Hands-on experience in Java / J2EE, web servers, application servers, and major RDBMS.
  • Experience with at least one CI/CD and IaC tools.
  • Experience with at least one cloud technology and its reliability tools.
  • Experience in Observability and performance tuning in distributed systems.
  • Linux experience and knowledge of performance monitoring.
  • Knowledge of microservices, Spring Boot, and cloud-native architectures.
  • Proficiency in Java runtimes, JVM tuning, and troubleshooting.
  • Automation scripting with Python is desirable.
  • Familiarity with Git/Jira/Confluence and collaborative development.

Responsibilities

  • Defining SLA/SLO/SLI for products and services.
  • Engineering resilient design and automation to reduce toil.
  • Designing Observability Solutions to track SLA adherence.
  • Optimizing IT infra and operations costs (FinOps).

Skills

Java
J2EE
CI/CD
IaC
Cloud platforms
Observability
Linux
Microservices
Spring Boot
Docker
Kubernetes
Automation Python
Git
Jira
Confluence
APM tools

Tools

Terraform
Ansible
Jenkins
CloudFormation
AWS
Azure
GCP
Dynatrace
AppDynamics
Splunk
ELK
OpenShift

Job description

At EY, you’ll have the chance to build a career as unique as you are, with the global scale, support, inclusive culture and technology to become the best version of you. And we’re counting on your unique voice and perspective to help EY become even better, too. Join us and build an exceptional experience for yourself, and a better working world for all.

Site Reliability Engineering (SRE) Architect / Consultant (M)
Description
  • Site Reliability Engineering (SRE) is a modern way of delivering IT Operations by imbibing Software engineering principles in Service Delivery to reduce IT Risk to business, improve business resilience, attain predictability & reliability, optimize cost of IT Infra and Ops
  • An SRE Architect / Consultant will help in designing the roadmap to SRE for Enterprise IT
  • They will also implement various SRE Solutions across the enterprise / line-of-businesses
  • They will be able to assess SRE Maturity of an IT Organization and provide strategy and roadmap to achieve higher maturity levels
Responsibilities
  • Defining SLA/SLO/SLI for a product / service
  • Engineering in resilient design and implementation practices into solutions as they go through the product life cycle
  • Designing & implementing Observability Solutions to track, report, and measure SLA adherence
  • Engineering out manual effort (Toil) through the development of automated processes and services (e.g., Automated Management of Systems, CI/CD improvements)
  • Optimize Cost of IT Infra & Operations - FinOps
Typical Skills And Background
  • 12+ years of experience in software product engineering principles, processes and systems
  • Hands-on experience in Java / J2EE, one of web server (Apache Tomcat or IBM HTTP Server), one of the application servers (Tomcat/WebSphere), and any major RDBMS like Oracle
  • Hands-on experience in at least one CI-CD (Azure DevOps, GitLab CI/CD, Jenkins) and IaC tools (Terraform, AWS CloudFormation, Ansible etc.)
  • Experience in at least one cloud technology (AWS/Azure/GCP etc. and Docker, Pivotal, Kubernetes, OpenShift etc.) and its reliability tools (Azure AppInsight, CloudWatch, Azure Monitor etc.)
  • Experience in Observability - APM tools (Dynatrace, AppDynamics etc.), metrics / log consolidation (Splunk) and ELK Stack
  • Experience in Linux (RHEL) operating system performance monitoring parameters and their interpretation, commands used for monitoring
  • Experience in Web Services, SOA, ESB (DataPower), RESTFul
  • Defining NFRs and SLA/SLO/SLI agreement for a product / platform / services
  • Knowledge on queuing models used, thread pools, request servicing processes etc.
  • Knowledge of application design patterns, J2EE application architectures, Microservices, Spring boot & Cloud native architectures
  • Proficiency in Java runtimes, Core Java, Garbage collection, JVM parameters tuning
  • Experience in performance tuning on Application Servers (Tomcat/WAS)
  • Experience in trouble shooting Performance / Scalability / Availability issues
  • Thread dump, heap dump generation & analysis
  • Knowledge on Query tuning and database architecture
  • Knowledge at least one automation scripting language like Python
  • Mastery of collaborative software development using Git, Jira, Confluence etc.
  • AI/ML & Data Analytics knowledge and experience is a desirable
EY | Building a better working world

EY exists to build a better working world, helping to create long-term value for clients, people and society and build trust in the capital markets.

Enabled by data and technology, diverse EY teams in over 150 countries provide trust through assurance and help clients grow, transform and operate.

Working across assurance, consulting, law, strategy, tax and transactions, EY teams ask better questions to find new answers for the complex issues facing our world today.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Tech S and T -Resilience And Reliability Architect-Manager-GDSF02
Tech S and T -Resilience And Reliability Architect-Manager-GDSF02

EY • Mumbai

On-site
INR 3,000,000 - 5,500,000
Tech S and T -Resilience And Reliability Architect-Manager-GDSF02
Tech S and T -Resilience And Reliability Architect-Manager-GDSF02

EY • Bengaluru

On-site
INR 3,500,000 - 7,000,000
Tech S and T-Resilience and Reliability Engineer-Senior-GDSF02
Tech S and T-Resilience and Reliability Engineer-Senior-GDSF02

EY • Kolkata District

On-site
INR 1,200,000 - 2,000,000
Tech S and T -Resilience And Reliability Architect-Manager-GDSF02
Tech S and T -Resilience And Reliability Architect-Manager-GDSF02

Ernst & Young Advisory Services Sdn Bhd • India

On-site
INR 5,000,000 - 8,000,000
Tech S And T-Resilience And Reliability Engineer-Senior-GDSF02
Tech S And T-Resilience And Reliability Engineer-Senior-GDSF02

EY • Mumbai

On-site
INR 1,500,000 - 2,100,000
Tech S And T-Resilience And Reliability Engineer-Senior-GDSF02
Tech S And T-Resilience And Reliability Engineer-Senior-GDSF02

EY • Chennai District

On-site
INR 1,200,000 - 2,000,000
Tech S And T-Resilience And Reliability Engineer-Senior-GDSF02
Tech S And T-Resilience And Reliability Engineer-Senior-GDSF02

EY • Hyderabad

On-site
INR 2,800,000 - 6,000,000
Tech S and T-Resilience and Reliability Engineer-Senior-GDSF02
Tech S and T-Resilience and Reliability Engineer-Senior-GDSF02

Ernst & Young Advisory Services Sdn Bhd • Dadri

On-site
INR 3,000,000 - 6,000,000
Resilience and Reliability Engineer
Resilience and Reliability Engineer

EY • Pune District, Gurugram District, Bengaluru

Hybrid
INR 1,800,000 - 2,800,000
Site Reliability Engineer - Vice President
Site Reliability Engineer - Vice President

Citi • Maharashtra

On-site
INR 4,000,000 - 7,000,000