Technical Lead, Cloud & Infra Engg — SRE & Observability

Birlasoft

New Jersey

On-site

USD 140,000 - 190,000

Full time

7 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Birlasoft is seeking a Technical Lead-Cloud & Infra Engg to guide SRE/DevOps efforts, drive automation, and manage middleware, CDN, and incident response. You will lead a team across incident management, change control, and production reliability while collaborating with product, QA, and development teams.

The role emphasizes leadership, governance, and hands-on optimization of Java EE environments, edge platform management, and comprehensive monitoring to ensure production resilience.

Qualifications

  • Proven experience in Site Reliability Engineering (SRE) and cloud/infrastructure environments.
  • Strong leadership and mentoring abilities for technical teams.
  • Experience with incident management, RCA, and ITIL-aligned processes.

Responsibilities

  • Ensure high availability, performance, and resilience of production systems.
  • Implement SRE best practices: error budgets, SLIs/SLOs, capacity planning, chaos testing, runbook creation.
  • Drive automation to reduce manual operational tasks and improve MTTR.
  • Conduct post‑incident reviews (PIRs) and implement long‑term corrective actions.
  • Manage deployments, rollbacks, and environment synchronization across Dev, QA, UAT, and Production.
  • Install, configure, upgrade, and maintain WebLogic, Tomcat, Apache, and Nginx servers; perform JVM tuning, thread pool optimization, connection pool management, and performance tuning.
  • Troubleshoot middleware issues related to memory leaks, thread contention, SSL, certificates, and clustering.
  • Configure dashboards, alerts, and performance insights using New Relic and Splunk.
  • Develop log‑based monitoring strategies and anomaly detection.
  • Implement proactive monitoring to reduce downtime and improve reliability.
  • Configure Akamai caching rules, WAF policies, edge redirects, and performance optimizations.
  • Troubleshoot CDN‑related latency, caching, and routing issues.
  • Collaborate with Akamai support for advanced troubleshooting.
  • Lead major incident bridges, coordinate cross functional teams, and provide timely updates.
  • Manage problem tickets, root cause analysis, and preventive action plans.
  • Ensure compliance with ITIL processes for change, release, and incident management.
  • Lead and mentor a team of SRE/DevOps engineers.
  • Provide technical guidance, training, and performance feedback.
  • Act as a customer facing technical SME for escalations and production issues.
  • Collaborate with product, QA, development, and business teams to ensure smooth delivery.
  • Ownership & Accountability: Takes responsibility for production stability and issue resolution.
  • Leadership: Guides team members, manages workload, and drives operational excellence.
  • Communication: Clear, structured communication with customers and internal teams.
  • Problem Solving: Strong analytical skills and ability to troubleshoot complex issues.
  • Collaboration: Works effectively across engineering, QA, product, and business teams.
  • Calm Under Pressure: Handles critical incidents with composure and clarity.

Job description

Birlasoft is seeking a Technical Lead-Cloud & Infra Engg to guide SRE/DevOps efforts, drive automation, and manage middleware, CDN, and incident response. You will lead a team across incident management, change control, and production reliability while collaborating with product, QA, and development teams.

The role emphasizes leadership, governance, and hands-on optimization of Java EE environments, edge platform management, and comprehensive monitoring to ensure production resilience.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Technical Lead-Cloud & Infra Engg
Technical Lead-Cloud & Infra Engg

Birlasoft • New Jersey

On-site
USD 140,000 - 190,000
Senior Java SRE Tech Lead: Cloud, APIs & Observability
Senior Java SRE Tech Lead: Cloud, APIs & Observability

Pacer Group • Phoenix (AZ)

On-site
USD 150,000 - 190,000
Medical insurance
Dental insurance
Vision insurance
+1
Technical Lead - Java Spring Boot & Cloud Apps
Technical Lead - Java Spring Boot & Cloud Apps

Birlasoft • Kettering (OH)

On-site
USD 110,000 - 140,000
Senior Java SRE Engineer - Cloud, Observability, Automation
Senior Java SRE Engineer - Cloud, Observability, Automation

Cognizant • Phoenix (AZ)

Hybrid
USD 112,000 - 132,000
Stock awards
Discretionary annual incentive program
Medical/Dental/Vision/Life Insurance
+1
SRE Lead: Production Reliability & Observability Architect
SRE Lead: Production Reliability & Observability Architect

TechDigital Group • Woonsocket (RI)

On-site
USD 140,000 - 190,000
Senior SRE Lead - Cloud Reliability & Observability
Senior SRE Lead - Cloud Reliability & Observability

EPAM Systems, Inc. • United States

Remote
USD 120,000 - 180,000
Healthcare benefits
Paid time off
Learning & development
+1
SRE Operations Lead — 24x7 Reliability & AI‑Driven Automation
SRE Operations Lead — 24x7 Reliability & AI‑Driven Automation

Tata Consultancy Services • Charlotte (NC)

On-site
USD 70,000 - 120,000
Remote SRE Technical Lead - Automation & Reliability
Remote SRE Technical Lead - Automation & Reliability

Bright Vision Technologies • Austin (TX), New York (NY)

On-site
USD 100,000 - 150,000
SRE Manager - AI-Driven Reliability & Observability
SRE Manager - AI-Driven Reliability & Observability

M&T Bank • South Dakota

On-site
USD 140,000 - 233,000
SRE Lead: Reliability & Cloud Observability Architect
SRE Lead: Reliability & Cloud Observability Architect

BlackCube Labs • San Diego (CA)

On-site
USD 190,000 - 280,000