Site Reliability Engineer

Tech Mahindra

Dublin

Hybrid

EUR 90,000 - 135,000

Full time

13 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Tech Mahindra is seeking a Senior Site Reliability Engineer to align product and customer priorities with operational needs. You will participate in end-to-end service lifecycle, design reviews, and continuous improvement to enhance reliability and customer experience.

The role emphasizes managing operational resilience, incident response, and collaboration with a global team across multiple geographies. A strong Unix/Linux background and AWS experience are essential.

Qualifications

  • Strong experience in operational resiliency and self-healing systems.
  • Proficiency with Unix/Linux systems and network concepts.
  • Hands-on experience with AWS infrastructure and secure access practices.
  • Knowledge of ITSM processes, incident management and compliance frameworks.

Responsibilities

  • Engage in and improve the lifecycle of services from design to operations.
  • Maintain services by monitoring availability, latency and health.
  • Lead DevOps practices and automation to improve reliability and velocity.
  • Support CI/CD pipelines and promote software through environments.
  • Collaborate with global teams across geographies and time zones.

Skills

Unix/Linux
AWS
DevOps
CI/CD
Troubleshooting

Tools

Jenkins
Ansible
Venafi
Git

Job description

Tech Mahindra offers technology consulting and digital solutions to global enterprises across industries, enabling transformative scale at unparalleled speed. With 149k+ professionals across 90+ countries helping 1100+ clients, TechM provides a full spectrum of services including consulting, information technology, enterprise applications, business process services, engineering services, network services, customer experience & design services, AI & analytics, and cloud & infrastructure services. It is the first Indian company in the world to have been awarded the Sustainable Markets Initiative’s Terra Carta Seal, in recognition of actively leading the charge to create a climate and nature-positive future. Tech Mahindra (NSE: TECHM) is part of the Mahindra Group, founded in 1945, one of the largest and most admired multinational federations of companies.

Experience: 7+ years

Work model: 3 days per week work from client office

Job Description

Ultimately, the role of SRE is to align Product and Customer Focused priorities with Operational needs. We regularly review our run state not only from an internal perspective, but also understanding and providing the feedback loop to our development partners on how we can improve the customer experience of our applications.

  • Engage in and improve the whole lifecycle of services—from inception and design, through deployment, operations, and refinement.
  • Analyze ITSM activities of the platform and provide feedback loop to development teams on operational gaps or resiliency concerns
  • Support services before they go live through activities such as system design consulting, capacity planning and launch reviews.
  • Maintain services once they are live by measuring and monitoring availability, latency and overall system health.
  • Scale systems sustainably through mechanisms like automation and evolve systems by pushing for changes that improve reliability and velocity.
  • Support the application CI/CD pipeline for promoting software into higher environments through validation and operational gating, and lead DevOps automation and best practices.
  • Practice sustainable incident response and blameless postmortems.
  • Take a holistic approach to problem solving, by connecting the dots during a production event thru the various technology stack that makes up the platform, to optimize mean time to recover
  • Collaborate with a global team spread across tech hubs in multiple geographies and time zones
  • Share knowledge and mentor junior resources.
  • Develop and maintain automation pipelines for certificate renewal, traffic routing, alerting, and compliance reporting using tools like Ansible, Venafi.
  • Drive improvements in ITSM and DQ SLOs, ensuring timely CRQ status updates and incident closure.
  • Lead initiatives for Safety & Soundness and Operational Excellence across quarterly EPICs, covering areas such as PCI compliance, threat/toil management, self-healing, and ITSM defect resolution
All About You
  • Background in operational resiliency and self-healing systems.
  • Understanding of two factor authentication.
  • Strong documentation and communication skills.
  • Strong in Unix/Linux
  • Intermediate understanding of Active Directory (Users / Groups), SAML, LTPA, SSO, Oauth.
  • Understanding of DEVOPS technologies like Chef, Jenkins, Groovy, shell scripting, bitbucket, GIT.
  • Experience in working with or implementing automation workflows and/or scripting development.
  • Understanding of:
  • Client-server relationships
  • Network concepts (Layer 1 to Layer 3)
  • Stack trace analysis (TCP dumps, heap dumps, CPU/memory analysis, thread dumps).
  • Load balancers and application firewalls.
  • Logging and monitoring methods, standards, and tools.
  • High availability and business continuity planning
  • Caching concepts
  • Configuration management
  • Awareness of security implementations, certificate management lifecycle, mutual TLS, SSL handshake, SSH keys, symmetric and asymmetric encryptions.
  • o Experience with AWS infrastructure and secure access practices.
  • Familiarity with ITSM processes, compliance frameworks, and incident management.
  • Excellent communication and collaboration skills across cross-functional teams

Tech Mahindra is an Equal Employment Opportunity employer. We promote and support a diverse workforce at all levels of the company. All qualified applicants will receive consideration for employment without regard to race, religion, color, sex, age, national origin or disability. All applicants will be evaluated solely on the basis of their ability, competence, and performance of the essential functions of their positions.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Cloud Infrastructure Engineer
Senior Cloud Infrastructure Engineer

Tech Mahindra • Dublin

Hybrid
EUR 90,000 - 130,000
Site Reliability Engineering Technical Lead
Site Reliability Engineering Technical Lead

AMCS Group • Dublin

On-site
EUR 110,000 - 150,000
Site Reliability Engineering Technical Lead
Site Reliability Engineering Technical Lead

AMCS Group • Leinster

On-site
EUR 90,000 - 130,000
SRE (Application Support + Dev-Ops + Automation)
SRE (Application Support + Dev-Ops + Automation)

Fulcrum Digital • Dublin

On-site
EUR 90,000 - 120,000
DevOps SRE
DevOps SRE

Test Triangle • Dublin

Hybrid
EUR 60,000 - 90,000
BizOps SRE
BizOps SRE

HCLTech • Dublin

Hybrid
EUR 70,000 - 120,000
Site Reliability Engineering Technical Lead
Site Reliability Engineering Technical Lead

AMCS Group • Limerick

On-site
EUR 110,000 - 140,000
DEVOPS LEAD L1(CONTRACT)
DEVOPS LEAD L1(CONTRACT)

Wipro Technologies • Dublin

On-site
EUR 90,000 - 120,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Harvey Nash • Dublin

On-site
EUR 90,000 - 130,000
Development & Product Management Site Reliability Engineering Technical Lead Dublin, Ireland
Development & Product Management Site Reliability Engineering Technical Lead Dublin, Ireland

AMCS Group • Dublin

On-site
EUR 90,000 - 130,000