SRE: Observability & AI-Driven Reliability

MetLife México

Fatih

On-site

TRY 320,000 - 540,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Private health insurance
Pension plan
Work from home allowance
Heritage Day off

Job summary

MetLife México is seeking a Site Reliability Engineer (SRE) – Observability & Elastic to join a high‑performing engineering organization that ships resilient, secure, and observable platforms across cloud and hybrid environments.

You will build and maintain observability pipelines (logs, metrics, traces) using Elastic Stack, create dashboards and alerts, define SLOs, and drive automated remediation in collaboration with development, security, and platform teams.

Qualifications

  • Bachelor’s degree in Computer Science, Engineering, or related field.
  • 3+ years of experience in SRE, DevOps, or production engineering roles.
  • Strong hands‑on experience with Elastic Stack (Elasticsearch, Logstash, Kibana).
  • Proficiency in monitoring tools and observability frameworks (metrics, distributed tracing, logging).

Responsibilities

  • Design, implement, and manage end‑to‑end observability solutions (metrics, logs, traces).
  • Build and maintain Elastic Stack based logging and monitoring platforms.
  • Develop dashboards, alerts, and visualization layers for proactive issue detection.
  • Define and continuously improve SLIs, SLOs, and alerting strategies.
  • Enable log, metric, and trace correlation to improve troubleshooting efficiency.

Skills

SRE/DevOps experience
English proficiency
Python scripting
Distributed systems
CI/CD familiarity

Education

Bachelor’s degree in Computer Science, Engineering

Tools

Elastic Stack (Elasticsearch, Logstash, Kibana)
Prometheus
Grafana
Azure Monitor
App Insights
Splunk

Job description

MetLife México is seeking a Site Reliability Engineer (SRE) – Observability & Elastic to join a high‑performing engineering organization that ships resilient, secure, and observable platforms across cloud and hybrid environments.

You will build and maintain observability pipelines (logs, metrics, traces) using Elastic Stack, create dashboards and alerts, define SLOs, and drive automated remediation in collaboration with development, security, and platform teams.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

MetLife México • Fatih

On-site
TRY 320,000 - 540,000
Private health insurance
Pension plan
Work from home allowance
+1
Senior SRE: GenAI-Driven Reliability & Cloud Ops
Senior SRE: GenAI-Driven Reliability & Cloud Ops

EPAM Systems, Inc. • Turkey

On-site
TRY 600,000 - 900,000
Private health insurance
Continuous upskilling & development
English courses
+1
SRE Leadership Manager - Scale & Reliability
SRE Leadership Manager - Scale & Reliability

n11 • Fatih

On-site
TRY 800,000 - 1,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

EPAM Systems, Inc. • Turkey

On-site
TRY 600,000 - 900,000
Private health insurance
Continuous upskilling & development
English courses
+1
Senior SRE: AI-Driven Cloud Reliability & Kubernetes
Senior SRE: AI-Driven Cloud Reliability & Kubernetes

Sezzle • Turkey

On-site
TRY 2,684,000 - 5,100,000
Site Reliability Engineering Manager
Site Reliability Engineering Manager

n11 • Fatih

On-site
TRY 800,000 - 1,000,000
Remote Site Reliability Engineer – AI-Driven Incident Response
Remote Site Reliability Engineer – AI-Driven Incident Response

Storyteller • Turkey

Remote
TRY 1,051,000 - 1,752,000
Senior DevOps Engineer: Observability & CI/CD (Remote)
Senior DevOps Engineer: Observability & CI/CD (Remote)

Fundraise Up • Fatih

On-site
31 days off
100% paid telemedicine plan
Home Office Setup Assistance
+5
Platform SRE: Reliability, Automation & Incident Ownership
Platform SRE: Reliability, Automation & Incident Ownership

ING Hubs Türkiye • Fatih

On-site
TRY 350,000 - 520,000
Meal Card
Transportation Allowance
MultiSport Membership
+4
Principal SRE: Scale Cloud Infra, Drive Reliability
Principal SRE: Scale Cloud Infra, Drive Reliability

Sezzle • Turkey

On-site
TRY 300,000 - 400,000