Senior Site Reliability Engineer (SRE) – Application Observability & Readiness (Azure)

Encora10

Comisión de Fomento de Perú

Vor Ort

ARS 121.858.000 - 182.788.000

Vollzeit

Vor 12 Tagen
Bewerbungsgenerator

Bekomme eine Antwort von diesem Arbeitgeber — ein Lebenslauf und ein Anschreiben, die genau auf die Eigenschaften eingehen, die gesucht werden.

Schaffe es an den ATS-Filtern vorbei

Zusammenfassung

Coforge is seeking a Senior Site Reliability Engineer (SRE) focused on Application Observability & Readiness in Azure. You will design monitoring, dashboards, and instrumentation across distributed applications, aligning with production readiness and incident learnings.

The role emphasizes hands-on Azure IaaS, DBT, Databricks, SQL, and strong English communication for global teams. Work remotely for residents across LATAM, collaborating with development teams to implement SLIs/SLOs, and drive

Qualifikationen

  • Strong experience with Microsoft Azure IaaS environments.
  • Hands-on observability, monitoring and reliability practices.
  • DBT, Databricks and SQL: minimum 1 year of experience.
  • Experience with APM solutions such as Application Insights or New Relic.
  • Dashboarding and monitoring design using Azure Monitor, Insights, and Log Analytics.
  • CI/CD familiarity with Azure DevOps and GitHub Actions.

Aufgaben

  • Collaborate with development teams to design monitoring, alerts, dashboards and APM instrumentation.
  • Lead implementation and optimization of APM solutions.
  • Apply observability best practices using Azure Monitor, Insights, New Relic and Log Analytics.
  • Enable code-level instrumentation, distributed tracing and structured logging.
  • Design and maintain monitoring dashboards and health metrics.
  • Define SLIs/SLOs and alerting strategies based on latency, errors, traffic and saturation.
  • Improve monitoring through production insights and incident learnings.
  • Participate in production readiness reviews; identify risks and gaps.
  • Support incident analysis and post-incident improvements.

Kenntnisse

Azure IaaS
Azure Monitor
Application Insights
New Relic
Databricks
DBT
SQL
Azure DevOps
GitHub Actions
Site Reliability Engineering (SRE)
Application Performance Monitoring (AP
Log Analytics (KQL)
Observability
Distributed Tracing
Structured Logging
English communication

Jobbeschreibung

Job Title

Senior Site Reliability Engineer (SRE) - Application Observability & Readiness (Azure)

Location and Work Mode

Location: Legal residents of Peru, Colombia, Bolivia, Costa Rica, Mexico, and Brazil. Work Mode: Remote

Key Skills

Azure IaaS, Azure Monitor, Application Insights, New Relic, Databricks, DBT, SQL, Azure DevOps, GitHub Actions, Site Reliability Engineering (SRE), Application Performance Monitoring (APM), Log Analytics (KQL), Observability, Distributed Tracing, Structured Logging

Experience

5+ years of experience in Site Reliability Engineering, Cloud Operations, or related roles. Mandatory minimum of 1 year of hands-on experience with DBT, Databricks, and SQL.

Main Responsibilities
  • Collaborate with development teams to design and implement monitoring, alerting, dashboards, and APM instrumentation across applications and services.
  • Lead the implementation, configuration, and optimization of Application Performance Monitoring (APM) solutions.
  • Apply observability best practices using tools such as Azure Monitor, Application Insights, New Relic, and Log Analytics (KQL).
  • Enable code-level instrumentation, distributed tracing, and structured logging to improve application visibility and reliability.
  • Design and maintain application-level monitoring dashboards and operational health metrics.
  • Define and implement Service Level Indicators (SLIs), Service Level Objectives (SLOs), and effective alerting strategies based on latency, error rates, traffic, and resource saturation.
  • Continuously improve monitoring and alerting mechanisms through production insights and incident learnings.
  • Participate in production readiness reviews, identifying operational risks, observability gaps, and potential failure scenarios before deployment.
  • Support incident analysis and post-incident improvements through enhanced telemetry and monitoring practices.
  • Partner with engineering teams to ensure applications are reliable, scalable, and production-ready.
Mandatory Requirements
  • Strong experience supporting and operating applications in Microsoft Azure IaaS environments.
  • Hands-on experience with application observability, monitoring, and reliability engineering practices.
  • Mandatory experience with DBT, Databricks, and SQL (minimum 1 year of experience).
  • Experience implementing and managing APM solutions such as Application Insights, New Relic, or similar platforms.
  • Experience designing dashboards and monitoring solutions using Azure Monitor, Application Insights, and Log Analytics (KQL).
  • Familiarity with CI/CD environments including Azure DevOps and GitHub Actions.
  • Solid understanding of cloud-native architectures and distributed application systems.
  • Practical SRE mindset with experience in incident analysis, root cause investigation, and proactive problem prevention.
  • Strong verbal and written English communication skills, with the ability to collaborate effectively with global teams.
Preferred Requirements
  • Experience with scripting and automation using PowerShell and/or Bash.
  • Knowledge of scalability, availability, and resilience patterns in modern cloud environments.
  • Experience driving production readiness and operational excellence initiatives.
  • Exposure to reliability engineering best practices in enterprise-scale environments.
Background and Commitment

At Coforge, we hire professionals solely based on their skills and qualifications and do not discriminate on the basis of age, disability, religion, gender, sexual orientation, socioeconomic status, or nationality.

Posting date

Posted on: 28-09-2026

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Senior Site Reliability Engineer (SRE) – Application Observability & Readiness (Azure)
Senior Site Reliability Engineer (SRE) – Application Observability & Readiness (Azure)

Encora10 • Argentinien

Vor Ort
MXN 900.000 - 1.300.000
Senior Site Reliability Engineer (SRE) – Application Observability & Readiness (Azure)
Senior Site Reliability Engineer (SRE) – Application Observability & Readiness (Azure)

Encora • Comisión de Fomento de Perú

Remote
ARS 136.430.000 - 227.383.000
Senior Site Reliability Engineer (Sre) Application Observability & Readiness (Azure)
Senior Site Reliability Engineer (Sre) Application Observability & Readiness (Azure)

Encora Inc. • Comisión de Fomento de Perú

Remote
ARS 137.078.000 - 198.002.000
Remote Azure SRE: Observability & Reliability Lead
Remote Azure SRE: Observability & Reliability Lead

Encora10 • Comisión de Fomento de Perú

Remote
ARS 121.858.000 - 182.788.000
Remote Azure SRE - Observability & Production Readiness
Remote Azure SRE - Observability & Production Readiness

Encora10 • Argentinien

Hybrid
MXN 900.000 - 1.300.000
Remote Senior SRE — Azure Observability & Readiness
Remote Senior SRE — Azure Observability & Readiness

Encora • Comisión de Fomento de Perú

Remote
ARS 136.430.000 - 227.383.000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

N-iX • Argentinien

Vor Ort
ARS 135.892.000 - 196.289.000
Flexible work format
Education reimbursement
Mentorship program
+2
Senior SRE - Azure Observability & Readiness (Remote)
Senior SRE - Azure Observability & Readiness (Remote)

Encora Inc. • Comisión de Fomento de Perú

Remote
ARS 137.078.000 - 198.002.000
Site Reliability Engineer
Site Reliability Engineer

Strategic Staffing Solutions • Argentinien

Remote
ARS 1.200.000 - 2.500.000
Full-time employment
Remote work 100%
Competitive salary in ARS
+1
SRE SR
SRE SR

Werben HR • Buenos Aires

Remote
ARS 136.732.000 - 197.502.000