Site Reliability Engineer (SRE)

APPLY

Comisión de Fomento de Perú

Híbrido

ARS 105.910.000 - 166.430.000

Jornada completa

hace 33 horas
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Destaca para este puesto: genera un currículum y una carta de presentación adaptados en cuestión de un minuto.

Supera los filtros ATS

Ventajas ofrecidas por este puesto de trabajo

Agentic Delivery
Inclusive culture
AI upskilling budget
Generous vacation
Flexible work arrangements

Descripción de la vacante

APPLY is seeking an SRE Engineer to join our globally distributed team. This hybrid role focuses on reliability engineering, incident response, and improving observability across multiple UK e-commerce clients.

You will collaborate with the Service Desk and EMEA teams to ensure smooth operations and on‑call coverage. Ideal candidates have 2–3 years of SRE or platform operations experience, strong English skills, and hands‑on work with Grafana/Prometheus, Docker, and Kubernetes.

Formación

  • Excellent written and verbal English communication.
  • 2–3 years in SRE, platform operations, or technical Service Desk.
  • Experience with monitoring/observability tools (Grafana/Prometheus) and incident management.
  • Experience supporting ecommerce platforms and working in Agile environments.
  • Scripting skills in shell and/or Python for automation.
  • Familiarity with containerization concepts (Docker, Kubernetes).
  • Strong documentation habits and proactive problem solving.

Responsabilidades

  • Monitor platform health across client environments and maintain observability dashboards.
  • Respond to incidents, triage and escalate per runbooks.
  • Contribute to postmortems and learnings from incidents.
  • Support Service Desk with triage and incident resolution.
  • Contribute to automation to reduce toil and improve reliability.
  • Participate in on‑call rotations and shift handoffs with EMEA teams.

Conocimientos

English proficiency
SRE experience
Monitoring tools
Incident management
Cloud knowledge
Docker/Kubernetes
Agile/Scrum
Scripting (Shell/Python)

Herramientas

Grafana
Prometheus
Docker
Kubernetes
Jira
GitHub Actions

Descripción del empleo

About Apply

APPLY is the Agentic Customer Experience (ACx) partner for the world's most ambitious consumer and entertainment brands. We bring together deep domain expertise across Retail, CPG, Sports, and Media with AI-native delivery capability, designing and delivering agentic solutions that turn CX vision into commercial reality. We are the partner of choice for brands like Arc'teryx, NFL, Lululemon, and Kraft Heinz. For more information, visit applydigital.com .


About Apply

APPLY is the Agentic Customer Experience (ACx) partner for the world's most ambitious consumer and entertainment brands. We bring together deep domain expertise across Retail, CPG, Sports, and Media with AI-native delivery capability, designing and delivering agentic solutions that turn CX vision into commercial reality. We are the partner of choice for brands like Arc'teryx, NFL, Lululemon, and Kraft Heinz. For more information, visit applydigital.com .


LOCATION

APPLY is hybrid/remote-friendly. The preferred candidate should be based in Latin America, preferably working in hours that align to PT (Pacific Timezone) or ET (Eastern Timezone) . Candidates located in Santiago, Chile are able to work out of our Santiago office as remote/hybrid employees. Candidates located outside of Santiago, Chile will be fully remote employees.


THE ROLE

Apply Digital is looking for an SRE Engineer to join our globally distributed team. This is a hybrid SRE/Service Desk role designed for someone who is passionate about reliability engineering and comfortable supporting day-to-day operational needs across multiple UK e-commerce clients.


You will be a key contributor in maintaining the health and performance of client platforms, responding to incidents, and continuously improving observability and operational processes. While your primary focus is SRE, you will collaborate closely with the Service Desk team to support triaging, escalation, and resolution workflows.


This role is ideal for someone with 2–3 years of experience who thrives in a fast-paced, multi-client environment, values clear documentation, and is comfortable working with a high degree of autonomy during their shift.


What You’ll Do


  • Monitor platform health across multiple client environments using tools like Grafana and Prometheus, or other monitoring tools

  • Respond to and triage incidents, following established runbooks and escalation paths

  • Participate in post-incident reviews and contribute to postmortem documentation

  • Support the Service Desk team with technical triaging, incident classification, and resolution

  • Maintain and improve observability dashboards, alerts, and SLI, and SLO tracking

  • Write and maintain runbooks, operational documentation, and knowledge base articles

  • Identify recurring issues and propose automation or process improvements to reduce toil

  • Participate in on-call rotation covering weekends (alternating schedule — one weekend on, one weekend off)

  • Collaborate with the EMEA team during shift overlap to ensure smooth handoffs and continuity

  • Support root cause analysis and contribute to continuous improvement initiatives


WHAT WE’RE LOOKING FOR


  • Strong proficiency in English (written and verbal communication) is required

  • 2–3 years of experience in SRE, platform operations, or a technical Service Desk role

  • Experience with monitoring and observability tools such as Grafana, Prometheus, or equivalent

  • Solid understanding of incident management processes (triaging, escalation, postmortems)

  • Experience supporting e-commerce platforms

  • Scripting skills in shell and/or Python for automation and operational tasks

  • Familiarity with containerization concepts (Docker, Kubernetes) at an operational level

  • Experience working in Agile environments and using ticketing tools (e.g. Jira)

  • Comfort working independently during early‑morning shifts with minimal supervision

  • Strong documentation habits and attention to detail

  • Experience with Agile processes, testing, and code review

  • Strong experience with scripting - shell, Python, etc.

  • Excellent customer service attitude, communication skills (written and verbal), and interpersonal skills

  • Excellent analytical and problem‑solving skills

  • Ability to communicate effectively with technical and non‑technical stakeholders. You should feel comfortable explaining technical concepts in simple terms

  • Experience working in fast-paced, Agile environments, balancing priorities across multiple projects


NICE TO HAVE


  • Experience with Google Cloud Platform (GCP) or other major cloud providers (AWS, Azure)

  • Familiarity with CI/CD pipelines (GitHub Actions, GitLab CI)

  • Basic experience with Infrastructure as Code tools such as Terraform

  • Basic knowledge of AIOps concepts and their application in operational workflows

  • SRE or cloud certifications (Google Cloud, AWS, Kubernetes)


LIFE AT APPLY

People are at the core of everything we do at APPLY. We provide you with modern tools, systems and approaches, value your time, safety, and health, and strive to build a work community where you can thrive and grow. Here are a few benefits we offer to support you:



  • Agentic Delivery: Our people work in a modern way to deliver client outcomes. Broaden your skills on a range of engagements with international brands that have a global impact.

  • An inclusive and safe environment: We’re truly committed to building a culture where you are celebrated and everyone feels welcome and safe.

  • AI & Strategic Upskilling: Accelerate your professional growth with generous training budgets and mentorship, with a specific focus on Agentic AI expertise and the critical human skills required for the future of work.

  • Generous vacation policy: Work‑life balance is key to our team’s success, so we offer ample time away from work to promote overall well‑being.

  • Flexible work arrangements: We work in a variety of ways, from remote, to in‑office, to a blend of both.


APPLY is a safe, respectful, and inclusive community where differences are celebrated. We are committed to equal opportunity and fostering a workplace where everyone belongs. Learn more in our Diversity, Equity, and Inclusion (DEI) section. For recruitment accommodations, please email [email protected] .


We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

APPLY • Argentina

Híbrido
ARS 75.496.000 - 105.694.000
Agentic Delivery
Inclusive culture
Training budget
+2
Service Desk Engineer (Weekend)
Service Desk Engineer (Weekend)

APPLY • Argentina

Híbrido
ARS 18.119.000 - 37.748.000
Flexible work arrangements
Generous vacation policy
AI upskilling
+1
Senior Product Manager
Senior Product Manager

APPLY • Argentina

Híbrido
ARS 135.892.000 - 211.388.000
Generous vacation policy
Flexible work arrangements
Inclusive culture
+1
Director, Solutions Strategy
Director, Solutions Strategy

APPLY • Argentina

Híbrido
ARS 181.190.000 - 271.784.000
Generous vacation
Flexible work
AI upskilling
+1
Director, Solutions Strategy
Director, Solutions Strategy

APPLY • Comisión de Fomento de Perú

Híbrido
ARS 226.487.000 - 286.883.000
Agentic Delivery
Inclusive environment
AI Upskilling budget
+2
Senior Site Reliability Engineer (SRE)
Senior Site Reliability Engineer (SRE)

EPAM Systems • Argentina

Presencial
ARS 63.416.000 - 102.674.000
Healthcare benefits
Employee financial programs
Paid time off and sick leave
+2
Senior Site Reliability Engineer IRC302878
Senior Site Reliability Engineer IRC302878

GlobalLogic • Buenos Aires

Híbrido
ARS 8.928.000 - 13.392.000
Exciting projects
Collaborative environment
Work-life balance
+2
Senior Site Reliability Engineer IRC302878
Senior Site Reliability Engineer IRC302878

GlobalLogic • Argentina

Presencial
ARS 1.200.000 - 1.800.000
Exciting Projects
Collaborative Environment
Work-Life Balance
+2
Service Delivery Engineer (Remote - Argentina, Mexico or Ecuador)
Service Delivery Engineer (Remote - Argentina, Mexico or Ecuador)

AppDome • Argentina

A distancia
ARS 59.548.000 - 104.209.000
Stock options
Senior Product Manager
Senior Product Manager

Opportunities with AppDirect's Advisor Partners (Recruitment as a Service) • Buenos Aires

Híbrido
ARS 2.600.000 - 4.200.000
Flexible hybrid schedule
PTO 15 days
Medical coverage OSDE Plan 310
+2