Site Reliability Engineer (SRE)

EPAM Systems

Argentina

Presencial

ARS 1.800.000 - 3.200.000

Jornada completa

Hace 7 días
Sé de los primeros/as/es en solicitar esta vacante

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Ventajas ofrecidas por este puesto de trabajo

Healthcare benefits
Paid time off and sick leave
Upskilling and certification courses
LinkedIn Learning access
Global career opportunities
Volunteer and community involvement

Descripción de la vacante

EPAM Systems is seeking a proactive Site Reliability Engineer (SRE) to bridge software development and operations. This role focuses on automating operations, scaling infrastructure, and keeping production environments highly available.

You will design, build, and maintain cloud infrastructure and CI/CD pipelines to support fast, safe software deployment. You will implement robust monitoring and alerting, define SLOs/SLIs, and respond to incidents with blameless post-mortems.

Formación

  • 2+ years in systems administration, DevOps, or SRE.
  • I​n-depth knowledge of automation and resilience practices.
  • English at least C1+.

Responsabilidades

  • Design, build, and maintain cloud infrastructure using IaC (Terraform, CloudFormation).
  • Build and optimize CI/CD pipelines for deployments and configuration management.
  • Implement logging, monitoring, and alerting with Prometheus, Grafana, Datadog.
  • Set SLOs and SLIs; plan capacity and performance.
  • Respond to incidents and lead root-cause analyses.

Conocimientos

Automation
English proficiency
To be familiar with DevOps culture

Herramientas

AWS
Azure
GCP
Docker
Kubernetes
Terraform
CloudFormation
Prometheus
Grafana
Datadog

Descripción del empleo

Our engineering team is looking to add a skilled and proactive Site Reliability Engineer (SRE). This position serves as the connective link between software development and systems operations. Software engineering principles will be applied to automate operations, scale infrastructure, and keep systems highly available, resilient, and performant. The core mission involves building, running, and safeguarding the production environments that power our applications, keeping downtime to a minimum while enabling fast, safe software deployment.

Responsibilities
  • Design, build, and maintain cloud infrastructure using modern Infrastructure as Code practices such as Terraform and CloudFormation
  • Build and optimize CI/CD pipelines to automate software deployments, configuration management, and repetitive operational tasks
  • Design and implement robust logging, monitoring, and alerting systems using tools such as Prometheus, Grafana, and Datadog
  • Establish clear Service Level Objectives (SLOs) and Service Level Indicators (SLIs)
  • Respond to production incidents and lead troubleshooting efforts to restore services
  • Conduct blameless post-mortems to identify root causes and prevent recurrence
  • Partner with software developers to optimize system performance and plan capacity
  • Ensure services can scale to handle growth and traffic spikes
Requirements
  • 2+ years of experience in systems administration, DevOps, or systems-focused software development
  • Proficiency in at least one scripting or programming language such as Python, Bash, Go, or Rust
  • Experience with public cloud providers such as AWS, Azure, or GCP, along with containerization tools such as Docker and Kubernetes
  • Understanding of Linux/Unix administration and networking fundamentals such as TCP/IP, DNS, and HTTP/SSL/TLS
  • Passion for automation, eliminating toil, and building resilient systems that fail gracefully
  • Advanced proficiency in English (C1+)
We offer
  • International projects with top brands
  • Work with global teams of highly skilled, diverse peers
  • Healthcare benefits
  • Employee financial programs
  • Paid time off and sick leave
  • Upskilling, reskilling and certification courses
  • Unlimited access to the LinkedIn Learning library and 22,000+ courses
  • Global career opportunities
  • Volunteer and community involvement opportunities
  • EPAM Employee Groups
  • Award-winning culture recognized by Glassdoor, Newsweek and LinkedIn
Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Senior Site Reliability Engineer (SRE)
Senior Site Reliability Engineer (SRE)

EPAM Systems • Argentina

Presencial
ARS 133.982.000 - 178.643.000
International projects
Global teams
LinkedIn Learning access
+2
Site Reliability Engineer: Build Resilient Cloud Infra & CI/CD
Site Reliability Engineer: Build Resilient Cloud Infra & CI/CD

EPAM Systems • Argentina

Presencial
ARS 1.800.000 - 3.200.000
Healthcare benefits
Paid time off and sick leave
Upskilling and certification courses
+3
Site Reliability Engineer
Site Reliability Engineer

EPAM Systems • Argentina

Presencial
ARS 134.672.073 - 179.562.764
Senior Site Reliability Engineer/Platform Engineer
Senior Site Reliability Engineer/Platform Engineer

Techunting • Córdoba

Presencial
ARS 89.314.000 - 119.087.000
Senior SRE: Automate, Scale, and Production Reliability
Senior SRE: Automate, Scale, and Production Reliability

EPAM Systems • Argentina

Presencial
ARS 133.982.000 - 178.643.000
International projects
Global teams
LinkedIn Learning access
+2
Senior DevOps / Site Reliability Engineer
Senior DevOps / Site Reliability Engineer

N-iX • Argentina

Híbrido
ARS 1.800.000 - 3.200.000
Flexible working format
Competitive salary
Career growth and mentorship
+3
Senior Site Reliability Engineer IRC302878
Senior Site Reliability Engineer IRC302878

GlobalLogic • Municipio de Rincón de los Sauces

Presencial
ARS 178.643.000 - 223.304.000
Competitive salary
Family medical insurance
Extended paternity leave
+2
Senior Site Reliability Engineer IRC302878
Senior Site Reliability Engineer IRC302878

GlobalLogic • Argentina

Presencial
ARS 1.200.000 - 1.800.000
Exciting Projects
Collaborative Environment
Work-Life Balance
+2
Senior Site Reliability Engineer IRC302878
Senior Site Reliability Engineer IRC302878

GlobalLogic • Buenos Aires

Híbrido
ARS 8.928.000 - 13.392.000
Exciting projects
Collaborative environment
Work-life balance
+2
Site Reliability Engineer: Automate, Monitor, and Scale
Site Reliability Engineer: Automate, Monitor, and Scale

EPAM Systems • Argentina

Presencial
ARS 134.672.073 - 179.562.764