SENIOR SITE RELIABILITY ENGINEER

Svitla

Argentina

Hybrid

ARS 91,242,000 - 136,863,000

Full time

45 hours ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Flexible workspace
15 vacation days
10 national holidays
10 sick leaves
Learning program
Tech webinars
Bonuses for referrals
Office or coworking options

Job summary

Svitla Systems Inc. is seeking a Senior Site Reliability Engineer in Argentina for a full-time role. You will lead migration of legacy AWS workloads to EKS, design reproducible infrastructure with Terraform, and package apps with Helm.

Youll develop Python/Bash automation, implement deep tracing with New Relic/Datadog, and manage incidents including on-call shifts. Strong Kubernetes, AWS, and observability experience are essential.

Qualifications

  • Bachelor's degree and 8+ years of professional experience handling large-scale production systems.
  • Hands-on experience designing and deploying EKS / AKS clusters.
  • Understanding of Kubernetes security best practices, including RBAC, network policies, and PodSecurityPolicies.
  • Ability to identify and resolve issues related to Kubernetes, networking, storage, and application deployments.
  • Experience migrating workloads to Kubernetes.
  • Experience with AWS or a comparable cloud provider, with certification.
  • Hands-on experience with Ruby, Terraform, and configuration management tools such as Chef, Ansible, or equivalent.
  • Excellent knowledge of large-scale web applications and distributed systems.
  • Experience with observability tools such as New Relic and Datadog.
  • Expertise in problem-solving and analyzing global-scale distributed systems.
  • Excellent written and verbal communication skills.
  • Critical thinking and a habit of continuously challenging how and why we do things in order to improve.

Responsibilities

  • Lead the migration of legacy AWS workloads to Amazon EKS (Elastic Kubernetes Service), using Terraform for reproducible infrastructure and Helm for application packaging.
  • Provide weekend support for migration activities.
  • Develop Python and Bash automation to streamline containerization, secret management (AWS Secrets Manager), and resource tagging.
  • Implement deep-trace monitoring with observability tools to maintain visibility during and after the migration.
  • Act as the primary point of contact for resolving Kubernetes incidents, including Pod CrashLoopBackOffs, OOMKills, and VPC CNI networking issues.
  • Manage the full incident lifecycle, from real-time troubleshooting (including on weekends) to post-mortem analysis, and build automated guardrails against recurrence.
  • Work closely with development, QA, and operations teams to ensure seamless collaboration and efficient workflows.
  • Coordinate incident, problem, and change management.
  • Participate in an on-call rotation for after-hours emergencies.

Skills

Kubernetes
AWS
Terraform
Python
Bash
Helm
RBAC
Network policies
PodSecurityPolicies
New Relic
Datadog
Incident management
Communication
On-call

Education

Bachelor's degree in a related field

Tools

EKS
AKS
Terraform
Helm
AWS
Chef
Ansible
Ruby
Python
Bash
New Relic
Datadog

Job description

Svitla Systems Inc. is looking for a Senior Site Reliability Engineer for a full-time position (40 hours per week) in Argentina.

Requirements
  • Bachelor's degree and 8+ years of professional experience handling large-scale production systems.
  • Hands-on experience designing and deploying EKS / AKS clusters.
  • Understanding of Kubernetes security best practices, including RBAC, network policies, and PodSecurityPolicies.
  • Ability to identify and resolve issues related to Kubernetes, networking, storage, and application deployments.
  • Experience migrating workloads to Kubernetes.
  • Experience with AWS or a comparable cloud provider, with certification.
  • Hands-on experience with Ruby, Terraform, and configuration management tools such as Chef, Ansible, or equivalent.
  • Excellent knowledge of large-scale web applications and distributed systems.
  • Experience with observability tools such as New Relic and Datadog.
  • Expertise in problem-solving and analyzing global-scale distributed systems.
  • Excellent written and verbal communication skills.
  • Critical thinking and a habit of continuously challenging how and why we do things in order to improve.
Responsibilities
  • Lead the migration of legacy AWS workloads to Amazon EKS (Elastic Kubernetes Service), using Terraform for reproducible infrastructure and Helm for application packaging.
  • Provide weekend support for migration activities.
  • Develop Python and Bash automation to streamline containerization, secret management (AWS Secrets Manager), and resource tagging.
  • Implement deep-trace monitoring with observability tools to maintain visibility during and after the migration.
  • Act as the primary point of contact for resolving Kubernetes incidents, including Pod CrashLoopBackOffs, OOMKills, and VPC CNI networking issues.
  • Manage the full incident lifecycle, from real-time troubleshooting (including on weekends) to post-mortem analysis, and build automated guardrails against recurrence.
  • Work closely with development, QA, and operations teams to ensure seamless collaboration and efficient workflows.
  • Coordinate incident, problem, and change management.
  • Participate in an on-call rotation for after-hours emergencies.
We offer
  • US and EU projects based on advanced technologies.
  • Competitive compensation based on skills and experience.
  • Regular performance appraisals to support your growth.
  • Flexibility in workspace, either remote, our welcoming office or local coworking.
  • Bonuses for recommendations of new employees.
  • Bonuses for article writing, public talks, other activities.
  • 15 vacation days, 10 national holidays, 10 sick leaves.
  • Personalized learning program tailored to your interests and skill development.
  • Free tech webinars and meetups organized by Svitla.
  • Fun corporate online\offline celebrations and activities.
  • Awesome team, friendly and supportive community!
Why join Svitla

Svitla Systems is a global digital solutions company headquartered in the U.S. and operating across the Americas, Europe, Asia, and APAC. Since 2003, we have served a wide range of clients - from innovative start-ups to Fortune 500 companies. Our success is built on partnership. By integrating seamlessly with clients' teams, we create lasting collaborations that drive real results. We are strong advocates of workplace flexibility, remote culture, individual approach to professional and personal growth.

Our global mission is to build a business that contributes to wellbeing of our partners, personnel, and their families, improves our communities, and makes a lasting difference in the world. Together, we are coding a brighter tomorrow - and living it.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Full Stack Engineer
Senior Full Stack Engineer

Svitla • Argentina

Hybrid
ARS 137,157,000 - 228,596,000
Remote or office
Bonuses for activities
Generous time off
+3
Backend Software Engineer (JavaKotlin)
Backend Software Engineer (JavaKotlin)

Svitla Systems, Inc. • Argentina

On-site
ARS 105,468,000 - 165,735,000
Remote/workspace flexibility
Learning program
Referral bonuses
+1
C++ Engineer
C++ Engineer

Svitla Systems, Inc. • Argentina

On-site
ARS 133,982,000 - 178,643,000
Bonuses for recommendations of new emp
Learning program tailored to interests
Webinars and meetups
Senior Full Stack Developer with Angular & Java/Spring Boot
Senior Full Stack Developer with Angular & Java/Spring Boot

Svitla Systems, Inc. • Argentina

On-site
ARS 60,397,000 - 135,892,000
Flexible workspace: remote/office/cowr
Bonuses for recommendations
Learning programs
+2
Analytics Engineer
Analytics Engineer

Svitla Systems, Inc. • Buenos Aires

On-site
ARS 1,200,000 - 1,800,000
Bonuses for recommendations
Personalized learning program
Flexible workspace options
+2
Senior Full Stack Engineer with Java and TypeScript
Senior Full Stack Engineer with Java and TypeScript

Svitla Systems, Inc. • Argentina

On-site
ARS 135,663,000 - 195,957,000
Remote options
Office or coworking flexibility
Learning program
+2
Senior Full Stack Engineer
Senior Full Stack Engineer

Svitla Systems, Inc. • Argentina

Hybrid
ARS 1,200,000 - 2,200,000
Private medical insurance
Performance appraisals
Remote work options
+3
Senior SRE: Kubernetes & Cloud Migration Lead (Remote)
Senior SRE: Kubernetes & Cloud Migration Lead (Remote)

Svitla • Argentina

Hybrid
ARS 91,242,000 - 136,863,000
Flexible workspace
15 vacation days
10 national holidays
+5
SENIOR FULLSTACK ENGINEER
SENIOR FULLSTACK ENGINEER

Svitla Systems, Inc. • Argentina

Remote
ARS 137,157,000 - 198,116,000
Remote work from LATAM
Bonuses for speaking and writing
Free tech webinars and meetups
+1
ELIGIBILITY AND ACCUMULATOR ANALYST
ELIGIBILITY AND ACCUMULATOR ANALYST

Svitla Systems, Inc. • Argentina

On-site
ARS 88,483,829 - 132,725,744