SRE SR

Werben HR

Buenos Aires

Remote

ARS 136,732,000 - 197,502,000

Full time

12 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Werben HR is seeking a Senior Site Reliability Engineer to join a distributed, on-call capable team. You will design and implement automated, scalable infrastructure across on-prem and cloud resources, focusing on reliability and performance.

The role emphasizes collaboration, code-first thinking, and rapid prototyping within an agile environment. The position offers remote work with global collaboration across Argentina, Brazil, Colombia, and beyond, requiring strong Linux, Python, Kubernetes,

Qualifications

  • Experience with Linux-based systems, optimization, and troubleshooting.
  • Strong Python scripting and automation abilities.
  • Hands-on experience with container orchestration (Kubernetes) and container tech (Docker/Podman).
  • Proficient with IaC tools (Helm, Terraform, Ansible) to manage scalable infra.
  • Experience configuring and maintaining CI/CD pipelines (GitLab preferred).
  • Knowledge of monitoring/logging stacks (Prometheus, Grafana, ELK).
  • Familiarity with relational databases and performance tuning in distributed systems.
  • Agile development experience and collaborative mindset.
  • Cloud exposure (AWS/GCP) a strong plus.
  • Effective communication and teamwork across distributed teams.

Responsibilities

  • Architect and deploy As-A-Service solutions to automate system management, scaling, and monitoring.
  • Develop tools to optimize deployment, monitoring, and incident management at scale.
  • Collaborate with development and operations to improve reliability and drive DevOps initiatives.
  • Set up and maintain monitoring, alerting, and on-call rotations for incident response.
  • Improve CI/CD pipelines and infrastructure provisioning with IaC techniques.
  • Leverage Kubernetes for container orchestration; familiarity with Slurm is a plus.
  • Utilize cloud resources (AWS/GCP) and IaC to automate provisioning and scaling.

Skills

Linux Systems
Python Proficiency
Container Orchestration
Infrastructure as Code
Container Technologies
CI/CD Pipelines
Monitoring & Logging
Relational Databases
Agile Development
Cloud Experience
Collaboration & Communication

Tools

GitLab
GitHub
Git
Helm
Terraform
Ansible
Docker
Podman
Kubernetes
Slurm
Prometheus
Grafana
ELK Stack

Job description

Locations*: Buenos Aires, Argentina, is preferable; other locations are in Argentina, Brazil, Colombia, Peru, Chile, Mexico, Bolivia, Spain, Serbia (Belgrade), Czechia (Prague), Ukraine, Portugal. Type of work: Remote, full-time.

We are seeking a Senior SRE engineer to join a team that works on a complex distributed architecture, spanning physical machines - and virtualizing on-prem host/cloud computing. The role is to help set up centralized DevOps and help existing teams adopt more centralized best practices. The ideal candidate will have the ability to manage complexity and tackle problems across multiple stack layers as a part of a small team championing operational excellence. Our environment is relaxed yet intellectually intense. Our teams are lean and agile, which means rapid prototyping of products with immediate user feedback. We seek people who think in code, aspire to solve undiscovered computer science challenges, and are motivated by being around like-minded people. In fact, of the 600 employees globally, approximately 500 of them code daily. The customer develops and deploys systematic financial strategies across a variety of asset classes and global markets. We seek to produce high-quality predictive signals (alphas) through our proprietary research platform to employ financial strategies focused on exploiting market inefficiencies. Our teams work collaboratively to drive the production of alphas and financial strategies – the foundation of a sustainable, global investment platform.

Key Responsibilities:
  • Architecture and Automation: Design and deploy As-A-Service solutions using open-source software to automate system management, scaling, and monitoring.
  • System Optimization: Develop tools to streamline deployment, monitoring, and incident management for large-scale, distributed environments.
  • Collaboration Across Teams: Work with development and operations teams to design and implement software solutions that enhance the overall reliability of services. Contribute to the ongoing DevOps and Agile transformation.
  • Monitoring & Incident Response: Set up, configure, and maintain monitoring and alerting systems to ensure real-time visibility into system performance. Participate in on-call rotations to respond to incidents and mitigate downtime.
  • CI/CD & Infrastructure Management: Continuously improve CI/CD pipelines using tools like GitLab, Helm, Terraform, and Ansible, ensuring fast, safe, and reliable deployments.
  • Container Orchestration: Leverage container orchestration platforms like Kubernetes (K8S) to manage distributed systems at scale. Experience with Slurm or similar cluster management is a plus.
  • Cloud and Automation Tools: Use cloud infrastructure (AWS, GCP, etc.) and Infrastructure as Code (IaC) tools to automate the provisioning and scaling of resources.
Key Skills and Requirements:
  • Linux Systems: Deep expertise and hands-on experience working with Linux-based systems, with a focus on optimization and troubleshooting.
  • Python Proficiency: Strong skills in Python for scripting, automation, and system management.
  • Containerization & Orchestration: In-depth knowledge of container orchestration technologies such as Kubernetes (K8S). Experience with other cluster management tools like Slurm is a plus.
  • Infrastructure as Code (IaC): Hands-on experience with tools like Helm, Terraform, and Ansible to manage infrastructure in a scalable and automated way.
  • Container Technologies: Strong working knowledge of Docker, Podman, or other containerization systems to enable efficient and consistent deployment.
  • CI/CD Pipelines: Experience working with CI/CD tools, especially GitLab (preferred), GitHub, or Git, to ensure smooth and rapid delivery cycles.
  • Monitoring & Logging: Experience with monitoring and logging solutions such as Prometheus, Grafana, and the ELK stack to provide comprehensive insights into system performance and health.
  • Relational Databases: Understanding of relational databases, their performance tuning, and management in distributed systems.
  • Agile Development: Familiarity with Agile development methodologies, with a focus on continuous improvement and collaboration.
  • Cloud Experience: Exposure to cloud technologies such as AWS or Google Cloud (GCP) is a strong plus.
  • Collaboration & Communication: A team-first attitude with excellent verbal and written communication skills in English, able to work collaboratively with peers across the organization.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Cloud/Devops Engineer Id92208
Senior Cloud/Devops Engineer Id92208

Agileengine • San Carlos de Bariloche

Remote
ARS 182,485,000 - 228,106,000
Professional growth : Mentorship, Tech
USD-based pay with education, fitness,
Exciting projects : Fortune 500
+1
SRE (U$D pay) – ID #00057
SRE (U$D pay) – ID #00057

Werben HR • Buenos Aires

Remote
ARS 83,717,000 - 146,504,000
Senior Devsecops Engineer — Aws Cloud Platform & Sre
Senior Devsecops Engineer — Aws Cloud Platform & Sre

Talan • Ciudad de Mendoza

On-site
ARS 212,695,000 - 273,465,000
Professional growth
Competitive compensation
Exciting projects
+1
Senior Site Reliability Engineer IRC304301
Senior Site Reliability Engineer IRC304301

t2s - Group International . your partner in executive search • Argentina

On-site
ARS 2,400,000 - 4,200,000
Senior Devops Engineer — Remote (Aws/Eks/Terraform)
Senior Devops Engineer — Remote (Aws/Eks/Terraform)

Flux It • Partido de Quilmes

Hybrid
ARS 182,310,000 - 273,465,000
Professional growth
Competitive pay
Exciting projects
+1
Senior Devops Engineer – Remote (Aws, Kubernetes)
Senior Devops Engineer – Remote (Aws, Kubernetes)

Flux It • Partido de Quilmes

Hybrid
ARS 182,771,000 - 274,156,000
Professional growth
Competitive compensation
Exciting projects
+1
Senior Cloud/Devops Engineer Id92208
Senior Cloud/Devops Engineer Id92208

Agileengine • Mar del Plata

Hybrid
ARS 182,485,000 - 273,727,000
Professional growth
Competitive compensation
Exciting projects
+1
Senior Backend Engineer — Salesforce Apis & Ai Automation
Senior Backend Engineer — Salesforce Apis & Ai Automation

Agileengine • Rosario

Remote
ARS 136,732,000 - 197,502,000
Growth without limits
Competitive compensation
Flexibility: remote work
+3
Staff Software Engineer (3) - R2
Staff Software Engineer (3) - R2

Marathon Talent • Buenos Aires

On-site
ARS 182,771,000 - 274,156,000
Senior Cloud/DevOps Engineer ID92208
Senior Cloud/DevOps Engineer ID92208

AgileEngine, LLC. • Rosario

Remote
ARS 136,430,000 - 197,065,000
Professional growth
Competitive USD-based pay
Exciting projects
+1