Senior Site Reliability Engineer IRC304301

GlobalLogic Inc.

Argentina

Presencial

ARS 182.249.000 - 273.373.000

Jornada completa

Hace 4 días
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Destaca en este puesto — genera un currículum adaptado y una carta de presentación en aproximadamente un minuto.

Supera los filtros ATS

Ventajas ofrecidas por este puesto de trabajo

Exciting Projects
Collaborative Environment
Work-Life Balance
Professional Development
Excellent Benefits

Descripción de la vacante

GlobalLogic seeks a Senior Site Reliability Engineer to lead reliability initiatives for an AI-first platform in Argentina. You will define observability patterns, build an Internal Developer Platform, and guide AI agent pipelines with a focus on robust SRE practices.

The role requires 6–8 years in SRE/DevOps, deep AWS knowledge, and strong scripting. You will collaborate across engineering, product, and leadership to ensure platform resilience.

Formación

  • 6–8 years in SRE, Platform Engineering, or DevOps in senior roles.
  • Deep AWS expertise (EKS, Lambda, CloudWatch) and multi-region architecture.
  • Terraform and GitOps proficiency.
  • Datadog dashboards, tracing, and SLO management experience.
  • Docker and Kubernetes experience.
  • CI/CD tools (Bitbucket Pipelines or GitHub Actions) and internal developer platforms like Backstage or Compass.
  • Python and/or Bash scripting proficiency.

Responsabilidades

  • Manage SLOs, SLIs, and error budgets for AI reliability pipelines.
  • Lead reliability initiatives and optimize CI/CD templates across teams.
  • Maintain Datadog telemetry and evolve the internal service catalog.
  • Oversee Terraform/GitOps deployments and cloud cost optimization in AWS.
  • Mentor team members and drive platform robustness and technical excellence.

Conocimientos

AWS
CloudWatch
Bash
DataDog
EKS/ECS
GitLab CI/CD
GitOps
Kubernetes
Lambda
Multi-Region Deployments
Python
SLO/SLI/SLA
Terraform

Herramientas

Backstage
Compass

Descripción del empleo

Senior Site Reliability Engineer IRC304301

Function

IT, Telecom & Internet

Experience

5-10 years

Location

Argentina

Skills

AWS, AWS Cloudwatch, Bash, DataDog, EKS/ECS, Gitlab CI/CD, GitOps, Kubernetes, Lambda, Multi-Region Deployments, Python, SLO/SLI/SLA Engineering, Terraform

The projectis building the reliability and AI operations foundation for its next chapter—an AI-first intelligence platform that runs the most demanding semiconductor intelligence workflows in the world. The SRE team operates as a technical leader within our engineering organization, responsible for defining reliability patterns for AI agent pipelines, architecting observability, and building an Internal Developer Platform (IDP). We provide a fast-scaling environment where reliability is prioritized from day one, offering the opportunity to apply deep SRE expertise to cutting-edge AI workloads and agentic systems.

Requirements

Technical Skills:
Experience: 6–8 years in SRE, Platform Engineering, or DevOps in senior or technical leadership roles.
Cloud & Infrastructure: Deep expertise in AWS (EKS, Lambda, CloudWatch) and multi-region architecture.
Infrastructure as Code: Proficiency in Terraform and GitOps practices.
Observability: Advanced operational proficiency with Datadog (dashboards, tracing, and SLO management).
Containerization: Solid experience with Docker and Kubernetes.
CI/CD & IDP: Experience with CI/CD tools (Bitbucket Pipelines or GitHub Actions) and internal developer platforms like Backstage or Compass.
Automation: Demonstrated proficiency in Python and/or Bash scripting.

Professional Skills:
Leadership: Ability to independently lead strategic reliability initiatives.
Problem Solving: A “first principles” approach to resolving complex technical issues.
Mentorship: A strong commitment to mentoring junior and intermediate engineers.
Communication: Excellent verbal and written communication skills to engage stakeholders across engineering, product, and leadership teams.

Job responsibilities

AI Reliability & Operations: Manage SLOs, SLIs, and error budgets. Design and implement patterns for LLM observability and recovery to ensure the reliability of AI-agent pipelines.
Engineering Enablement: Act as the primary SRE liaison for software and AI teams. Optimize CI/CD pipelines and drive the adoption of “golden path” templates and SRE best practices across the organization.
Observability & Service Catalog: Maintain Datadog for telemetry and performance tracing of pipeline services. Manage and evolve the internal service catalog to provide clear visibility into architecture.
FinOps & Infrastructure Management: Oversee infrastructure-as-code deployments (Terraform/GitOps) while actively managing and optimizing cloud infrastructure costs in AWS.
Strategic Growth: Lead cross-team reliability initiatives, mentor team members, and ensure the engineering culture prioritizes platform robustness and technical excellence.

What we offer

Exciting Projects: Come take your place at the forefront of digital transformation! With clients across all industries and sectors, we offer an opportunity to work on market-defining products using the latest technologies.

Collaborative Environment:Expand your skills by collaborating with a diverse team of highly talented people in an open, laidback environment — or even abroad in one of our global centers or client facilities!

Work-Life Balance:GlobalLogic prioritizes work-life balance, which is why we offer flexible work schedules.We offer you the best quality of work life so that you exceed the expectations of our clients, while achieving your professional and personal ambitions.

Professional Development:Our dedicated Learning & Development team regularly organizes English classes, professional certifications, and technical and soft skill trainings. We also offer the chance to travel internationally

Excellent Benefits:We provide our employees with competitive salaries, family medical insurance, extended paternity leave, annual performance bonuses, and referral bonuses.

About GlobalLogic

GlobalLogic, a Hitachi Group Company, is a trusted digital engineering partner to the world’s largest and most forward-thinking companies. Since 2000, we’ve been at the forefront of the digital revolution – helping create some of the most innovative and widely used digital products and experiences. Today we continue to collaborate with clients in transforming businesses and redefining industries through intelligent products, platforms, and services.

Gender * The gender information on this form helps us understand the makeup of our applicant pool in this key area, and to continuously improve our efforts to make our workforce more inclusive.

Consigue la evaluación confidencial y gratuita de tu currículum.

o arrastra y suelta tu archivo aquí

Similar jobs

Puestos de trabajo similares que vale la pena comparar

Senior Site Reliability Engineer IRC304301
Senior Site Reliability Engineer IRC304301

t2s - Group International . your partner in executive search • Argentina

A distancia
ARS 2.400.000 - 4.200.000
Senior Site Reliability Engineer IRC304301
Senior Site Reliability Engineer IRC304301

Hitachi Vantara Corporation • Argentina

Híbrido
ARS 1.800.000 - 3.200.000
Exciting Projects
Collaborative Environment
Work-Life Balance
+2
Principal DevOps AI IRC303057
Principal DevOps AI IRC303057

GlobalLogic • Argentina

Presencial
ARS 133.982.000 - 193.530.000
Senior SRE - AI Reliability & Platform Engineer
Senior SRE - AI Reliability & Platform Engineer

t2s - Group International . your partner in executive search • Argentina

A distancia
ARS 2.400.000 - 4.200.000
Principal DevOps AI IRC303057
Principal DevOps AI IRC303057

GlobalLogic • Municipio de Rincón de los Sauces

Presencial
ARS 3.000.000 - 6.000.000
Competitive salary
Global exposure and travel
Professional development
+1
Principal AI Engineer I (Python + AWS) – Advanced English) IRC303058
Principal AI Engineer I (Python + AWS) – Advanced English) IRC303058

GlobalLogic • Argentina

Presencial
ARS 89.322.000 - 133.982.000
Exciting Projects
Collaborative Environment
Work-Life Balance
+2
Senior Backend Java Software Engineer (with AWS experience & Advanced English) IRC304295
Senior Backend Java Software Engineer (with AWS experience & Advanced English) IRC304295

GlobalLogic • Buenos Aires

Presencial
ARS 181.560.000 - 242.079.000
Flexible work schedules
Global centers
English classes
+2
Senior SRE: AI Reliability & Internal Platform Lead
Senior SRE: AI Reliability & Internal Platform Lead

GlobalLogic Inc. • Argentina

Presencial
ARS 182.249.000 - 273.373.000
Exciting Projects
Collaborative Environment
Work-Life Balance
+2
Fullstack .NET/Phyton IRC303631
Fullstack .NET/Phyton IRC303631

GlobalLogic • Argentina

Presencial
ARS 90.697.000 - 181.395.000
Exciting projects
Collaborative environment
Work-life balance
+2
Senior Cloud DevSecOp Engineer IRC303764
Senior Cloud DevSecOp Engineer IRC303764

GlobalLogic • Argentina

Presencial
ARS 90.697.000 - 166.279.000
Exciting projects
Collaborative environment
Work-life balance
+2