Senior Site Reliability Engineer IRC304301

Hitachi Vantara Corporation

Argentina

Híbrido

ARS 1.800.000 - 3.200.000

Jornada completa

Hace 3 días
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Consigue una respuesta de este empleador — un currículum y una carta de presentación adaptados exactamente a lo que busca para contratar.

Supera los filtros ATS

Ventajas ofrecidas por este puesto de trabajo

Exciting Projects
Collaborative Environment
Work-Life Balance
Professional Development
Excellent Benefits

Descripción de la vacante

GlobalLogic is seeking an experienced Senior SRE/Platform Engineer to lead reliability efforts for AI workflows and internal developer platforms. You will shape observability patterns, drive CI/CD enablement, and mentor engineers across multiple teams worldwide.

The role emphasizes hands-on leadership, architecture planning, and cost optimization in AWS, with opportunities to work across global centers and client facilities.

Formación

  • Experience leading reliability initiatives and platform engineering efforts.
  • Strong knowledge of AWS multi-region architectures and observability.
  • Proficient with Terraform and GitOps for IaC and deployment.
  • Deep experience with Datadog dashboards, traces, and SLOs.
  • Hands-on with Docker and Kubernetes in production environments.
  • Familiarity with CI/CD tools and internal developer platforms.

Responsabilidades

  • Manage SLOs, SLIs, and error budgets for AI reliability pipelines.
  • LeadCI/CD enablement and promote golden path templates across teams.
  • Maintain Datadog telemetry and service catalog visibility.
  • Oversee Terraform/GitOps deployments and cloud cost optimization in AWS.
  • Drive cross-team reliability initiatives and mentor engineers.

Conocimientos

SRE leadership
Cloud architecture
Observability
Platform engineering
Mentorship
Communication
Problem solving
First principles thinking
Cross-team collaboration
Leadership
Strategic initiative

Herramientas

AWS
EKS/Lambda/CloudWatch
Terraform
GitOps
Datadog
Docker
Kubernetes
Bitbucket Pipelines
GitHub Actions
Backstage
Compass
Python
Bash scripting

Descripción del empleo

Description

The project is building the reliability and AI operations foundation for its next chapter-an AI-first intelligence platform that runs the most demanding semiconductor intelligence workflows in the world. The SRE team operates as a technical leader within our engineering organization, responsible for defining reliability patterns for AI agent pipelines, architecting observability, and building an Internal Developer Platform (IDP). We provide a fast-scaling environment where reliability is prioritized from day one, offering the opportunity to apply deep SRE expertise to cutting-edge AI workloads and agentic systems.

Requirements
  • Technical Skills:
  • Experience: 6-8 years in SRE, Platform Engineering, or DevOps in senior or technical leadership roles.
  • Cloud & Infrastructure: Deep expertise in AWS (EKS, Lambda, CloudWatch) and multi-region architecture.
  • Infrastructure as Code: Proficiency in Terraform and GitOps practices.
  • Observability: Advanced operational proficiency with Datadog (dashboards, tracing, and SLO management).
  • Containerization: Solid experience with Docker and Kubernetes.
  • CI/CD & IDP: Experience with CI/CD tools (Bitbucket Pipelines or GitHub Actions) and internal developer platforms like Backstage or Compass.
  • Automation: Demonstrated proficiency in Python and/or Bash scripting.
  • Professional Skills:
  • Leadership: Ability to independently lead strategic reliability initiatives.
  • Problem Solving: A "first principles" approach to resolving complex technical issues.
  • Mentorship: A strong commitment to mentoring junior and intermediate engineers.
  • Communication: Excellent verbal and written communication skills to engage stakeholders across engineering, product, and leadership teams.
Job responsibilities
  • AI Reliability & Operations: Manage SLOs, SLIs, and error budgets. Design and implement patterns for LLM observability and recovery to ensure the reliability of AI-agent pipelines.
  • Engineering Enablement: Act as the primary SRE liaison for software and AI teams. Optimize CI/CD pipelines and drive the adoption of "golden path" templates and SRE best practices across the organization.
  • Observability & Service Catalog: Maintain Datadog for telemetry and performance tracing of pipeline services. Manage and evolve the internal service catalog to provide clear visibility into architecture.
  • FinOps & Infrastructure Management: Oversee infrastructure-as-code deployments (Terraform/GitOps) while actively managing and optimizing cloud infrastructure costs in AWS.
  • Strategic Growth: Lead cross-team reliability initiatives, mentor team members, and ensure the engineering culture prioritizes platform robustness and technical excellence.
What we offer
  • Exciting Projects: Come take your place at the forefront of digital transformation! With clients across all industries and sectors, we offer an opportunity to work on market-defining products using the latest technologies.
  • Collaborative Environment: Expand your skills by collaborating with a diverse team of highly talented people in an open, laidback environment - or even abroad in one of our global centers or client facilities!
  • Work-Life Balance: GlobalLogic prioritizes work-life balance, which is why we offer flexible work schedules.We offer you the best quality of work life so that you exceed the expectations of our clients, while achieving your professional and personal ambitions.
  • Professional Development: Our dedicated Learning & Development team regularly organizes English classes, professional certifications, and technical and soft skill trainings. We also offer the chance to travel internationally
  • Excellent Benefits: We provide our employees with competitive salaries, family medical insurance, extended paternity leave, annual performance bonuses, and referral bonuses.
About GlobalLogic

GlobalLogic, a Hitachi Group Company, is a trusted digital engineering partner to the world's largest and most forward-thinking companies. Since 2000, we’ve been at the forefront of the digital revolution - helping create some of the most innovative and widely used digital products and experiences. Today we continue to collaborate with clients in transforming businesses and redefining industries through intelligent products, platforms, and services.

Consigue la evaluación confidencial y gratuita de tu currículum.

o arrastra y suelta tu archivo aquí

Similar jobs

Puestos de trabajo similares que vale la pena comparar

Senior Site Reliability Engineer IRC304301
Senior Site Reliability Engineer IRC304301

t2s - Group International . your partner in executive search • Argentina

A distancia
ARS 2.400.000 - 4.200.000
Senior Site Reliability Engineer IRC304301
Senior Site Reliability Engineer IRC304301

GlobalLogic Inc. • Argentina

Presencial
ARS 182.249.000 - 273.373.000
Exciting Projects
Collaborative Environment
Work-Life Balance
+2
Principal DevOps AI IRC303057
Principal DevOps AI IRC303057

GlobalLogic • Argentina

Presencial
ARS 133.982.000 - 193.530.000
Principal DevOps AI IRC303057
Principal DevOps AI IRC303057

GlobalLogic • Municipio de Rincón de los Sauces

Presencial
ARS 3.000.000 - 6.000.000
Competitive salary
Global exposure and travel
Professional development
+1
Senior SRE - AI Reliability & Platform Engineer
Senior SRE - AI Reliability & Platform Engineer

t2s - Group International . your partner in executive search • Argentina

A distancia
ARS 2.400.000 - 4.200.000
AI Reliability SRE — Observability & Platform Lead
AI Reliability SRE — Observability & Platform Lead

Hitachi Vantara Corporation • Argentina

Híbrido
ARS 1.800.000 - 3.200.000
Exciting Projects
Collaborative Environment
Work-Life Balance
+2
Senior SRE: AI Reliability & Internal Platform Lead
Senior SRE: AI Reliability & Internal Platform Lead

GlobalLogic Inc. • Argentina

Presencial
ARS 182.249.000 - 273.373.000
Exciting Projects
Collaborative Environment
Work-Life Balance
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

N-iX • Argentina

Presencial
ARS 135.892.000 - 196.289.000
Flexible work format
Education reimbursement
Mentorship program
+2
Senior SQL Database Developer (.NET Environment) IRC295949
Senior SQL Database Developer (.NET Environment) IRC295949

Hitachi Vantara Corporation • Buenos Aires

Presencial
ARS 2.000.000 - 3.500.000
SR Advisory Consultant
SR Advisory Consultant

Globallogic • Buenos Aires

Híbrido
ARS 135.820.000 - 226.367.000
Competitive salaries
Family medical insurance
Extended paternity leave
+2