Senior Site Reliability Engineer

N-iX

México

Híbrido

MXN 1.200.000 - 1.600.000

Jornada completa

Hace 2 días
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Destaca en este puesto — genera un currículum adaptado y una carta de presentación en aproximadamente un minuto.

Supera los filtros ATS

Ventajas ofrecidas por este puesto de trabajo

Flexible working format
Competitive compensation package
Professional development

Descripción de la vacante

N-iX is seeking a Senior Site Reliability Engineer to join our LATAM team. You will work on scalable, fault-tolerant cloud services hosted in Azure and on-premise data centers, focusing on reliability, uptime, and performance.

You will design and implement automated systems, lead major software components, and participate in post-incident analyses to drive continuous improvement. This role offers a flexible Office/Remote work arrangement and a competitive compensation package.

Formación

  • Experience with cloud services and large-scale distributed systems.
  • Strong background in DevOps practices and site reliability engineering.
  • Familiarity with monitoring, alerting, and incident response processes.

Responsabilidades

  • Develop and improve the entire lifecycle of services.
  • Enhance monitoring to reduce outages and duration.
  • Automate operations to increase reliability and velocity.
  • Lead designs of major components to improve availability and efficiency.
  • Conduct post-incident reviews and drive continuous improvement.

Conocimientos

Azure
Kubernetes
Terraform
Chef
CI/CD
Unix/Linux
Monitoring/Observability

Educación

Bachelor's degree in Computer Science or similar

Herramientas

Kubernetes
Terraform
Chef

Descripción del empleo

Senior Site Reliability Engineer (#5784)

LATAM

Work type:

Office/Remote

Technical Level:

Senior

Job Category:

Software Development

Project:

Top tech for managing staff and stock in retail & hospitality

We are looking for a Senior Site Reliability Engineer who is interested in an opportunity to work for an innovative hospitality company with cutting-edge technologies, with new development activities and challenges ahead. Our international team members share a common desire to develop brilliant products on reliable and resilient systems, along with their own skills. We run our services in Azure and traditional data centers. Take a chance to make a valuable contribution and enhance your professional skills.

About the job:

Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that cloud services—both our internally critical and our externally-visible systems—have reliability, uptime appropriate to customer needs, and a fast rate of improvement. Additionally, SREs will keep an ever-watchful eye on our systems' capacity and performance.

On the SRE team, you’ll have the opportunity to manage the complex challenges of scale that are unique to the project while using your expertise in coding, algorithms, complexity analysis, and large-scale system design. You will provide scalable, reliable, durable, and secure services using a customer-first approach while innovating technically. You will understand our customer needs and how we can meet them.

Responsibilities:
  • Develop and improve the whole lifecycle of services
  • Establish and improve monitoring capabilities to reduce outage frequency and duration
  • Create sustainable systems through automation and uplifts
  • Develop and scale systems sustainably through mechanisms such as automation, and evolve systems by pushing for changes that improve reliability and velocity.
  • Lead designs of major software components, systems, and features to improve the availability, scalability, latency, and efficiency of our services
  • Analyze and support services before they go live via system design consulting, developing software platforms and frameworks, capacity planning
  • Conduct post-incident analysis and reviews with an attitude of continuous improvement
Requirements:
  • Ideally, strong experience in Azure Services and capabilities, but other cloud services (AWS, Google Cloud Platform etc.) will be considered
  • Confidence and strong experience with KubernetesRecent and fluent Terraform and (Chef platform experience nice to have)
  • Extensive expertise in software development/testing, development operations, and site reliability engineering
  • Experience of Unix/Linux administration - an appreciation of systems internals (e.g., filesystems, system calls) is a bonus
  • Experience with Continuous Integration and Deployment (CI/CD) and release orchestration and Configuration Management of VMs
  • Cloud-agnostic approach, with flexibility to work across various cloud platforms
Nice to have:
  • Bachelor's degree in Computer Science, similar technical field of study, or equivalent practical experience
  • Experience designing, analyzing, and troubleshooting large-scale distributed systems
  • Systematic problem-solving approach, combined with excellent communication skills and a sense of ownership and drive
  • Experience in configuring application monitoring with Azure Monitor and Application Insight
  • Experience with Service Mesh
  • Previous experience as a DevOps engineer is preferred
We offer*:
  • Flexible working format - remote, office-based or flexible
  • A competitive salary and good compensation package
  • Professional development tools (mentorship program, tech talks and trainings, centers of excellence, and more)
  • Active tech communities with regular knowledge sharing

Project: Leading platform for electronic agreements

Consigue la evaluación confidencial y gratuita de tu currículum.

o arrastra y suelta tu archivo aquí

Similar jobs

Puestos de trabajo similares que vale la pena comparar

Site Reliability Engineer ID53670
Site Reliability Engineer ID53670

AgileEngine • Rosarito

Híbrido
MXN 870.000 - 1.306.000
Professional growth
Competitive compensation
Exciting projects
+1
Site Reliability Engineer
Site Reliability Engineer

CTC • Estado de México

A distancia
MXN 1.433.000 - 1.793.000
Senior AWS Site Reliability Engineers - 2850
Senior AWS Site Reliability Engineers - 2850

Xideral • Región Centro

Híbrido
MXN 1.200.000 - 1.500.000
Premium Benefits
Performance bonuses
SGMM Medical insurance
SRE (Engineering & Administration Background)
SRE (Engineering & Administration Background)

Fulcrum Digital • Ciudad de México

Híbrido
MXN 900.000 - 1.500.000
Site Reliability Engineer ID60188
Site Reliability Engineer ID60188

AgileEngine • Ciudad de México

Híbrido
MXN 1.049.000 - 1.400.000
Professional growth: Mentorship, TechTalks, and personalized growth roadmaps.
Competitive compensation: USD-based pay with education, fitness, and team activity budgets.
Exciting projects: Modern solutions with Fortune 500 and top product companies.
+1
Site Reliability Engineer ID45689
Site Reliability Engineer ID45689

AgileEngine • Rosarito

A distancia
MXN 1.531.000 - 2.212.000
Mentorship and TechTalks
Competitive USD-based compensation
Work on modern solutions
+1
Site Reliability Engineer
Site Reliability Engineer

Tata Consultancy Services • Ciudad de México

Presencial
Senior Site Reliability Engineer Lead
Senior Site Reliability Engineer Lead

HCL Technologies Limited • Región Centro

Presencial
MXN 900.000 - 1.200.000
Site Reliability Engineering (SRE) Lead - 2770
Site Reliability Engineering (SRE) Lead - 2770

Xideral • Región Centro

A distancia
MXN 700.000 - 900.000
Attractive Salary
Performance bonuses
SGMM Medical insurance
Site Reliability Engineer
Site Reliability Engineer

KI people • México

Híbrido
MXN 745.000 - 1.118.000
Payroll
Direct hire by client
Multicultural teams
+1