Senior Site Reliability Engineer

N-iX

Colombia

Híbrido

COP 120.000.000 - 180.000.000

Jornada completa

hace 15 horas
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Transforma esta oferta en una entrevista: un currículum y una carta de presentación creados pensando en lo que quiere el empleador.

Supera los filtros ATS

Ventajas ofrecidas por este puesto de trabajo

Flexible work format
Education reimbursement
Professional development

Descripción de la vacante

N-iX is seeking a Senior Site Reliability Engineer to join an innovative hospitality company with Azure-based services and on-premises data centers. You will enhance reliability, scalability, and performance across distributed systems, and lead critical design choices for uptime and velocity.

The role emphasizes cloud and on-premises hybrid environments, automation, and continuous improvement in a global team. Remote-friendly with flexible format and growth opportunities.

Formación

  • Experience with cloud platforms (Azure preferred).
  • Strong knowledge of Kubernetes and container orchestration.
  • Hands-on experience with CI/CD pipelines and release orchestration.
  • Familiarity with configuration management tools.

Responsabilidades

  • Develop and improve the lifecycle of services.
  • Improve monitoring to reduce outages.
  • Automate operations to improve reliability.
  • Scale systems and improve performance and efficiency.
  • Lead designs for high availability and latency improvements.
  • Support pre-live validation and capacity planning.
  • Conduct post-incident reviews and continuous improvement.

Conocimientos

Azure
Kubernetes
CI/CD
Automation
Unix/Linux

Educación

Bachelor's degree in CS or related

Herramientas

Terraform
Chef
Ansible

Descripción del empleo

We are looking for a Senior Site Reliability Engineer who is interested in an opportunity to work for an innovative hospitality company with cutting-edge technologies, with new development activities and challenges ahead. Our international team members share a common desire to develop brilliant products on reliable and resilient systems, along with their own skills. We run our services in Azure and traditional data centers. Take a chance to make a valuable contribution and enhance your professional skills.

About the job:

Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that cloud services—both our internally critical and our externally-visible systems—have reliability, uptime appropriate to customer needs, and a fast rate of improvement. Additionally, SREs will keep an ever-watchful eye on our systems' capacity and performance.

Responsibilities:
  • Develop and improve the whole lifecycle of services
  • Establish and improve monitoring capabilities to reduce outage frequency and duration
  • Create sustainable systems through automation and uplifts
  • Develop and scale systems sustainably through mechanisms such as automation, and evolve systems by pushing for changes that improve reliability and velocity.
  • Lead designs of major software components, systems, and features to improve the availability, scalability, latency, and efficiency of our services
  • Analyze and support services before they go live via system design consulting, developing software platforms and frameworks, capacity planning
  • Conduct post-incident analysis and reviews with an attitude of continuous improvement
Requirements:
  • Ideally, strong experience in Azure Services and capabilities, but other cloud services (AWS, Google Cloud Platform etc.) will be considered
  • Confidence and strong experience with KubernetesRecent and fluent Terraform and (Chef platform experience nice to have)
  • Extensive expertise in software development/testing, development operations, and site reliability engineering
  • Experience of Unix/Linux administration - an appreciation of systems internals (e.g., filesystems, system calls) is a bonus
  • Experience with Continuous Integration and Deployment (CI/CD) and release orchestration and Configuration Management of VMs
  • Cloud-agnostic approach, with flexibility to work across various cloud platforms
  • Experience programming in one or more of the following languages: C#,, C++, Java, Python, JavaScript, Go, Perl, or Ruby
Nice to have:
  • Bachelor's degree in Computer Science, similar technical field of study, or equivalent practical experience
  • Experience in distributed systems, storage systems, or databases
  • Experience designing, analyzing, and troubleshooting large-scale distributed systems
  • Systematic problem-solving approach, combined with excellent communication skills and a sense of ownership and drive
  • Experience in configuring application monitoring with Azure Monitor and Application Insight
  • Experience with Service Mesh
  • Previous experience as a DevOps engineer is preferred
We offer*:
  • Flexible working format - remote, office-based or flexible
  • A competitive salary and good compensation package
  • Personalized career growth
  • Professional development tools (mentorship program, tech talks and trainings, centers of excellence, and more)
  • Active tech communities with regular knowledge sharing
  • Education reimbursement
  • Memorable anniversary presents
  • Corporate events and team buildings
  • Other location-specific benefits
  • not applicable for freelancers
Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Site Reliability Engineer
Site Reliability Engineer

Intraway • Bogotá

Presencial
COP 96.000.000 - 140.000.000
Unlimited PTO
Training access
English classes
+2
Site Reliability Engineer ID53670
Site Reliability Engineer ID53670

AgileEngine • Metropolitana

Híbrido
COP 142.369.000 - 213.554.000
Professional growth: Mentorship, TechTalks, and personalized growth roadmaps.
Competitive compensation: USD-based pay with education, fitness, and team activity budgets.
Exciting projects: Modern solutions with Fortune 500 and top product companies.
+1
Senior Site Reliability Engineer (Azure)
Senior Site Reliability Engineer (Azure)

OpsGenius • Colombia

Presencial
COP 380.409.000 - 475.511.000
Paid time off
Health insurance
Learning credits
+3
Senior Site Reliability Engineer
Senior Site Reliability Engineer

LanceSoft, Inc. • Colombia

Presencial
COP 90.000.000 - 150.000.000
Site Reliability Engineer
Site Reliability Engineer

Source Meridian • Colombia

Presencial
COP 244.021.000 - 406.703.000
Workout bonus
Home office setup bonus
Learning platform bonus
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

MPS Group LLC • Bogotá

Presencial
COP 120.000.000 - 180.000.000
Site Reliability Engineering (SRE)
Site Reliability Engineering (SRE)

adidas • Colombia

Presencial
COP 152.495.848 - 220.271.781
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Michael Page Colombia • San Gil

Presencial
COP 228.815.000 - 343.224.000
Crecimiento profesional a través de desafíos técnicos
Trabajo con tecnologías cloud de vanguardia
Exposición a prácticas SRE modernas
Site Reliability Engineering (SRE)
Site Reliability Engineering (SRE)

FashionUnited Group • Bogotá ciudad

Híbrido
COP 144.000.000 - 216.000.000
Service Reliability Engineer
Service Reliability Engineer

1083 Amadeus IT Group Colombia, S.A.S. • Colombia

Presencial
COP 182.089.000 - 254.926.000
Competitive remuneration
Vacation and holiday paid time off
Health insurances
+3