Senior Site Reliability Engineer (#5784)

N-iX

Philippines

Hybrid

PHP 1,500,000 - 2,100,000

Full time

31 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Flexible work options

Job summary

N-iX is seeking a Senior Site Reliability Engineer to join our international team in a hybrid setup (Office/Remote). You will work on a leading platform for managing services in Azure and data centers, focusing on reliability, scalability, and performance for large distributed systems.

You will design and implement monitoring, automation, and capacity planning, collaborating across teams to deliver resilient cloud-based services with a customer-first approach.

Qualifications

  • Experience with cloud services, preferably Azure, and other clouds.
  • Strong Kubernetes experience and containerized workloads.
  • Extensive expertise in DevOps, SRE practices, and automation.
  • Experience with Unix/Linux administration.
  • Familiarity with CI/CD and release orchestration.
  • Cloud-agnostic mindset with cross-cloud flexibility.
  • Good communication and ownership skills.

Responsibilities

  • Develop and improve the full lifecycle of services.
  • Establish and improve monitoring capabilities to reduce outages.
  • Create sustainable systems through automation and uplift.
  • Develop and scale systems sustainably through automation and improvement.
  • Lead designs of major software components for availability and latency.
  • Analyze and support services before go-live via design reviews.

Skills

Azure expertise
Kubernetes experience
CI/CD
Unix/Linux administration
System design & reliability
Communication & ownership

Education

Bachelor's degree in Computer Science or equivalent

Tools

Azure
Kubernetes
Terraform
Chef
CI/CD tools
Azure Monitor/Application Insights

Job description

Senior Site Reliability Engineer (#5784)

LATAM

Work type:

Office/Remote

Technical Level:

Senior

Job Category:

Software Development

Project:

Top tech for managing staff and stock in retail & hospitality

We are looking for a Senior Site Reliability Engineer who is interested in an opportunity to work for an innovative hospitality company with cutting-edge technologies, with new development activities and challenges ahead. Our international team members share a common desire to develop brilliant products on reliable and resilient systems, along with their own skills. We run our services in Azure and traditional data centers. Take a chance to make a valuable contribution and enhance your professional skills.

About the job:

Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that cloud services—both our internally critical and our externally-visible systems—have reliability, uptime appropriate to customer needs, and a fast rate of improvement. Additionally, SREs will keep an ever-watchful eye on our systems' capacity and performance.

On the SRE team, you’ll have the opportunity to manage the complex challenges of scale that are unique to the project while using your expertise in coding, algorithms, complexity analysis, and large-scale system design. You will provide scalable, reliable, durable, and secure services using a customer-first approach while innovating technically. You will understand our customer needs and how we can meet them.

Responsibilities:

  • Develop and improve the whole lifecycle of services
  • Establish and improve monitoring capabilities to reduce outage frequency and duration
  • Create sustainable systems through automation and uplifts
  • Develop and scale systems sustainably through mechanisms such as automation, and evolve systems by pushing for changes that improve reliability and velocity.
  • Lead designs of major software components, systems, and features to improve the availability, scalability, latency, and efficiency of our services
  • Analyze and support services before they go live via system design consulting, developing software platforms and frameworks, capacity planning
  • Conduct post-incident analysis and reviews with an attitude of continuous improvement

Requirements:

  • Ideally, strong experience in Azure Services and capabilities, but other cloud services (AWS, Google Cloud Platform etc.) will be considered
  • Confidence and strong experience with KubernetesRecent and fluent Terraform and (Chef platform experience nice to have)
  • Extensive expertise in software development/testing, development operations, and site reliability engineering
  • Experience of Unix/Linux administration - an appreciation of systems internals (e.g., filesystems, system calls) is a bonus
  • Experience with Continuous Integration and Deployment (CI/CD) and release orchestration and Configuration Management of VMs
  • Cloud-agnostic approach, with flexibility to work across various cloud platforms

Nice to have:

  • Bachelor's degree in Computer Science, similar technical field of study, or equivalent practical experience
  • Experience designing, analyzing, and troubleshooting large-scale distributed systems
  • Systematic problem-solving approach, combined with excellent communication skills and a sense of ownership and drive
  • Experience in configuring application monitoring with Azure Monitor and Application Insight
  • Experience with Service Mesh
  • Previous experience as a DevOps engineer is preferred

We offer*:

  • Flexible working format - remote, office-based or flexible
  • A competitive salary and good compensation package
  • Professional development tools (mentorship program, tech talks and trainings, centers of excellence, and more)
  • Active tech communities with regular knowledge sharing

Project: Leading platform for electronic agreements

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Technical Lead - Site Reliability Engineering
Technical Lead - Site Reliability Engineering

LSEG • Taguig

On-site
PHP 4,914,000 - 7,372,000
Healthcare
Retirement planning
Paid volunteering days
+1
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Acquire Intelligence • Taguig

On-site
PHP 900,000 - 1,500,000
Senior Engineer - Site Reliability Engineering
Senior Engineer - Site Reliability Engineering

LSEG • Philippines

On-site
PHP 1,000,000 - 1,600,000
Staff SRE Engineer
Staff SRE Engineer

Stellar Cyber • España

On-site
PHP 5,528,000 - 7,372,000
Site Reliability Engineer
Site Reliability Engineer

IDEMIA • Philippines

On-site
PHP 900,000 - 1,500,000
Staff Site Reliability Engineer – Cloud Efficiency
Staff Site Reliability Engineer – Cloud Efficiency

Super • España

On-site
PHP 1,200,000 - 1,600,000
Medical / Health Insurance
Employee Assistance Programme
Technical Lead - Site Reliability Engineering
Technical Lead - Site Reliability Engineering

LSEG • Philippines

On-site
PHP 2,000,000 - 4,000,000
Site Reliability Engineer
Site Reliability Engineer

IDEMIA PHILIPPINES INC. • Philippines

On-site
PHP 900,000 - 1,350,000
Site Reliability Engineer
Site Reliability Engineer

AgileEngine • Mexico

Hybrid
PHP 8,734,000 - 13,100,000
Professional growth
Competitive USD-based pay
Exciting projects
+1
Associate Site Reliability Engineer
Associate Site Reliability Engineer

Railway Corp • Mexico

On-site
PHP 5,474,000 - 7,908,000