infrastructure engineer (weekends, shift) #IKR

RECRUIT EXPRESS PTE LTD

Singapore

On-site

SGD 120,000 - 180,000

Full time

9 days ago
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

RECRUIT EXPRESS PTE LTD is seeking an Infrastructure Engineer to build, operate, and improve reliable, scalable production systems. The role blends infrastructure engineering with SRE practices, focusing on availability, automation, incident management, and continuous improvement.

Key responsibilities include managing high-severity incidents, driving SRE/chaos engineering for strategic systems, and delivering data-driven reliability improvements.

Qualifications

  • Bachelor's degree in computer science or related field.
  • 5+ years of relevant experience in infrastructure/ops or SRE.
  • Experience driving major production incidents and retrospectives.
  • Proficiency with Core Java 8, Cloud Foundry, NoSQL, Linux/Unix, and scripting.

Responsibilities

  • Manage high severity incidents and rapid recovery.
  • Champion production resilience and availability with ops and dev teams.
  • Drive SRE and chaos engineering initiatives for key systems.
  • Communicate production reliability and performance to business stakeholders.
  • Lead continuous improvement of processes using reliability methods.
  • Analyze incidents to reduce impact and implement preventative solutions.
  • Improve system reliability with data-driven design and metrics.
  • Provide expert guidance on technology choices for reliability.

Skills

Incident management
Automation
SRE practices
Production reliability

Education

Bachelor's degree in Computer Science or related field

Tools

Core Java 8
Cloud Foundry
NoSQL databases
Linux
Unix
Python
Unix scripting

Job description

Note: Successful candidates need to be comfortable with working on weekends and shifts as and when required

We are looking for Infrastructure Engineer responsible for building, operating, and improving reliable, scalable, and high-performing production systems. This role combines infrastructure engineering expertise with SRE practices, focusing on system availability, automation, incident management, and continuous improvement.

Key Responsibilities
  • Manage high severity incidents and high customer impact incidents focusing on fast recovery
  • Champions production resilience and availability, focusing on superior client experience, by working with the operation team and technology development teams
  • Drive the implementation of Site Reliability Engineer (SRE) and Chaos Engineering design for all strategic systems
  • Drive effective communication between business and technology with regards to production service reliability and performance
  • Drive continuous improvements in processes or systems leveraging Site Reliability Engineering methods
  • Respond to, evaluate and analyze production incidents to minimise their impact as well as devise innovative solutions to prevent them in the future
  • Improve the reliability and availability of systems by gathering hard data, designing systems for increased service reliability and performance
  • Provide expert advice and training to our engineers as to which technology solutions and advanced reliability techniques to use on each situation
  • Any other ad-hoc duties as assigned by supervisor
Requirements
  • Bachelor's degree in computer science or related field
  • 5+ years of relevant experience
  • Experience driving major production incidents and organize incident retrospective meetings
  • Experience with Core Java 8, Cloud Foundry and non-relational databases, and Linux, Unix systems
  • Experience with high availability, high-scale, and performant systems
  • Experience with python and Unix scripting
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Infrastructure Engineer — SRE & Reliability
Senior Infrastructure Engineer — SRE & Reliability

RECRUIT EXPRESS PTE LTD • Singapore

On-site
SGD 120,000 - 180,000
Site Reliability Engineer( SRE)
Site Reliability Engineer( SRE)

XIAOMI TECHNOLOGIES SINGAPORE PTE. LTD. • Singapore

On-site
SGD 90,000 - 130,000
Site Reliability Engineer
Site Reliability Engineer

SINGAPORE EXCHANGE LIMITED • Singapore

On-site
SGD 120,000 - 160,000
SL2564 - SRE & Service Delivery Lead
SL2564 - SRE & Service Delivery Lead

FPT Asia Pacific Pte Ltd • Singapore

On-site
SGD 120,000 - 180,000
SL2564 - SRE & Service Delivery Lead
SL2564 - SRE & Service Delivery Lead

FPT Asia Pacific • Singapore

On-site
SGD 90,000 - 130,000
Site Reliability Engineer
Site Reliability Engineer

TEKsystems • Singapore

Hybrid
SGD 120,000 - 180,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

kidentify pte. ltd. • Singapore

On-site
SGD 80,000 - 120,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Kidentify • Singapore

On-site
SGD 120,000 - 180,000
Site Reliability Engineer
Site Reliability Engineer

TP-LINK CORPORATION PTE. LTD. • Singapore

On-site
SGD 90,000 - 150,000
L2 SRE
L2 SRE

NTT Data Singapore • Singapore

On-site
SGD 120,000 - 180,000