Site Reliability Engineer

IDEMIA Public Security

Singapore

On-site

SGD 120,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

IDEMIA Public Security is seeking an experienced DevOps/SRE to ensure reliability, scalability and performance of our systems. You will collaborate with development and operations to build and maintain robust infrastructure, observability, and rapid deployment pipelines.

The role involves monitoring platforms, incident response, on-call rotations, CI/CD expansion, and working with Docker, Kubernetes and cloud services to deliver resilient services in Singapore.

Qualifications

  • 2–5 years of experience in software development, DevOps or SRE.
  • Strong communication and ability to work in a fast-paced environment.
  • Degree in Electrical / Electronics / Computer Engineering / Computer Science or a relevant discipline.
  • Basic understanding of Linux/Unix systems and shell scripting.

Responsibilities

  • Maintains platforms after go-live by monitoring availability and performance.
  • Assists in incident response and root-cause analysis.
  • Develops tools and scripts to improve operational efficiency.
  • Maintains and enhances CI/CD pipelines.
  • Collaborates with engineers to design scalable, resilient systems.
  • Participates in on-call rotations and reduces alert fatigue.
  • Documents processes, configurations, and best practices.

Skills

Strong communication
Problem solving
Team player
Curious

Education

Bachelor's degree in Computer Engineering / Computer Science

Tools

Docker
Kubernetes
Prometheus
Grafana
ELK
Jenkins
GitLab
Bitbucket
Jira
Python
Bash
Java
Linux

Job description

This role plays a critical part in ensuring reliability, scalability, and performance of our systems and services. You will work closely with development and operations teams to build and maintain robust infrastructure and tools that support high availability, monitoring and rapid deployment.

Job description:

Purpose

This role plays a critical part in ensuring reliability, scalability, and performance of our systems and services. You will work closely with development and operations teams to build and maintain robust infrastructure and tools that support high availability, monitoring and rapid deployment.

Key Missions
  • Maintains platforms or products after go live by measuring and monitoring their availability, performance and overall system health
  • Recovers platforms or products during production incidents to meet targeted service-level agreements
  • Set up, enhance and maintain observability tools.
  • Assist in incident response, perform root cause analysis, and postmortem documentation.
  • Develop tools/applications/scripts to improve operational efficiency.
  • Maintain and enhance CI/CD pipelines.
  • Collaborate with software engineers to design scalable and resilient systems.
  • Participate in on-call and on-site rotations and contribute to reducing alert fatigue.
  • Document processes, configurations, and best practices.
  • Support other software efficiency improvement initiatives.
Profile & Other Information
  • At least 2-5 years’ experience in software development, Devops or SRE.
  • Curious, Strong communicator and ready to work in a fast-paced environment and willing to pick up new skills and technologies as necessary.
  • Degree in Electrical / Electronics / Computer Engineering / Computer Science or a relevant discipline
  • Basic understanding of Linux/Unix systems and shell scripting.
  • Familiarity with cloud platforms (e.g., AWS, Azure, GCP).
  • Exposure to containerization tools (e.g., Docker, Kubernetes).
  • Experience with monitoring tools (e.g., Prometheus, Grafana, ELK).
  • Knowledge of CI/CD tools (e.g., Jenkins, Gitlab, Bitbucket, Jira).
  • Programming/scripting skills in Python, Java, or Bash.
  • Understanding of networking fundamentals and system security.
  • Good written and verbal communication skills.
  • Self-motivated, independent and a good team player
  • Able to work under pressure in a fast-paced environment
  • Innovative, proactive mindset and with a focus on continuous improvement
  • Strong analytical and problem-solving skills
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer (SRE)
Senior Site Reliability Engineer (SRE)

VANGUARD SOFTWARE PTE. LTD. • Singapore

On-site
SGD 100,000 - 150,000
Technical Leadership
Career Growth
High-Performance Team
+1
Site Reliability Engineer(Senior SRE)
Site Reliability Engineer(Senior SRE)

XIAOMI TECHNOLOGIES SINGAPORE PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Software Engineer/ Site Reliability Engineer
Software Engineer/ Site Reliability Engineer

United States Digital Space LLC • Singapore

On-site
SGD 90,000 - 150,000
Site Reliability Engineer
Site Reliability Engineer

Jobtailor • Singapore

On-site
SGD 120,000 - 180,000
Site Reliability Engineer, Enterprise Technology Services
Site Reliability Engineer, Enterprise Technology Services

United States Digital Space LLC • Singapore

On-site
SGD 120,000 - 200,000
Site Reliability Engineer - Data Availability
Site Reliability Engineer - Data Availability

SIX • Singapore

Hybrid
SGD 100,000 - 180,000
Site Reliability Engineer - HM: Mukesh
Site Reliability Engineer - HM: Mukesh

NTT Data Singapore • Singapore

On-site
SGD 80,000 - 120,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

PURVIEW ASIA PACIFIC PTE. LTD. • Singapore

On-site
SGD 90,000 - 130,000
Site Reliability Engineer
Site Reliability Engineer

SGX Group • Singapore

On-site
SGD 180,000 - 300,000
Site Reliability Engineer
Site Reliability Engineer

Singapore Exchange Limited • Singapore

On-site
SGD 180,000 - 250,000