DevOps and SRE

United States Digital Space LLC

Singapore

On-site

SGD 120,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

United States Digital Space LLC is seeking an experienced hands-on DevOps and SRE lead to establish and operate a modern DevOps practice for a B2B2C/D2C Digital Platform programme at a customer site. The role collaborates with Enterprise Architecture, Product Owner, and external partners to ensure alignment with enterprise standards and delivery goals.

It is a 1-year contract. You will drive CI/CD improvements, observability, SLOs, incident readiness, and production readiness while guiding

Qualifications

  • Minimum 10 years of software engineering, platform engineering, or technical delivery roles including at least 5 years in cloud-based solution delivery
  • Proven experience leading DevOps and SRE practices, including CI/CD, monitoring and logging, incident management, RCA, SLOs, and production readiness
  • Strong hands-on experience in solution architecture for digital applications, APIs, and integrations in cloud environments.
  • Experience delivering digital platforms involving customer-facing journeys and partner ecosystems is desirable, especially insurance digital distribution or policy servicing platforms
  • Practical experience with Microsoft Azure and/or Amazon Web Services (AWS), including services such as API Management (APIM), serverless functions, messaging services (e.g. Service Bus), Key Vault, Application Gateway and Web Application Firewall (WAF), storage, and observability tooling
  • Experience designing and implementing resilience patterns for distributed systems and external integrations
  • Ability to guide engineering teams and vendors on implementation quality, runtime reliability, and operational discipline
  • Practical experience in cloud cost optimisation in collaboration with delivery and finance stakeholders
  • Exposure to Infrastructure as Code, container platforms, and cloud-native delivery models is desirable
  • Strong communication and stakeholder management skills across programme, technology, architecture, infrastructure, and security teams
  • Experience working in regulated environments with multiple governance stakeholders is desirable
  • Relevant cloud architecture or platform certifications are preferred

Responsibilities

  • Lead hands-on DevOps enablement for the Digital Platform Programme, including CI/CD pipeline improvements, release automation, environment promotion controls, and deployment quality practices
  • Establish and operate SRE practices, including observability, service monitoring, alerting, SLOs, incident readiness, and RCA discipline
  • Improve engineering outcomes by increasing deployment frequency, reducing lead time for changes, and strengthening release safety
  • Strengthen production readiness through structured practices such as runbooks, dashboards, logging standards, alert tuning, and operational handover
  • Reduce high-severity incidents (Severity 1 and Severity 2) and improve MTTR
  • Provide day-to-day solution architecture support, guiding design and implementation decisions in collaboration with the Enterprise Architect
  • Ensure solutions align with enterprise architecture principles, approved standards, and non-functional requirements
  • Apply practical design patterns to enhance resilience, scalability, recoverability, and operability (e.g., retries, queuing, idempotency, circuit breakers, failover handling)
  • Escalate material architecture changes and design deviations through appropriate governance channels for review and approval
  • Work closely with engineering teams and vendors to translate approved designs into deliverable, maintainable, and supportable solutions
  • Support the implementation of API and integration patterns with a focus on reliability, performance, observability, and supportability
  • Work with internal and external teams to improve integration consistency, service behaviour, and partner onboarding readiness
  • Ensure event-driven and messaging-based patterns are implemented in a scalable and recoverable manner
  • Work with cybersecurity and architecture teams to implement approved security and identity requirements
  • Support secure-by-design practices, including secrets management, identity integration, application protection, and logging standards
  • Drive improved engineering discipline to reduce recurring code and security issues and ensure timely remediation
  • Collaborate with the relevant security governance function on cybersecurity governance, risk oversight, and security approvals
  • Partner with the organisation’s cloud engineer to enhance cloud cost visibility and optimisation through tagging discipline, right-sizing, utilisation reviews, and engineering usage improvements
  • Promote strong delivery governance and engineering discipline, including code quality, operational readiness, and consistency of implementation.
  • Ensure delivery partners and vendors adhere to agreed technical, operational, and security standards
  • Provide hands-on technical leadership, especially for backend and platform-related delivery.
  • Coach engineering and delivery teams on DevOps, SRE, resilience, observability, and production support practices
  • Act as a technical focal point during delivery, resolving day-to-day issues and aligning teams on execution standards.

Skills

DevOps
SRE
CI/CD
Observability
Back-end engineering
Cloud platforms
Stakeholder management
Leadership

Tools

Azure
AWS
APIM
Key Vault
Service Bus
WAF

Job description

Roles & Responsibilities

This is a hands-on role responsible for establishing and operating a modern DevOps practice, and performing Site Reliability Engineering (SRE) to support the successful delivery and operation of a B2B2C/D2C Digital Platform Programme at Customer site.

In addition to DevOps and SRE responsibilities, the role provides day-to-day solution architecture support to ensure delivery aligns with enterprise standards. The incumbent will work closely with the Enterprise Architect to ensure alignment to Group architecture practices and standards, support the Product Owner on technical matters, and coordinate with internal technology team leads and external delivery partners. The role reports to the Technology Solutions department and operates as an embedded specialist within a cross-functional delivery team (“pod”) supporting the Digital Platform Programme.

Please note that this is a 1-year contract role.

Key Responsibilities
DevOps and SRE
  • Lead hands-on DevOps enablement for the Digital Platform Programme, including Continuous Integration and Continuous Delivery (CI/CD) pipeline improvements, release automation, environment promotion controls, and deployment quality practices.
  • Establish and operate SRE practices, including observability, service monitoring, alerting, Service Level Objectives (SLOs), incident readiness, and Root Cause Analysis (RCA) discipline.
  • Improve engineering outcomes by increasing deployment frequency, reducing lead time for changes, and strengthening release safety
  • Strengthen production readiness through structured practices such as runbooks, dashboards, logging standards, alert tuning, and operational handover.
  • Reduce high-severity incidents (Severity 1 and Severity 2) and improve Mean Time to Recovery (MTTR) through enhanced system resilience and runtime support.
Solution Architecture and Technical Delivery
  • Provide day-to-day solution architecture support, guiding design and implementation decisions in collaboration with the Enterprise Architect
  • Ensure solutions align with enterprise architecture principles, approved standards, and non-functional requirements.
  • Apply practical design patterns to enhance resilience, scalability, recoverability, and operability (e.g., retries, queuing, idempotency, circuit breakers, failover handling).
  • Escalate material architecture changes and design deviations through appropriate governance channels for review and approval.
  • Work closely with engineering teams and vendors to translate approved designs into deliverable, maintainable, and supportable solutions.
API, Integration and Runtime Enablement
  • Support the implementation of API and integration patterns with a focus on reliability, performance, observability, and supportability
  • Work with internal and external teams to improve integration consistency, service behaviour, and partner onboarding readiness
  • Ensure event-driven and messaging-based patterns are implemented in a scalable and recoverable manner
Security and Identity Implementation
  • Work with cybersecurity and architecture teams to implement approved security and identity requirements
  • Support secure-by-design practices, including secrets management, identity integration, application protection, and logging standards
  • Drive improved engineering discipline to reduce recurring code and security issues and ensure timely remediation
  • Collaborate with the relevant security governance function on cybersecurity governance, risk oversight, and security approvals
Cloud Cost, Quality and Governance
  • Partner with the organisation’s cloud engineer to enhance cloud cost visibility and optimisation through tagging discipline, right-sizing, utilisation reviews, and engineering usage improvements
  • Promote strong delivery governance and engineering discipline, including code quality, operational readiness, and consistency of implementation.
  • Ensure delivery partners and vendors adhere to agreed technical, operational, and security standards
Leadership and Ways of Working
  • Provide hands-on technical leadership, especially for backend and platform-related delivery.
  • Coach engineering and delivery teams on DevOps, SRE, resilience, observability, and production support practices
  • Act as a technical focal point during delivery, resolving day-to-day issues and aligning teams on execution standards.
Requirements
  • Minimum 10 years of experience in software engineering, platform engineering, or technical delivery roles, including at least 5 years in cloud-based solution delivery
  • Proven experience leading DevOps and SRE practices, including CI/CD, monitoring and logging, incident management, RCA, SLOs, and production readiness
  • Strong hands-on experience in solution architecture for digital applications, APIs, and integrations in cloud environments.
  • Experience delivering digital platforms involving customer-facing journeys and partner ecosystems is desirable, especially insurance digital distribution or policy servicing platforms
  • Practical experience with Microsoft Azure and/or Amazon Web Services (AWS), including services such as API Management (APIM), serverless functions, messaging services (e.g. Service Bus), Key Vault, Application Gateway and Web Application Firewall (WAF), storage, and observability tooling
  • Experience designing and implementing resilience patterns for distributed systems and external integrations
  • Ability to guide engineering teams and vendors on implementation quality, runtime reliability, and operational discipline
  • Practical experience in cloud cost optimisation in collaboration with delivery and finance stakeholders
  • Exposure to Infrastructure as Code, container platforms, and cloud-native delivery models is desirable
  • Strong communication and stakeholder management skills across programme, technology, architecture, infrastructure, and security teams
  • Experience working in regulated environments with multiple governance stakeholders is desirable
  • Relevant cloud architecture or platform certifications are preferred
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Director, APAC Head of Platform Services
Director, APAC Head of Platform Services

mufg bank, ltd. singapore branch • Singapore

On-site
SGD 250,000 - 380,000
Site Reliability Engineer
Site Reliability Engineer

IDEMIA Public Security • Singapore

On-site
SGD 120,000 - 180,000
Operations Lead - SRE, Incident Management
Operations Lead - SRE, Incident Management

Sciente Consulting • Singapore

On-site
SGD 120,000 - 160,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Autodesk • Singapore

On-site
SGD 80,000 - 120,000
Senior Program Engineer
Senior Program Engineer

Neerinfo Solutions • Singapore

On-site
SGD 90,000 - 140,000
Senior DevOps & SRE Architect for Cloud Platform
Senior DevOps & SRE Architect for Cloud Platform

United States Digital Space LLC • Singapore

On-site
SGD 120,000 - 180,000
Site Reliability Engineer
Site Reliability Engineer

SGX Group • Singapore

On-site
SGD 180,000 - 300,000
DevOps SRE
DevOps SRE

Codigo - The Mobile App Company • Singapore

On-site
SGD 70,000 - 100,000
Software Engineer/ Site Reliability Engineer
Software Engineer/ Site Reliability Engineer

United States Digital Space LLC • Singapore

On-site
SGD 90,000 - 150,000
Program Engineer
Program Engineer

NeerInfo Solutions • Singapore

On-site
SGD 120,000 - 180,000