Software Engineer/ Site Reliability Engineer

United States Digital Space LLC

Singapore

On-site

SGD 90,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

United States Digital Space LLC is seeking a Software Engineer + SRE hybrid to ensure reliability and performance of our cloud platform. You will own tools like GitHub, Jenkins, Jira, and Artifactory, triage issues, and automate resolution with a focus on developer productivity.

The role requires hands-on experience with CI/CD pipelines, observability, and distributed systems. Expect in‑office collaboration at least 3 days a week, with 24/7 on-call rotation as part of the team.

Qualifications

  • Bachelor's degree or 3+ years of relevant work experience.
  • 2 years experience with CI/CD tooling such as Jenkins, Github, Bitbucket, ArgoCD, Artifactory, Azure DevOps in a large-scale environment.
  • 2 or more years working in a Platform, SRE or Production Engineering group for high availability/critical platforms/applications.
  • Bachelor's degree in IT, CS or related field and/or 3+ Years Working Experience IT Operations and Delivery.

Responsibilities

  • You will be the primary point of contact for developers using tools like GitHub, Jenkins, Jira, or Artifactory.
  • Troubleshoot and resolve tool-related issues promptly to minimize developer downtime.
  • Maintain and optimize CI/CD pipelines and integrations for reliability and scalability.
  • Collaborate with development teams to improve workflows and automation.
  • Design, implement, and maintain systems for high availability, scalability, and performance.
  • Monitor and improve application reliability through proactive measures and incident response.
  • Develop and maintain observability solutions (metrics, logging, tracing).
  • Participate in on‑call rotations and drive root cause analysis for incidents.

Skills

CI/CD tooling
DevTools support
Linux / containers
Cloud platforms
Python/Go/Java scripting
Observability tooling
On-call support

Education

Bachelor's degree or 3+ years of experience

Tools

GitHub
Jenkins
Artifactory
ArgoCD
Azure DevOps
Jira
Bitbucket
Grafana
Prometheus
Docker
Kubernetes

Job description

Job Description

Site Reliability Engineering (SRE) is essential to the company’s Cloud platform strategy. In this role, you’ll ensure our development platform and tools let engineers focus on innovation instead of infrastructure. You’ll promote observability best practices and automate resolution of recurring issues, working closely with software engineering teams to support security, availability, and performance. Responsibilities include triaging issues, collaborating on infrastructure management, and setting up monitoring for full coverage. Hands‑on expertise is required, especially with major DevTools like GitHub, Jenkins, Jira, and Artifactory.

We seek a Software Engineer + SRE hybrid engineer. The ideal candidate deeply understands at least one major DevTool, quickly resolves tool-related issues in collaboration with developers, and applies systems thinking to maintain reliable applications and infrastructure while improving developer productivity.

Key Responsibilities
DevTools Support
  • You will be the primary point of contact for developers using tools like GitHub, Jenkins, Jira, or Artifactory.
  • Troubleshoot and resolve tool-related issues promptly to minimize developer downtime.
  • Maintain and optimize CI/CD pipelines and integrations for reliability and scalability.
  • Collaborate with development teams to improve workflows and automation.
Site Reliability Engineering
  • Design, implement, and maintain systems for high availability, scalability, and performance.
  • Monitor and improve application reliability through proactive measures and incident response.
  • Develop and maintain observability solutions (metrics, logging, tracing).
  • Participate in on‑call rotations and drive root cause analysis for incidents.
Collaboration & Continuous Improvement
  • Partner with engineering teams to identify reliability risks and implement best practices.
  • Document processes, troubleshooting guides, and reliability playbooks.
  • Advocate for automation and self‑service solutions to reduce operational overhead.

the company requires at least 3 days in office, expectations of these days will be confirmed by your Hiring Manager.

Qualifications
Basic Qualifications
  • Bachelor's degree, OR 3+ years of relevant work experience
Preferred Qualifications
  • Bachelor's degree, OR 3+ years of relevant work experience
  • 2 years experience with CI/CD tooling such as Jenkins, Github, Bitbucket, ArgoCD, Artifactory, Azure DevOps in a large-scale environment
  • 2 or more years working in a Platform, SRE or Production Engineering group for high availability/critical platforms/applications
  • Bachelor's degree in IT, CS or related field and/or 3+ Years Working Experience IT Operations and Delivery.
  • Beginner level programming and/or scripting in 2 or more of the following: Python, Java, Go, PowerShell, JavaScript, Terraform, Ansible, Helm, Chef, Cloud Formation.
  • Basic understanding of YAML, JSON, HTML, XML.
  • Hands on experience in Linux and /or Windows systems and good understanding of distributed computing environments.
  • 2 years experience with observability tooling such as Grafana, Prometheus, Splunk, Datadog, New Relic, DynaTrace, Sentry, etc. in a large-scale environment
  • 2 years experience supporting relational and non-relational databases (MySQL, MongoDB, PostgreSQL, etc.), including creating and running queries, managing performance and scaling
  • Experience: 3 years in SRE and/or DevTools support roles.
  • Proficiency in at least one DevTool (GitHub, Jenkins, ArgoCD, Jira, Artifactory,
  • Strong understanding of CI/CD principles and pipelines.
  • Solid knowledge of Linux systems, networking, and containerization (Docker/Kubernetes).
  • Hands‑on experience with cloud platforms.
  • Programming/Scripting: Proficiency in Python, Ansible, or similar languages.
  • Mindset: Strong problem‑solving skills, systems thinking, self‑starter, and a passion for reliability.
  • Experience managing container infrastructure and supporting development transformation to a container first model.
  • This role requires on call support as the team provides 24/7 operational support.
  • Experience managing a distributed container platform, including but not limited to deployment/release management, provisioning, capacity management, workload management

the company is an EEO Employer

Qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability or protected veteran status. the company will also consider for employment qualified applicants with criminal histories in a manner consistent with EEOC guidelines and applicable local law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer/ Site Reliability Engineer
Software Engineer/ Site Reliability Engineer

Visa • Singapore

On-site
SGD 120,000 - 180,000
Senior Site Reliability Engineer (SRE)
Senior Site Reliability Engineer (SRE)

VANGUARD SOFTWARE PTE. LTD. • Singapore

On-site
SGD 100,000 - 150,000
Technical Leadership
Career Growth
High-Performance Team
+1
Site Reliability Engineer
Site Reliability Engineer

IDEMIA Public Security • Singapore

On-site
SGD 120,000 - 180,000
Software Engineer - Engineering Enablement (SRE Focus)
Software Engineer - Engineering Enablement (SRE Focus)

United States Digital Space LLC • Singapore

On-site
SGD 120,000 - 180,000
Site Reliability Engineer, Enterprise Technology Services
Site Reliability Engineer, Enterprise Technology Services

United States Digital Space LLC • Singapore

On-site
SGD 120,000 - 200,000
Sr. SRE
Sr. SRE

United States Digital Space LLC • Singapore

On-site
SGD 120,000 - 180,000
On-site in Singapore (3 days/wk)
Site Reliability Engineer(Senior SRE)
Site Reliability Engineer(Senior SRE)

XIAOMI TECHNOLOGIES SINGAPORE PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Site Reliability Engineer, Enterprise Technology Services
Site Reliability Engineer, Enterprise Technology Services

United States Digital Space LLC • Singapore

On-site
SGD 100,000 - 130,000
Site Reliability Engineer - Data Availability
Site Reliability Engineer - Data Availability

SIX • Singapore

Hybrid
SGD 100,000 - 180,000
Site Reliability Engineer
Site Reliability Engineer

Singapore Exchange Limited • Singapore

On-site
SGD 180,000 - 250,000