Software Engineer/ SRE (Linux)

United States Digital Space LLC

Basingstoke

Hybrid

GBP 60,000 - 90,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

United States Digital Space LLC is seeking a Site Reliability Engineer with a strong focus on SRE and DevTools support to help harden our cloud platform and tooling. The role emphasizes observability, automation, and collaboration with software engineering teams to improve security, availability, and performance.

The hybrid position requires in-office days to be confirmed by the Hiring Manager; candidates should be comfortable working with GitHub, Jenkins, Jira, and Artifactory, and have

Qualifications

  • Bachelor’s degree in IT, CS or related field, or 3+ years of relevant work experience.

Responsibilities

  • DevTools Support—primary contact for developers using GitHub, Jenkins, Jira, or Artifactory.
  • Troubleshoot and resolve tool-related issues promptly to minimize downtime.
  • Maintain and optimize CI-CD pipelines and integrations for reliability and scalability.
  • Collaborate with development teams to improve workflows and automation.
  • Site Reliability Engineering design, implement, and maintain systems for high availability, scalability, and performance.
  • Monitor and improve application reliability through proactive measures and incident response.
  • Develop and maintain observability solutions (metrics, logging, tracing).
  • Participate in on-call rotations and drive root cause analysis for incidents.
  • Collaboration & Continuous Improvement—partner with engineering teams to identify reliability risks and implement best practices.
  • Document processes, troubleshooting guides, and reliability playbooks.
  • Advocate for automation and self-service solutions to reduce operational overhead.

Skills

Python
Java
Go
PowerShell
JavaScript
Terraform
Ansible
Helm
CI/CD

Education

Bachelor’s degree in IT/CS or related field

Tools

GitHub
Jenkins
Jira
Artifactory
ArgoCD
Bitbucket
Azure DevOps
Grafana
Prometheus
Splunk
Datadog
New Relic
Dynatrace
Sentry
Docker
Kubernetes
CloudFormation

Job description

Job Description

Site Reliability Engineering (SRE) is essential to the company’s Cloud platform strategy. In this role, you’ll ensure our development platform and tools let engineers focus on innovation instead of infrastructure. You’ll promote observability best practices and automate resolution of recurring issues, working closely with software engineering teams to support security, availability, and performance. Responsibilities include triaging issues, collaborating on infrastructure management, and setting up monitoring for full coverage. Hands‑on expertise is required, especially with major DevTools like GitHub, Jenkins, Jira, and Artifactory.

We seek a Software Engineer + SRE hybrid engineer. The ideal candidate deeply understands at least one major DevTool, quickly resolves tool‑related issues in collaboration with developers, and applies systems thinking to maintain reliable applications and infrastructure while improving developer productivity.

This is a hybrid position. Expectation of days in the office will be confirmed by your Hiring Manager. The company requires at least three days in office, expectations of these days will be confirmed by your Hiring Manager.

Key Responsibilities
  • DevTools Support—You will be the primary point of contact for developers using tools like GitHub, Jenkins, Jira, or Artifactory.
  • Troubleshoot and resolve tool‑related issues promptly to minimize developer downtime.
  • Maintain and optimize CI‑CD pipelines and integrations for reliability and scalability.
  • Collaborate with development teams to improve workflows and automation.
  • Site Reliability Engineering design, implement, and maintain systems for high availability, scalability, and performance.
  • Monitor and improve application reliability through proactive measures and incident response.
  • Develop and maintain observability solutions (metrics, logging, tracing).
  • Participate in on‑call rotations and drive root cause analysis for incidents.
  • Collaboration & Continuous Improvement—Partner with engineering teams to identify reliability risks and implement best practices.
  • Document processes, troubleshooting guides, and reliability playbooks.
  • Advocate for automation and self‑service solutions to reduce operational overhead.
Qualifications
  • Bachelor’s degree, OR 3+ years of relevant work experience.
  • Bachelor’s degree in IT, CS or related field and/or 3+ years working experience in IT Operations and Delivery.
  • Experience: 0.5–3 years in SRE and/or DevTools support roles.
  • Beginner level programming and/or scripting in 2 or more of the following: Python, Java, Go, PowerShell, JavaScript, Terraform, Ansible, Helm, Chef, CloudFormation.
  • Basic understanding of YAML, JSON, HTML, XML.
  • Hands‑on experience in Linux and/or Windows systems and good understanding of distributed computing environments.
  • Experience with CI‑CD tooling such as Jenkins, GitHub, Bitbucket, ArgoCD, Artifactory, Azure DevOps in a large‑scale environment.
  • Experience with observability tooling such as Grafana, Prometheus, Splunk, Datadog, New Relic, Dynatrace, Sentry, etc. in a large‑scale environment.
  • Experience supporting relational and non‑relational databases (MySQL, MongoDB, PostgreSQL, etc.), including creating and running queries, managing performance and scaling.
  • Working in a Platform, SRE or Production Engineering group for high availability‑critical platforms or applications.
  • Experience managing a distributed container platform including but not limited to deployment‑release management, provisioning, capacity management, workload management.
  • Experience managing container infrastructure and supporting development transformation to a container‑first model.
  • This role requires on‑call support as the team provides 24‑7 operational support.
  • Technical expertise: Proficiency in at least one DevTool (GitHub, Jenkins, ArgoCD, Jira, Artifactory).
  • Strong understanding of CI‑CD principles and pipelines.
  • Solid knowledge of Linux systems, networking, and containerization (Docker‑Kubernetes).
  • Hands‑on experience with cloud platforms.
  • Programming‑Scripting: Proficiency in Python, Ansible, or similar languages.
  • Mindset: Strong problem‑solving skills, systems thinking, self‑starter, and a passion for reliability.

Qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability, or protected veteran status. The company will also consider for employment qualified applicants with criminal histories in a manner consistent with EEOC guidelines and applicable local law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

SRE
SRE

Technopride Ltd • Hove

Hybrid
GBP 60,000 - 80,000
Software Engineer II, Site Reliability Engineering, Labs SRE
Software Engineer II, Site Reliability Engineering, Labs SRE

United States Digital Space LLC • Greater London

On-site
GBP 60,000 - 90,000
SRE Architect (68019)
SRE Architect (68019)

Hitachi Digital Services • Greater London

On-site
GBP 90,000 - 150,000
Site Reliability Engineer
Site Reliability Engineer

iXceed Solutions • Basildon

On-site
GBP 55,000 - 75,000
SRE Architect (68019) (DEAI DS) Cloud & Data Engineering United Kingdom
SRE Architect (68019) (DEAI DS) Cloud & Data Engineering United Kingdom

Hitachids • Greater London

On-site
GBP 90,000 - 140,000
SRE Engineer
SRE Engineer

Savant Recruitment • Greater London

On-site
GBP 60,000 - 80,000
Software Reliability Engineer/Devops
Software Reliability Engineer/Devops

RE Partners • Greater London

Hybrid
GBP 80,000 - 110,000
Site Reliability Engineer
Site Reliability Engineer

Insight International (UK) Ltd • Bournemouth

On-site
GBP 55,000 - 75,000
Site Reliability Engineer
Site Reliability Engineer

DNS INFO LTD • City Of London

On-site
GBP 70,000 - 95,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Reward Gateway • Greater London

Hybrid
GBP 60,000 - 65,000
Hybrid work option