Senior Site Reliability Engineer

Loft Orbital Solutions

Abu Dhabi

On-site

AED 450,000 - 750,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Loft Orbital Solutions in Abu Dhabi is seeking a Senior Site Reliability Engineer to join our Infrastructure Team. You will work across development, operations and IT to ensure reliable ground segment infrastructure for our space missions, applying DevOps principles to spacecraft control.

The role focuses on building scalable cloud solutions, automating CI/CD and IaC, enhancing observability, and maintaining secure, high-availability systems in a hybrid cloud environment.

Qualifications

  • Experience with public cloud infrastructure, preferably GCP.
  • Deep expertise in Kubernetes architecture deployment ops and resource optimization.
  • Design and build scalable highly available systems.
  • Familiarity with Software Defined Networking (SDN) concepts and tools.
  • Experience implementing and maintaining observability stacks (Grafana Prometheus Loki).
  • Proficiency in at least one backend language: Go Python Rust C/C or Java.
  • DevOps practices: CI/CD, IaC, and automation.
  • Proven track record in fast-paced high-growth technical environments.
  • Strong networking knowledge (TCP/IP DNS routing switching firewalls VPNs).
  • Deep experience in Systems Administration.
  • Excellent problem-solving and proactive results-driven mindset.
  • Strong communication in a multicultural cross-functional team.

Responsibilities

  • Collaborate with developers, test engineers and satellite operators to foster a SatDevOps culture.
  • Design and roll-out cloud solutions for testing and operations infrastructure to scale resources.
  • Design, implement and maintain scalable, reliable, and secure infrastructure in a hybrid cloud.
  • Improve developer and test engineers' experience by building better tools and workflows.
  • Lead automation of CI/CD pipelines, IaC, and deployment workflows for ground tests and space operations.
  • Own and evolve observability stack (metrics, tracing, logs). Grafana stacks are a plus.
  • Implement and advocate for best practices in reliability, fault tolerance and performance tuning.
  • Identify, investigate and resolve system reliability issues with root-cause analyses.
  • Partner with teams to design and operate Software Defined Networking (SDN) solutions.
  • Contribute to a collaborative, inclusive team culture with continuous learning.
  • Initially handle cloud-network-hardware interface and assume IT responsibilities as needed.

Skills

Public cloud infrastructure
Kubernetes
Scalable systems
SDN
Observability stacks
Backend languages
DevOps practices
Fast-paced environments
Networking
Systems Administration
Communication

Tools

Grafana
Prometheus
Loki
ArgoCD
FluxCD

Job description

About the Role:

Orbitworks is revolutionizing access to space by building reliable shareable satellites that drastically reduce the time and complexity traditionally required to get to orbit. We operate satellites fly customer payloads and handle entire missions from end-to-end. Orbitworks is a joint venture between Marlan Space (UAE-based) and Loft Orbital.

As a Senior Site Reliability Engineer on our Infrastructure Team youll play a pivotal role in maintaining and scaling our ground segment infrastructure. Youll collaborate across development operations and IT to ensure the integration delivery and reliability of services that support our test infrastructure and our space operations on Earth and in orbit. This is an exciting opportunity to work on cutting-edge technology and help build modern automated space infrastructure. This is not your typical SRE role we apply DevOps principles even to spacecraft control.

Responsibilities
  • Collaborate with developers test engineers and satellite operators to foster a strong SatDevOps culture.
  • Design and roll-out cloud solutions for our testing and operations infrastructure. Find the best trade-offs between existing and additional cloud resources to scale and help Orbitworks achieve its mission.
  • Design implement and maintain scalable reliable and secure infrastructure in a hybrid cloud environment.
  • Improve our developer and test engineers experience by building better tools workflows and environment to streamline
  • Lead efforts to automate and optimize systems including CI/CD pipelines infrastructure provisioning (IaC) and deployment workflows for test on the ground and operations in space.
  • Own and evolve our observability stack (metrics tracing logs) to improve usability and performance. Grafana-centric ecosystems are a plus.
  • Implement and advocate for best practices in software reliability fault tolerance and performance tuning.
  • Proactively identify investigate and resolve system reliability issues performing root cause analyses and implementing long-term fixes.
  • Partner with teams to design and operate Software Defined Network (SDN) solutions.
  • Contribute to a collaborative and inclusive team culture where respectful debate and continuous learning are celebrated.
  • Initially handle and manage the link between cloud and network/software/hardware infrastructure. Assume I&T (Information technology) responsibilities as much as necessary to start with.
Must Haves:
  • Strong experience with public cloud infrastructure ideally GCP.
  • Deep expertise in Kubernetes architecture deployment ops and resource optimization.
  • Demonstrated ability to design and build scalable highly available systems.
  • Familiarity with Software Defined Networking (SDN) concepts and tools.
  • Experience implementing and maintaining observability stacks (Grafana Prometheus Loki etc.).
  • Proficiency in at least one backend language: Go Python Rust C/C or Java.
  • Deep understanding and hands-on experience with DevOps practices: CI/CD infrastructure as code (IaC) and automation.
  • Proven track record of working in fast-paced high-growth technical environments.
  • Strong networking knowledge (TCP/IP DNS routing switching firewalls VPNs secure networks).
  • Deep experience in Systems Administration.
  • Excellent problem-solving skills and ability to operate independently with a proactive results-driven mindset.
  • Strong communication skills; thrives in a multicultural cross-functional team.
Nice to Have:
  • Hands-on experience with GitOps frameworks (ArgoCD FluxCD).
  • Interest or experience in FinOps and cost-optimized architectures
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Loft Orbital • Abu Dhabi

On-site
AED 300,000 - 450,000
Site Reliability Engineer
Site Reliability Engineer

Cerebras • Abu Dhabi

On-site
AED 330,000 - 441,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Good co India • United Arab Emirates

On-site
AED 240,000 - 480,000
Senior Full-Stack Web Developer
Senior Full-Stack Web Developer

jobs.frontdoordefense.com - Jobboard • Abu Dhabi

On-site
AED 200,000 - 250,000
Senior Full-Stack Web Developer
Senior Full-Stack Web Developer

Loft Orbital • Abu Dhabi

On-site
AED 320,000 - 520,000
Onboard Software Solutions Engineer
Onboard Software Solutions Engineer

ATX Venture Partners • Abu Dhabi

On-site
AED 257,000 - 331,000
DevOps Engineer
DevOps Engineer

Client of Huzzle • Abu Dhabi

On-site
AED 150,000 - 200,000
Competitive salary package
Opportunity to work with cutting-edge technologies
Strong career progression opportunities
Head of Site Reliability Engineering (SRE)
Head of Site Reliability Engineering (SRE)

Mark Williams Recruitment • Abu Dhabi

On-site
AED 380,000 - 700,000
Site Reliability Engineering Manager
Site Reliability Engineering Manager

Open Innovation AI • Abu Dhabi Emirate

On-site
AED 150,000 - 210,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

31 CONCEPT • United Arab Emirates

On-site
AED 300,000 - 460,000