Site Reliability Engineer (DataCosmos)

Open Cosmos

Barcelona

On-site

EUR 45,000 - 65,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Open Cosmos, located in Barcelona, seeks a Site Reliability Engineer to enhance the reliability and scalability of its data platform. The role involves ensuring peak performance, monitoring systems end-to-end, and collaborating with engineering to design resilient architectures.

You will manage deployments, automate processes, and troubleshoot incidents, all while working in a fast-paced environment focused on continuous improvement. Strong experience with Linux systems, Kubernetes, and cloud platforms is essential.

Qualifications

  • Strong demonstrable ability to work with Linux systems and cloud platforms.
  • Solid Kubernetes knowledge and ability to run production systems.
  • A clear understanding of observability including monitoring, logging, and tracing.
  • Capable of designing or operating high-availability, distributed systems.
  • A mindset focused on automation, scalability, and continuous improvement.

Responsibilities

  • Ensure our data platform is reliable, scalable, and performing optimally.
  • Monitor systems end-to-end, ensuring full visibility across the infrastructure.
  • Respond to incidents, troubleshoot issues, and drive long-term fixes.
  • Improve deployments and contribute to CI/CD pipelines for repeatable releases.
  • Work closely with engineering teams to design resilient, scalable systems.
  • Automate processes and reduce operational overhead.

Skills

Linux systems
Kubernetes
Cloud platforms (AWS, GCP, Azure)
Observability (monitoring, logging, tracing)
High-availability systems design
Automation
Scalability

Job description

Requirements
  • Strong demonstrable ability to work with Linux systems and cloud platforms (AWS, GCP or Azure)
  • Solid Kubernetes knowledge and ability to run production systems
  • A clear understanding of observability (monitoring, logging, tracing)
  • Capable of designing or operating high-availability, distributed systems
  • A mindset focused on automation, scalability, and continuous improvement
  • Confidence working in fast-moving environments where reliability really matters
What the job involves
  • At Open Cosmos, our Data division transforms satellite data into meaningful insights that drive real-world impact. The team delivers all data products generated by Open Cosmos and its partners, curates and develops DataCosmos (our geospatial data platform) and builds integrations that make satellite imagery easy to access and act on
  • We’re now looking for a Site Reliability Engineer to help us ensure our data platform is reliable, scalable, and performing at its best as we grow
  • Owning the reliability, performance, and scalability of our data platform and processing pipelines
  • Monitoring systems end-to-end, ensuring full visibility across infrastructure and data flows
  • Responding to incidents, troubleshooting issues, and driving long-term fixes
  • Improving deployments and contributing to CI/CD pipelines for safe, repeatable releases
  • Working closely with engineering teams to design resilient, scalable systems
  • Automating processes and reducing operational overhead
  • Supporting customer-impacting issues alongside Customer Success teams
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Backend Engineer (DataCosmos)
Backend Engineer (DataCosmos)

Open Cosmos • Barcelona

On-site
EUR 40,000 - 60,000
Customer Support Specialist (DataCosmos)
Customer Support Specialist (DataCosmos)

Open Cosmos • Barcelona

On-site
EUR 30,000 - 40,000
Software Engineer - On Board Processing
Software Engineer - On Board Processing

Open Cosmos • Barcelona

On-site
EUR 60,000 - 90,000
Staff Site Reliability Engineer
Staff Site Reliability Engineer

Hydrolix • Spain

On-site
EUR 70,000 - 90,000
Customer Support Specialist - Space Data Platform
Customer Support Specialist - Space Data Platform

Open Cosmos • Barcelona

On-site
EUR 30,000 - 40,000
Operations Engineer (AIT Production)
Operations Engineer (AIT Production)

Open Cosmos • Barcelona

On-site
EUR 40,000 - 60,000
Cutting-edge technology
Supportive team environment
Mission-driven work
Site Reliability Engineer ID53670
Site Reliability Engineer ID53670

AgileEngine • Madrid

On-site
EUR 45,000 - 60,000
Professional growth
Competitive compensation
Exciting projects
+1
Site Reliability Engineer ID53670
Site Reliability Engineer ID53670

AgileEngine • Ribarroja del Turia

On-site
EUR 40,000 - 70,000
Professional growth: Mentorship, TechTalks, and personalized growth roadmaps
Competitive compensation: USD-based pay with education, fitness, and team activity budgets
Exciting projects: Modern solutions with Fortune 500 and top product companies
+1
Staff Site Reliability Engineer
Staff Site Reliability Engineer

Doghouse Recruitment • Spain

Remote
EUR 110,000 - 170,000
Systems Engineer
Systems Engineer

Open Cosmos • Spain

On-site
EUR 45,000 - 65,000
Global customers
Mission-driven company
Diverse and supportive team