Senior Site Reliability Expert

Upserve

Deutschland

Hybrid

EUR 80.000 - 110.000

Vollzeit

14 Tage+

Erhalte mehr Antworten von Arbeitgebern

Versende in nur wenigen Minuten einen passgenauen Lebenslauf.

Zusammenfassung

Upserve's SRE team designs, operates and ensures the reliability of the product infrastructure. The role collaborates with developers, QA and PMs across locations to accelerate product development.

We emphasize automation, IaC and scalable cloud architectures, with on-call participation during incidents and a focus on cost-efficiency and security.

Qualifikationen

  • Strong knowledge of Amazon Web Services.
  • Experience with Docker, Kubernetes and Linux systems.
  • Experience with configuration management tools: Chef, Puppet, Ansible, Salt.
  • Experience with Infrastructure as code: Terraform and OpenTofu.
  • Shell scripting; programming languages like Python, Ruby, Go.

Aufgaben

  • Design, operate and ensure reliability of Upserve's product infrastructure.
  • Improve software delivery processes in a multi-location team.
  • Use automation to design, configure, manage and monitor systems.
  • Build scalable, cost-efficient cloud infrastructure with developers.
  • Follow IaC, monitoring, HA, DR, security and SRE/DevOps practices.
  • Provide on-call remediation during production incidents.

Kenntnisse

Amazon Web Services
Docker
Kubernetes
Linux
Shell scripting
Python
Ruby
Go
Agile
Team collaboration

Tools

Terraform
OpenTofu
Chef
Puppet
Ansible
Salt

Jobbeschreibung

Role Summary

Our SRE team is responsible for the design, operation and reliability of Upserve's product infrastructure. We collaborate with teams across the company to make this happen: Developers, QA, PMs, etc.

Key Responsibilities
  • Initiate and contribute to continuous improvement of our software delivery processes and practices in a multi-location, multidisciplinary team to empower and accelerate product development
  • Use automation extensively to design, configure, manage, and monitor systems in support of our product development teams
  • Design and architect operational solutions with the specific goal of increasing the standardization, automation, repeatability, cost-efficiency and consistency of operational tasks
  • Working with developers and other SREs to design and build scalable, reliable and cost-efficient Cloud infrastructure
  • Adhere to and advocate for best practices, including Infrastructure as Code, monitoring, high availability, disaster recovery, security, and SRE/DevOps methodologies
  • Provide timely assistance and remediation solutions during critical situations and production incidents to help resolve service problems (You will be on call for periods of time)
Required Qualifications
  • Strong knowledge of Amazon Web Services
  • Strong experience with Docker, Kubernetes & Linux Systems
  • Experience with configuration management tools such as Chef, Puppet, Ansible, Salt
  • Experience with Infrastructure as code practices: we use Terraform & OpenTofu
  • Ability to read & write complex scripts using Shell
  • Ability to read & understand programming languages: Python, Ruby, Go, etc.
  • Good understanding of Agile development and continuous delivery best practices, software engineering tools, processes, methods and testing
  • Ability to collaborate effectively with other teams
  • Ability to plan, organize, prioritize and stay focused
  • Good experience provisioning and managing infrastructures with high availability constraints
  • Good experience with cloud cost optimization
First 90 Days: Success Outcomes
  • You are a problem solver who does not shy away from tackling complexity and critical thinking
  • You have a strong will to learn, grow and get out of your comfort zone
  • You have great energy and passion for technology
  • You are able to express yourself flawlessly in English
  • You have strong interpersonal skills
Opportunity
  • Lots of autonomy, flexible work culture and possibility of remote work
  • Development of high traffic products, used at the global scale
  • Exposure to modern and proven technology
  • Opportunity to learn and expand your skill set
  • Tons of growth opportunities into technical or people management roles
  • Opportunity to join a fast-paced, high-growth company
Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Senior Site Reliability Engineer
Senior Site Reliability Engineer

PointClickCare • Deutschland

Hybrid
EUR 90.000 - 130.000
Senior SRE
Senior SRE

CloudFactory • Deutschland

Hybrid
EUR 90.000 - 120.000
Great Mission and Culture
Meaningful Work
Market competitive salary
+5
Site Reliability Engineer
Site Reliability Engineer

Apprize Technology Solutions • Deutschland

Vor Ort
EUR 70.000 - 90.000
Senior Site Reliability Engineer / SRE - Kubernetes & Hybrid Cloud (m/f/d)
Senior Site Reliability Engineer / SRE - Kubernetes & Hybrid Cloud (m/f/d)

FactFinder • Berlin

Hybrid
EUR 90.000 - 140.000
Hybrid work model (3 office days/week)
Staff Site Reliability Engineer (x/f/m)
Staff Site Reliability Engineer (x/f/m)

Meyandy LLC • Berlin

Hybrid
EUR 110.000 - 140.000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Forto • Berlin

Vor Ort
EUR 90.000 - 140.000
Manager, Site Reliability Engineering
Manager, Site Reliability Engineering

delinea • Deutschland

Hybrid
EUR 120.000 - 180.000
Healthcare insurance
Pension plan
Life insurance
+1
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Meyandy LLC • Berlin

Hybrid
EUR 90.000 - 130.000
Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d)
Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d)

FACT-Finder • Pforzheim

Hybrid
EUR 90.000 - 125.000
Hybrid work model
Senior Site Reliability Engineer
Senior Site Reliability Engineer

CloudFactory Limited • Berlin

Vor Ort
EUR 90.000 - 130.000