Senior Site Reliability Engineer

Nordic Investin Group Aktiebolag

Stockholms kommun

On-site

SEK 900,000 - 1,100,000

Full time

5 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Nordic Investin in Sweden is seeking a Senior Site Reliability Engineer to turn reliability into measurable targets, telemetry and resilient systems. You will partner with development teams to improve availability, latency and production insights.

The role blends hands-on delivery with design leadership: define SLIs/SLOs, enhance telemetry, automate tasks, and guide incident response. Strong SRE experience and knowledge of Python/Go/Java, cloud, Kubernetes and automation are key.

Qualifications

  • Strong production engineering or SRE experience.
  • Experience with Python, Go or Java.
  • Cloud, Kubernetes and infrastructure automation experience.
  • Experience designing metrics, logs, traces and alerts.
  • Calm incident leadership and structured problem solving.
  • Professional English and collaborative communication.

Responsibilities

  • Define service level indicators and objectives with product teams.
  • Improve telemetry, alert quality and operational dashboards.
  • Automate recurring operational tasks and recovery procedures.
  • Lead or support incident response and blameless reviews.
  • Test capacity, failure modes and disaster recovery.
  • Influence architecture through production evidence.

Skills

SRE experience
Python/Go/Java
Cloud/Kubernetes
Metrics/Logs/Traces/Alerts
Incident leadership
English communication

Job description

On behalf of a partner company, Nordic Investin is looking for a Senior Site Reliability Engineer. You turn reliability from a vague ambition into explicit service targets, useful telemetry and engineering work with clear priorities.

The partner runs digital services where availability and latency directly affect customers. You will work with development teams to improve resilience, incident response and the systems used to understand production behaviour.

The platform is treated as a service used by engineers and business teams. Success will be measured through safer change, lower operational friction and clearer ownership rather than the number of tools introduced.

How you will work

The work combines planned platform development with investigation of real production behaviour. You will collaborate with application teams, security and operations, using their feedback to decide what should become a shared service, an automated control or clear documentation. Ownership continues after the first release.

As a senior colleague, you will own substantial outcomes and help others make stronger decisions. You are expected to recognise risk early, communicate it without drama and move work forward with practical alternatives. The role still includes hands on delivery; seniority here means broader judgement, not distance from the work.

What you will do
  • Define service level indicators and objectives with product teams.

  • Improve telemetry, alert quality and operational dashboards.

  • Automate recurring operational tasks and recovery procedures.

  • Lead or support incident response and blameless reviews.

  • Test capacity, failure modes and disaster recovery.

  • Influence architecture through production evidence.

What you will bring
  • Strong production engineering or SRE experience.

  • Good software development ability in Python, Go, Java or similar.

  • Cloud, Kubernetes and infrastructure automation knowledge.

  • Experience designing metrics, logs, traces and alerts.

  • Calm incident leadership and structured problem solving.

  • Professional English and collaborative communication.

Experience that would add value
  • High traffic consumer or transaction platforms.

  • Chaos engineering and resilience testing.

  • Experience establishing SRE practices in a product organisation.

A background that can succeed here

You may have grown from operations, software engineering or a platform team. What matters is that you can improve reliability through code and design, and that you understand both the technical and human sides of incidents.

What makes the opportunity interesting

The partner offers services with enough scale for reliability work to matter and enough openness for you to change how it is done. Your improvements will be visible in customer experience and engineering focus.

The exact partner, employment model, compensation, start date and working arrangement will be explained openly during the process. Nordic Investin will make sure you understand the context, expectations and decision path before you are asked to commit significant time.

The recruitment conversation

During the process, Nordic Investin will focus on concrete decisions you have made: the context you received, the alternatives you considered, the result you observed and what you would change today. You do not need every optional technology if your core experience transfers and you can explain how you would close the gap.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer: Resilience & Incidents
Senior Site Reliability Engineer: Resilience & Incidents

Nordic Investin Group Aktiebolag • Stockholms kommun

On-site
SEK 900,000 - 1,100,000
Senior Observability Platform Engineer
Senior Observability Platform Engineer

Nordic Investin Group Aktiebolag • Stockholms kommun

On-site
SEK 900,000 - 1,200,000
Senior SRE / Site Reliability Engineer
Senior SRE / Site Reliability Engineer

Aurora Engineering AB • Göteborgs kommun

On-site
SEK 900,000 - 1,300,000
Senior AWS Platform Engineer
Senior AWS Platform Engineer

Nordic Investin Group Aktiebolag • Göteborgs kommun

On-site
SEK 900,000 - 1,150,000
Junior IT Operations Engineer
Junior IT Operations Engineer

Nordic Investin Group Aktiebolag • Stockholms kommun

On-site
SEK 379,000 - 580,000
Kubernetes Platform Engineer
Kubernetes Platform Engineer

Nordic Investin Group Aktiebolag • Stockholms kommun

Hybrid
SEK 900,000 - 1,300,000
Senior Java Platform Engineer
Senior Java Platform Engineer

Nordic Investin Group Aktiebolag • Göteborgs kommun

On-site
SEK 900,000 - 1,200,000
Engineering Manager
Engineering Manager

Nordic Investin Group Aktiebolag • Stockholms kommun

On-site
SEK 1,200,000 - 1,900,000
Site Reliability Engineer
Site Reliability Engineer

Menlo Ventures • Stockholms kommun

On-site
Confidential
Senior Azure Cloud Engineer
Senior Azure Cloud Engineer

Nordic Investin Group Aktiebolag • Stockholms kommun

On-site
SEK 900,000 - 1,150,000