Senior Site Reliability Engineer (Cloud and Networking) - Remote

Akamai Technologies

Kraków

On-site

PLN 90,000 - 130,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health benefits
Financial support
Family support
Flexible working hours

Job summary

Akamai Technologies is looking for a Senior Site Reliability Engineer to own the SRE lifecycle for NodeBalancer and Network Load Balancer in Kraków, Poland. This role involves designing SLO/SLI frameworks, leading incident responses, and building automation for cloud infrastructure.

Candidates should have extensive experience in SRE, strong expertise in Linux networking, and proficiency in L4/L7 load balancers. Benefits include health, finance, and family support, as well as time for personal pursuits.

Qualifications

  • Extensive experience in SRE, platform engineering, or infrastructure engineering.
  • Deep expertise with Linux networking fundamentals and diagnosing at packet level.
  • 4+ years in SRE or infrastructure engineering, with at least 2 years at cloud scale.

Responsibilities

  • Own the SRE lifecycle for NodeBalancer and Network Load Balancer.
  • Design and implement SLO/SLI frameworks.
  • Lead technical incident response for complex NB/NLB failures.

Skills

SRE experience
Linux networking fundamentals
L4/L7 load balancing technologies
Kubernetes and containerization
Automation using Python or Go

Tools

Prometheus
Grafana
SaltStack
Ansible
Terraform

Job description

Job Description

Do you want to own the reliability of cloud load balancing infrastructure that serves thousands of customers at global scale? Are you a senior technical leader who can drive solutions across distributed teams while mentoring the engineers around you? Join our Cloud Networking SRE Team.

The Cloud Networking SRE team (CNETSRE) is part of Akamai's Infrastructure Engineering & Operations (IE&O) organization. We design, deploy, and manage the reliability of Akamai's core cloud networking products — including NodeBalancer, our production L4/L7 load balancer, and NLB, our next-generation high‑throughput L4 load balancing platform. These products are foundational to the Akamai Cloud Compute platform, serving customer workloads across dozens of global regions.

Responsibilities
  • Own the SRE lifecycle for NodeBalancer and Network Load Balancer—from design reviews and pre‑rollout readiness assessments through production sign‑off and ongoing reliability management.
  • Design and implement SLO/SLI frameworks that reflect true customer experience for L4 and L7 load balancing services, and drive action when error budgets are at risk.
  • Build and maintain observability pipelines for NB/NLB infrastructure, including Prometheus metrics and Grafana dashboards that enable rapid incident triage.
  • Lead technical incident response for complex NB/NLB failures—BGP/VIP issues, failover failures, data plane degradations, and configuration problems—acting as the technical commander and driving root cause analysis and preventive follow‑through.
  • Develop and automate safe deployment workflows for phased NB/NLB releases, including bake period monitoring, feature‑flag management, and GO/NO‑GO validation across global datacenter rollouts.
  • Review design documents and product requirement documents and provide actionable SRE input on operational risks, capacity implications, Day‑2 concerns, and product strategy gaps.
  • Build automation and tooling using Python or Go that reduces operational toil and improves team‑wide operational capability.
  • Mentor SRE II engineers on the NB team, providing hands‑on technical guidance, code/config reviews, and raising the bar for the team's SRE practice.
  • Participate in an on‑call rotation for NB/NLB production systems, responding to incidents and driving resolution for customer‑facing load balancing infrastructure.
Qualifications
  • Extensive experience in SRE, platform engineering, or infrastructure engineering, working with large‑scale distributed systems.
  • Deep expertise with Linux networking fundamentals—routing, BGP, nftables/iptables, ARP, VXLAN—and comfort diagnosing at the packet level using tcpdump, netstat, and similar tools.
  • Hands‑on experience with L4/L7 load balancing technologies—proxy‑based or kernel‑level load balancers—covering configuration, health checking, high availability, and failure modes at scale.
  • Track record of defining SLO/SLI frameworks, building observability platforms from scratch, and running incident management processes at scale.
  • Expertise in Kubernetes and containerization at scale—including workload scheduling, networking (CNI, Services, ingress), resource management, and operating stateful or network‑intensive workloads in a cluster environment.
  • Build automation and tooling using Python or Go, with infrastructure‑as‑code experience (SaltStack, Ansible, or Terraform) and strong deployment safety instincts.
  • 4+ years in SRE or infrastructure engineering, with at least 2 years at cloud scale.
Benefits
  • Your health
  • Your finances
  • Your family
  • Your time at work
  • Your time pursuing other endeavors
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Cloud Networking SRE — Lead L4/L7 Load Balancer
Senior Cloud Networking SRE — Lead L4/L7 Load Balancer

Akamai Technologies • Kraków

On-site
PLN 90,000 - 130,000
Health benefits
Financial support
Family support
+1
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Akamai Technologies • Kraków

On-site
PLN 180,000 - 270,000
Senior II Site Reliability Engineer
Senior II Site Reliability Engineer

Akamai Technologies • Kraków

On-site
PLN 300,000 - 520,000
Benefits at Akamai
FlexBase program
Hybrid/work flexibility
Senior Cloud Networking SRE — NodeBalancer & NLB
Senior Cloud Networking SRE — NodeBalancer & NLB

Akamai Technologies GmbH • Kraków

Remote
PLN 253,000 - 381,000
Health benefits
Development opportunities
Work-life balance support
Senior II Site Reliability Engineer
Senior II Site Reliability Engineer

Akamai Technologies GmbH • Kraków

On-site
PLN 200,000 - 320,000
FlexBase program
Comprehensive benefits
Senior Site Reliability Engineer (Core Linux Platforms) - Remote
Senior Site Reliability Engineer (Core Linux Platforms) - Remote

Akamai Technologies • Kraków

On-site
PLN 180,000 - 260,000
FlexBase work flexibility
Health and wellbeing benefits
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Akamai Technologies GmbH • Kraków

Hybrid
PLN 90,000 - 130,000
Flexible working options
Health benefits
Professional development opportunities
Senior Site Reliability Engineer (Linux Performance) - Remote
Senior Site Reliability Engineer (Linux Performance) - Remote

Akamai Technologies • Kraków

Hybrid
PLN 120,000 - 160,000
Health benefits
Well-being support
FlexBase program
Senior Site Reliability Engineer (Server Enablement & Qualification) - Remote
Senior Site Reliability Engineer (Server Enablement & Qualification) - Remote

Akamai Technologies • Kraków

Hybrid
PLN 80,000 - 110,000
Comprehensive benefits
Flexibility to work from home or in-office
Hybrid work model
Senior Site Reliability Engineer - Remote
Senior Site Reliability Engineer - Remote

Akamai Technologies • Kraków

On-site
PLN 90,000 - 120,000
Health benefits
Financial benefits
Family support
+2