AI-First SRE Lead: Multi-Cloud Reliability & Automation

Socotra, Inc.

Poland

Remote

PLN 521,000 - 707,000

Full time

12 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Bonuses
Private medical, life, and disability

Job summary

Avalara is seeking a senior Site Reliability Engineer to advance reliability across our global SaaS platform. You will build automation-first, AI-first reliability ecosystems that improve stability and enable faster product delivery across multi-cloud environments.

You will mentor engineers, shape standards, and drive measurable improvements in reliability and performance, including incident response and post-incident reviews.

Qualifications

  • 10+ years of experience in SaaS, distributed systems, or site reliability engineering.
  • Programming skills in Go, Java, or Python.
  • Deep experience with observability tools such as Prometheus, Grafana, and OpenTelemetry.
  • Hands-on experience with Kubernetes, containerisation, and multi-cloud platforms (AWS, GCP, Azure, or OCI).
  • Strong understanding of Linux systems, networking, and cloud-native architectures.
  • Proven ability to design automation, improve system reliability, and apply AI or machine learning to operational workflows.

Responsibilities

  • Own and evolve the reliability strategy for distributed SaaS systems across multi-cloud platforms.
  • Design and implement AI-driven operations, including predictive monitoring, anomaly detection, and automated root cause analysis.
  • Build and scale observability solutions using Prometheus, Grafana, and OpenTelemetry.
  • Create self-healing systems and automation frameworks that reduce manual operational work.
  • Improve deployment practices using feature flags, progressive delivery, and safe rollout strategies.
  • Ensure reliability and performance of CI/CD pipelines and infrastructure as code environments.
  • Strengthen system availability, scalability, and fault tolerance across Kubernetes-based platforms.
  • Lead incident response, improve recovery times, and implement lasting fixes through post-incident reviews.
  • Integrate AI-driven workflows into incident detection, triage, and resolution to improve operational efficiency.
  • Mentor engineers and drive adoption of automation-first and AI-first reliability practices.

Skills

Distributed systems
Programming Go/Java/Python
Observability
Kubernetes & multi-cloud
Linux & networking
Automation & AI in operations
Incident response

Tools

Prometheus
Grafana
OpenTelemetry
Kubernetes
Containerisation
AWS
GCP
Azure
OCI

Job description

Avalara is seeking a senior Site Reliability Engineer to advance reliability across our global SaaS platform. You will build automation-first, AI-first reliability ecosystems that improve stability and enable faster product delivery across multi-cloud environments.

You will mentor engineers, shape standards, and drive measurable improvements in reliability and performance, including incident response and post-incident reviews.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI-Driven SRE Lead: Scale Reliability Across Multi-Cloud
AI-Driven SRE Lead: Scale Reliability Across Multi-Cloud

Avalara • Poland

Hybrid
PLN 250,000 - 420,000
Private medical insurance
Life insurance
Disability insurance
+2
Remote AI-Driven Senior SRE for Multi-Cloud Reliability
Remote AI-Driven Senior SRE for Multi-Cloud Reliability

Avalara, Inc. • Poland

Remote
PLN 260,000 - 420,000
Senior SRE: Multi-Cloud Reliability & Automation Lead
Senior SRE: Multi-Cloud Reliability & Automation Lead

Sii Poland • Kraków

On-site
PLN 180,000 - 300,000
Great Place to Work
Centre of internal trainings
Profit sharing
+2
Senior Multi-Cloud SRE: Automate & Scale Reliability
Senior Multi-Cloud SRE: Automate & Scale Reliability

Sii Poland • Piła

On-site
PLN 230,000 - 350,000
Great Place to Work
Profit sharing
Internal trainings
+2
Senior SRE: AI-Driven Reliability & Automation
Senior SRE: AI-Driven Reliability & Automation

Jobtailor • Poland

On-site
PLN 180,000 - 320,000
Senior Site Reliability Engineer - Multi-Cloud Automation
Senior Site Reliability Engineer - Multi-Cloud Automation

Sii Poland • Łódź

On-site
PLN 240,000 - 360,000
Great Place to Work
Centre of internal trainings
Profit sharing
+2
Senior SRE: Multi-Cloud Automation & Reliability Lead
Senior SRE: Multi-Cloud Automation & Reliability Lead

Sii Poland • Wrocław

On-site
PLN 260,000 - 360,000
Great Place to Work
Profit sharing
Medical care
+3
Senior SRE: Multi-Cloud Reliability & Automation Lead
Senior SRE: Multi-Cloud Reliability & Automation Lead

Sii Poland • Toruń

On-site
PLN 240,000 - 360,000
Great Place to Work
Solid financial situation
Contracts with the biggest brands
+9
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Avalara • Poland

Hybrid
PLN 250,000 - 420,000
Private medical insurance
Life insurance
Disability insurance
+2
Senior SRE: Cloud-Native Reliability & Automation
Senior SRE: Cloud-Native Reliability & Automation

OneRail USA • Kraków

Hybrid