Lead Site Reliability Engineer

Avalara

Poland

Hybrid

PLN 250,000 - 420,000

Full time

9 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Private medical insurance
Life insurance
Disability insurance
Bonuses
Employee resource groups

Job summary

Avalara seeks a senior site reliability engineer to shape reliability strategy across the global SaaS platform. You will build an automation-first, AI-led reliability ecosystem, improve observability, and enable faster, safer product delivery in multi-cloud environments.

As a senior IC, you’ll mentor engineers, raise the technical bar, design self-healing systems, improve deployment practices with feature flags and progressive delivery, and strengthen CI/CD pipelines and IaC across

Qualifications

  • 10+ years in SaaS/distributed systems or site reliability engineering
  • Proficiency in Go, Java, or Python
  • Deep experience with observability tools (Prometheus, Grafana, OpenTelemetry)
  • Hands-on with Kubernetes, containers, multi-cloud (AWS, GCP, Azure, OCI)
  • Strong understanding of Linux, networking, and cloud-native architectures
  • Proven ability to automate and apply AI/ML to operational workflows

Responsibilities

  • Lead reliability strategy for distributed SaaS systems across multi-cloud platforms
  • Design and implement AI-driven operations (predictive monitoring, anomaly detection, RCA)
  • Build observability solutions using Prometheus, Grafana, OpenTelemetry
  • Create self-healing systems and automation frameworks
  • Improve deployment practices with feature flags, progressive delivery, safe rollout
  • Ensure reliability and performance of CI/CD and IaC environments
  • Strengthen availability, scalability, fault tolerance across Kubernetes platforms
  • Lead incident response, improve MTTR, drive post-incident reviews
  • Integrate AI-driven workflows into incident detection, triage, resolution
  • Mentor engineers and promote automation-first and AI-first reliability practices

Skills

SRE leadership
Go/Java/Python
Observability
Kubernetes & multi-cloud
Automation/AI in operations
Linux & networking

Tools

Prometheus
Grafana
OpenTelemetry
Kubernetes
AWS
GCP/Azure/OCI

Job description

What You'll Do

You will lead how reliability is engineered across Avalara's global SaaS platform as we scale and move toward an AI-first operating model. You will focus on building a modern, automation-first reliability ecosystem that improves system stability, reduces operational risk, and enables faster, safer product delivery. You will work across multi-cloud environments to design self-healing systems, advance observability, and modernise deployment practices. As a senior individual contributor, you will also raise the technical bar by shaping standards, mentoring engineers, and driving measurable improvements in reliability and performance.

What Your Responsibilities Will Be
  • Own and evolve the reliability strategy for distributed SaaS systems across multi-cloud platforms
  • Design and implement AI-driven operations, including predictive monitoring, anomaly detection, and automated root cause analysis
  • Build and scale observability solutions using tools such as Prometheus, Grafana, and OpenTelemetry
  • Create self-healing systems and automation frameworks that reduce manual operational work
  • Improve deployment practices using feature flags, progressive delivery, and safe rollout strategies
  • Ensure reliability and performance of CI/CD pipelines and infrastructure as code environments
  • Strengthen system availability, scalability, and fault tolerance across Kubernetes-based platforms
  • Lead incident response, improve recovery times, and implement lasting fixes through post-incident reviews
  • Integrate AI-driven workflows into incident detection, triage, and resolution to improve operational efficiency
  • Mentor engineers and drive adoption of automation-first and AI-first reliability practices
What You'll Need To Be Successful
  • 10+ years of experience in SaaS, distributed systems, or site reliability engineering
  • Programming skills in Go, Java, or Python
  • Deep experience with observability tools such as Prometheus, Grafana, and OpenTelemetry
  • Hands-on experience with Kubernetes, containerisation, and multi-cloud platforms (AWS, GCP, Azure, or OCI)
  • Strong understanding of Linux systems, networking, and cloud-native architectures
  • Proven ability to design automation, improve system reliability, and apply AI or machine learning to operational workflows
Avalara is an AI-first Company

AI is embedded in our workflows, decision-making, and products. Success here requires embracing AI as an essential capability.

  • You'll bring experience using AI and AI-related technologies, ready to thrive here.
  • You'll apply AI every day to business challenges - improving efficiency, contributing solutions, and driving results for your team, our company, and our customers.
  • You'll grow with AI by staying curious about new trends and best practices, and by sharing what you learn so others can benefit too.
How We'll Take Care Of You
Total Rewards

In addition to a great compensation package, paid time off, and paid parental leave, many Avalara employees are eligible for bonuses.

Health & Wellness

Benefits vary by location but generally include private medical, life, and disability insurance.

Inclusive culture and diversity

Avalara strongly supports diversity, equity, and inclusion, and is committed to integrating them into our business practices and our organizational culture. We also have a total of 8 employee-run resource groups, each with senior leadership and exec sponsorship.

What You Need To Know About Avalara

We're defining the relationship between tax and tech.

We've already built an industry-leading cloud compliance platform, processing over 54 billion customer API calls and over 6.6 million tax returns a year. Our growth is real - we're a billion dollar business - and we're not slowing down until we've achieved our mission - to be part of every transaction in the world.

We're bright, innovative, and disruptive, like the orange we love to wear. It captures our quirky spirit and optimistic mindset. It shows off the culture we've designed, that empowers our people to win. We've been different from day one. Join us, and your career will be too.

We're An Equal Opportunity Employer

Supporting diversity and inclusion is a cornerstone of our company — we don't want people to fit into our culture, but to enrich it. All qualified candidates will receive consideration for employment without regard to race, color, creed, religion, age, gender, national orientation, disability, sexual orientation, US Veteran status, or any other factor protected by law. If you require any reasonable adjustments during the recruitment process, please let us know.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead Site Reliability Engineer
Lead Site Reliability Engineer

Avalara, Inc. • Poland

Remote
PLN 260,000 - 420,000
Remote AI-Driven Senior SRE for Multi-Cloud Reliability
Remote AI-Driven Senior SRE for Multi-Cloud Reliability

Avalara, Inc. • Poland

Remote
PLN 260,000 - 420,000
AI-First SRE Lead: Multi-Cloud Reliability & Automation
AI-First SRE Lead: Multi-Cloud Reliability & Automation

Socotra, Inc. • Poland

Remote
PLN 521,000 - 707,000
Bonuses
Private medical, life, and disability
Staff Site Reliability Engineer
Staff Site Reliability Engineer

Assured • Poland

Hybrid
PLN 80,000 - 120,000
Competitive salary and equity packages
Platinum medical, dental, and vision healthcare plan
Unlimited PTO
+4
AI-Driven SRE Lead: Scale Reliability Across Multi-Cloud
AI-Driven SRE Lead: Scale Reliability Across Multi-Cloud

Avalara • Poland

Hybrid
PLN 250,000 - 420,000
Private medical insurance
Life insurance
Disability insurance
+2
Product Lead, Data & Integrations
Product Lead, Data & Integrations

Assured • Poland

Remote
PLN 335,000 - 448,000
Competitive salary
Platinum healthcare plan
Free life insurance
+5
QA Automation Team Lead
QA Automation Team Lead

LM Wind Power / GE • Poland

Hybrid
PLN 180,000 - 240,000
Generous Paid Time Off
Volunteer Time Off
Competitive benefits
+2
Senior AI Machine Learning Engineer
Senior AI Machine Learning Engineer

BlackLine • Kraków

On-site
PLN 140,000 - 210,000
Hybrid work model
Senior AI Machine Learning Engineer
Senior AI Machine Learning Engineer

Blackline SP ZOO (Poland) • Kraków

Hybrid
PLN 180,000 - 260,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Akamai Technologies GmbH • Kraków

Hybrid
PLN 90,000 - 130,000
Flexible working options
Health benefits
Professional development opportunities