Senior Site Reliability Engineer, AI-First & Multi-Cloud

Avalara, Inc.

Town of Poland (NY)

On-site

USD 180,000 - 240,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Bonuses
Paid time off
Medical insurance

Job summary

Avalara, Inc. is seeking a senior reliability engineer to lead how reliability is engineered across Avalara's global SaaS platform as we move toward an AI-first operating model.

You will build a modern, automation-first reliability ecosystem that improves stability, reduces risk, and speeds safe product delivery across multi-cloud environments. You will mentor engineers, shape standards, and drive measurable improvements in reliability and performance, including AI-driven monitoring and

Qualifications

  • 10+ years of experience in SaaS, distributed systems, or site reliability engineering.
  • Programming skills in Go, Java, or Python.
  • Deep experience with observability tools such as Prometheus, Grafana, and OpenTelemetry.
  • Hands-on experience with Kubernetes, containerisation, and multi-cloud platforms (AWS, GCP, Azure, or OCI).
  • Strong understanding of Linux systems, networking, and cloud-native architectures.
  • Proven ability to design automation, improve system reliability, and apply AI or machine learning to operational workflows.

Responsibilities

  • Own the reliability strategy for distributed SaaS systems across multi-cloud platforms.
  • Design and implement AI-driven operations, including predictive monitoring and automated root-cause analysis.
  • Build and scale observability using Prometheus, Grafana, and OpenTelemetry.
  • Create self-healing systems and automation to reduce manual work.
  • Improve deployments with feature flags, progressive delivery, and safe rollouts.
  • Ensure reliability of CI/CD pipelines and IaC environments.
  • Strengthen availability, scalability, and fault tolerance on Kubernetes platforms.
  • Lead incident response and drive post-incident improvements.
  • Integrate AI-driven workflows into incident detection and resolution.
  • Mentor engineers and champion automation-first reliability.

Skills

Go
Java
Python

Tools

Prometheus
Grafana
OpenTelemetry

Job description

Avalara, Inc. is seeking a senior reliability engineer to lead how reliability is engineered across Avalara's global SaaS platform as we move toward an AI-first operating model.

You will build a modern, automation-first reliability ecosystem that improves stability, reduces risk, and speeds safe product delivery across multi-cloud environments. You will mentor engineers, shape standards, and drive measurable improvements in reliability and performance, including AI-driven monitoring and

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead Site Reliability Engineer
Lead Site Reliability Engineer

Avalara, Inc. • Town of Poland (NY)

On-site
USD 180,000 - 240,000
Bonuses
Paid time off
Medical insurance
Senior Backend Engineer AI-Driven Reliability (Remote)
Senior Backend Engineer AI-Driven Reliability (Remote)

Affirm • Boise (ID)

On-site
USD 173,000 - 233,000
Health coverage for you and dependents
FSAs - tech spending
Time off
+1
Senior AI-Driven Platform Engineer
Senior AI-Driven Platform Engineer

Avalara, Inc. • United States

On-site
USD 130,000 - 185,000
Bonuses
Medical, life, and disabilityInsurance
Diversity & inclusion
+1
Remote Senior Backend Engineer — AI-Driven Reliability
Remote Senior Backend Engineer — AI-Driven Reliability

Affirm • Richmond (VA)

On-site
USD 173,000 - 233,000
Health care coverage
Flexible Spending Wallets
Time off
+1
Senior Site Reliability Engineer — AI-First Observability
Senior Site Reliability Engineer — AI-First Observability

Qualitest • Riverwoods (IL)

On-site
USD 110,000 - 130,000
Diversity and inclusion
Internal rotation
Career progression
+3
Senior AI‑First DevOps Platform Engineer
Senior AI‑First DevOps Platform Engineer

Socotra, Inc. • United States

Hybrid
USD 120,000 - 150,000
Private medical insurance
Paid parental leave
Inclusive culture initiatives
Senior Site Reliability Engineer - AI-Driven Cloud
Senior Site Reliability Engineer - AI-Driven Cloud

Sight Machine • San Francisco (CA)

Hybrid
USD 200,000 - 260,000
Stock Options
Health Care Coverage
Flexible Vacation Policy
+5
Senior Site Reliability Engineer – AI-First Infra
Senior Site Reliability Engineer – AI-First Infra

Evidently Ltd. • San Francisco (CA), Northern (KY)

Hybrid
USD 200,000 - 250,000
Senior Software Engineer - AI-Driven Cloud Platform
Senior Software Engineer - AI-Driven Cloud Platform

Socotra, Inc. • United States

Remote
USD 100,000 - 130,000
Paid time off
Health and wellness benefits
Potential bonuses
Senior Backend Engineer, AI‑Driven Reliability Platform
Senior Backend Engineer, AI‑Driven Reliability Platform

Affirm • Detroit (MI)

On-site
USD 173,000 - 255,000
Health care coverage
FSA wallets
Time off
+1