Senior Site Reliability Engineer, Infrastructure

DraftKings

Boston (MA)

On-site

USD 128,000 - 160,000

Full time

7 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

DraftKings in Boston is seeking a Senior Site Reliability Engineer to build and scale Kubernetes infrastructure across public clouds and on‑prem environments, delivering automation‑first solutions that strengthen performance and reliability.

You’ll own Kubernetes scaling, capacity strategies, and fault‑tolerant design, build internal tooling, lead GitOps deployments with Rancher Fleet, Flux, and Helm, and contribute to on‑call rotations and architectural decisions.

Qualifications

  • Bachelor’s degree in CS or related field or equivalent education and training.
  • 4+ years of distributed cloud and on-premise experience at scale, with AWS; GCP, vSphere or Nutanix a plus.
  • Deep Kubernetes and container orchestration expertise; scalable, fault‑tolerant systems.
  • Strong Go and Python for automation and tooling.
  • Solid networking and Linux fundamentals; Docker/containerd experience.
  • Experience with IaC and configuration management for scalable infra.

Responsibilities

  • Drive stability, performance, and scalability across global compute platforms in cloud and on‑prem.
  • Build self‑healing, fault‑tolerant infrastructure and automation tooling to reduce toil.
  • Operate and evolve GitOps delivery with Rancher Fleet, Flux, and Helm.
  • Own Kubernetes scaling and capacity strategies using Karpenter, HPA, and KEDA.
  • Define and monitor SLOs and reliability metrics with Datadog and the logging pipeline.
  • Share knowledge, contribute to architecture, and participate in on‑call rotations.

Skills

Kubernetes
Go
Python
AWS
GCP
Linux
Docker
Infrastructure as Code
Networking
Nutanix

Education

Bachelor’s Degree in Computer Science or related field

Tools

vSphere
Nutanix
Flux
Rancher Fleet
Helm
Karpenter
HPA

Job description

At DraftKings, AI is becoming an integral part of both our present and future, powering how work gets done today, guiding smarter decisions, and sparking bold ideas. It’s transforming how we enhance customer experiences, streamline operations, and unlock new possibilities. Our teams are energized by innovation and readily embrace emerging technology. We’re not waiting for the future to arrive. We’re shaping it, one bold step at a time. To those who see AI as a driver of progress, come build the future together.

The Crown Is Yours

As a Senior Site Reliability Engineer, you’ll build and scale the critical Kubernetes infrastructure that powers our platforms and services. You’ll solve complex reliability challenges across public cloud and on-premise environments, designing automation-first solutions that strengthen performance and simplify operations. You’ll help shape architectural decisions, advance stability at scale, and build tools that give our teams the confidence to move quickly and deliver reliably.

What you’ll do as a Senior Site Reliability Engineer
  • Drive stability, performance, and scalability across our global compute platform spanning multiple public clouds and on-premise environments.

  • Build self-healing, fault-tolerant infrastructure and internal tooling that automates repetitive operational work and reduces toil for Platform and Application teams.

  • Operate and evolve our GitOps delivery model, using Rancher Fleet, Flux, and Helm to deploy core Kubernetes services and application workloads consistently and reliably.

  • Own Kubernetes scaling and capacity strategies using technologies including Karpenter, Horizontal Pod Autoscaler (HPA), Kubernetes Event-Driven Autoscaling (KEDA), and predictive scaling based on event and calendar data.

  • Define and monitor service-level objectives and reliability metrics for platform components using Datadog and our logging pipeline.

  • Strengthen our engineering practices by sharing knowledge, contributing to architectural and design discussions, and participating in an on-call rotation.

What you’ll bring
  • A Bachelor’s Degree in Computer Science or a related field, or equivalent education, experience, and training.

  • At least 4 years of experience managing distributed cloud and on-premise environments at scale, including strong hands-on experience with Amazon Web Services; experience with Google Cloud Platform, vSphere, or Nutanix is a plus.

  • Deep expertise in Kubernetes and container orchestration, with experience designing, scaling, and troubleshooting complex workloads.

  • Strong software development experience using languages such as Go and Python to build automation and infrastructure tooling.

  • Working knowledge of networking and Linux-based systems, including container runtimes such as Docker and containerd, packet-level debugging, and kernel troubleshooting.

  • Experience with Infrastructure as Code and configuration management tools to build scalable, consistent, and repeatable infrastructure.

Join Our Team

We’re a publicly traded (NASDAQ: DKNG) technology company headquartered in Boston. As a regulated gaming company, you may be required to obtain a gaming license issued by the appropriate state agency as a condition of employment. Don’t worry, we’ll guide you through the process if this is relevant to your role.

The US base salary range for this full-time position is 128,000.00 USD - 160,000.00 USD, plus bonus, equity, and benefits as applicable. Our ranges are determined by role, level, and location. The compensation information displayed on each job posting reflects the range for new hire pay rates for the position across all US locations. Within the range, individual pay is determined by work location and additional factors, including job-related skills, experience, and relevant education or training. Your recruiter can share more about the specific pay range and how that was determined during the hiring process. It is unlawful in Massachusetts to require or administer a lie detector test as a condition of employment or continued employment. An employer who violates this law shall be subject to criminal penalties and civil liability.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead Site Reliability Engineer
Lead Site Reliability Engineer

DraftKings • Boston (MA), Northern (KY)

Hybrid
USD 148,000 - 185,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

DraftKings Inc. • Boston (MA)

On-site
USD 148,000 - 185,000
AI-Driven Database Reliability Engineer
AI-Driven Database Reliability Engineer

DraftKings Inc. • Boston (MA)

On-site
USD 112,000 - 140,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

National Geographic • Boston (MA)

On-site
USD 148,000 - 185,000
Bonus
Equity
Benefits
Senior Lead Database Reliability Engineer
Senior Lead Database Reliability Engineer

DraftKings Inc. • Massachusetts

On-site
USD 168,000 - 210,000
Database Reliability Engineer
Database Reliability Engineer

DraftKings Inc. • Boston (MA)

On-site
USD 112,000 - 140,000
Data Engineering Manager, Customer
Data Engineering Manager, Customer

DraftKings • Boston (MA)

On-site
USD 153,000 - 192,000
Software Architect
Software Architect

DraftKings Inc. • Boston (MA)

On-site
USD 185,000 - 232,000
Senior Manager, HR Systems
Senior Manager, HR Systems

DraftKings Inc. • Boston (MA)

On-site
USD 149,000 - 186,000
Data Engineering Manager, Customer
Data Engineering Manager, Customer

DraftKings Inc. • Boston (MA)

On-site
USD 153,000 - 192,000
Equity
Benefits