Senior Site Reliability Engineer, Infrastructure

National Geographic

Boston (MA)

On-site

USD 128,000 - 160,000

Full time

6 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

DraftKings is seeking a Senior Site Reliability Engineer in Boston to build and scale our Kubernetes infrastructure across public clouds and on-prem environments. You will design automation-first solutions, implement GitOps delivery with Rancher Fleet, Flux, and Helm, and own scaling and capacity strategies with Karpenter, HPA, and KEDA.

You will develop reliable tooling using Go and Python, monitor reliability with Datadog, and contribute to architectural discussions while participating in

Qualifications

  • Bachelor’s Degree in Computer Science or a related field, or equivalent education, experience, and training.
  • At least 4 years of experience managing distributed cloud and on-premise environments at scale, including strong hands-on experience with Amazon Web Services; experience with Google Cloud Platform, vSphere, or Nutanix is a plus.
  • Deep expertise in Kubernetes and container orchestration, with experience designing, scaling, and troubleshooting complex workloads.
  • Strong software development experience using languages such as Go and Python to build automation and infrastructure tooling.
  • Working knowledge of networking and Linux-based systems, including container runtimes such as Docker and containerd, packet-level debugging, and kernel troubleshooting.
  • Experience with Infrastructure as Code and configuration management tools to build scalable, consistent, and repeatable infrastructure.

Responsibilities

  • Drive stability, performance, and scalability across our global compute platform spanning multiple public clouds and on-premise environments.
  • Build self-healing, fault-tolerant infrastructure and internal tooling that automates repetitive operational work and reduces toil for Platform and Application teams.
  • Operate and evolve our GitOps delivery model, using Rancher Fleet, Flux, and Helm to deploy core Kubernetes services and application workloads consistently and reliably.
  • Own Kubernetes scaling and capacity strategies using technologies including Karpenter, Horizontal Pod Autoscaler (HPA), Kubernetes Event-Driven Autoscaling (KEDA), and predictive scaling based on event and calendar data.
  • Define and monitor service-level objectives and reliability metrics for platform components using Datadog and our logging pipeline.
  • Strengthen our engineering practices by sharing knowledge, contributing to architectural and design discussions, and participating in an on-call rotation.

Skills

Kubernetes
Go
Python
AWS
Linux
Networking
Container runtimes
IaC tools

Education

Bachelor's degree in Computer Science

Tools

Docker
containerd
Terraform

Job description

At DraftKings, AI is becoming an integral part of both our present and future, powering how work gets done today, guiding smarter decisions, and sparking bold ideas. It’s transforming how we enhance customer experiences, streamline operations, and unlock new possibilities. Our teams are energized by innovation and readily embrace emerging technology. We’re not waiting for the future to arrive. We’re shaping it, one bold step at a time. To those who see AI as a driver of progress, come build the future together.

The Crown Is Yours As a Senior Site Reliability Engineer, you’ll build and scale the critical Kubernetes infrastructure that powers our platforms and services. You’ll solve complex reliability challenges across public cloud and on-premise environments, designing automation-first solutions that strengthen performance and simplify operations. You’ll help shape architectural decisions, advance stability at scale, and build tools that give our teams the confidence to move quickly and deliver reliably.

What you’ll do as a Senior Site Reliability Engineer
  • Drive stability, performance, and scalability across our global compute platform spanning multiple public clouds and on-premise environments.
  • Build self-healing, fault-tolerant infrastructure and internal tooling that automates repetitive operational work and reduces toil for Platform and Application teams.
  • Operate and evolve our GitOps delivery model, using Rancher Fleet, Flux, and Helm to deploy core Kubernetes services and application workloads consistently and reliably.
  • Own Kubernetes scaling and capacity strategies using technologies including Karpenter, Horizontal Pod Autoscaler (HPA), Kubernetes Event-Driven Autoscaling (KEDA), and predictive scaling based on event and calendar data.
  • Define and monitor service-level objectives and reliability metrics for platform components using Datadog and our logging pipeline.
  • Strengthen our engineering practices by sharing knowledge, contributing to architectural and design discussions, and participating in an on-call rotation.
What you’ll bring
  • A Bachelor’s Degree in Computer Science or a related field, or equivalent education, experience, and training.
  • At least 4 years of experience managing distributed cloud and on-premise environments at scale, including strong hands-on experience with Amazon Web Services; experience with Google Cloud Platform, vSphere, or Nutanix is a plus.
  • Deep expertise in Kubernetes and container orchestration, with experience designing, scaling, and troubleshooting complex workloads.
  • Strong software development experience using languages such as Go and Python to build automation and infrastructure tooling.
  • Working knowledge of networking and Linux-based systems, including container runtimes such as Docker and containerd, packet-level debugging, and kernel troubleshooting.
  • Experience with Infrastructure as Code and configuration management tools to build scalable, consistent, and repeatable infrastructure.
Join Our Team

We’re a publicly traded (NASDAQ: DKNG) technology company headquartered in Boston. As a regulated gaming company, you may be required to obtain a gaming license issued by the appropriate state agency as a condition of employment. Don’t worry, we’ll guide you through the process if this is relevant to your role.

The US base salary range for this full-time position is 128,000.00 USD - 160,000.00 USD, plus bonus, equity, and benefits as applicable. Our ranges are determined by role, level, and location. The compensation information displayed on each job posting reflects the range for new hire pay rates for the position across all US locations. Within the range, individual pay is determined by work location and additional factors, including job-related skills, experience, and relevant education or training.

It is unlawful in Massachusetts to require or administer a lie detector test as a condition of employment or continued employment. An employer who violates this law shall be subject to criminal penalties and civil liability.

DraftKings Inc. (Nasdaq: DKNG) is a digital sports entertainment and gaming company. It’s simple, at DraftKings, we believe life’s more fun with skin in the game. For that reason, we’re committed to responsibly creating the world’s favorite games and betting experiences. Headquartered in Boston, with offices around the globe, we believe we can continue to define what it means to be a technology company in sports entertainment. We love what we do, and think you will too.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer, Infrastructure
Senior Site Reliability Engineer, Infrastructure

DraftKings • Boston (MA)

On-site
USD 128,000 - 160,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

National Geographic • Boston (MA)

On-site
USD 148,000 - 185,000
Bonus
Equity
Benefits
Principal Site Reliability Engineer
Principal Site Reliability Engineer

DraftKings • Boston (MA)

On-site
USD 200,000 - 250,000
Bonus
Equity
Benefits
Senior Platform Engineer
Senior Platform Engineer

DraftKings Inc. • Boston (MA)

On-site
USD 120,000 - 149,000
Senior Platform Engineer
Senior Platform Engineer

DraftKings • Boston (MA)

On-site
USD 120,000 - 149,000
Senior Platform Engineer
Senior Platform Engineer

DraftKings Inc. • United States

Remote
USD 120,000 - 149,000
Senior Lead Site Reliability Engineer, Database
Senior Lead Site Reliability Engineer, Database

DraftKings • United States

On-site
USD 168,000 - 210,000
Senior Lead Software Engineer, Backend
Senior Lead Software Engineer, Backend

DraftKings • United States

On-site
USD 161,000 - 201,000
Data Engineering Manager, Customer
Data Engineering Manager, Customer

National Geographic • Boston (MA)

On-site
USD 153,000 - 192,000
Bonus
Equity
Benefits
Software Engineer, iOS
Software Engineer, iOS

National Geographic • Boston (MA)

On-site
USD 104,000 - 130,000