Staff Software Engineer, Billing

Docker, Inc.

Seattle, Northern (WA, KY)

Hybrid

USD 170,000 - 276,000

Full time

5 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Remote-first by design
Time off / PTO
Equity for all full-time employees
Parental leave
Technology stipend
Learning stipend
Illness and holidays coverage

Job summary

Docker, Inc. is seeking an experienced Staff Infrastructure Engineer to own and evolve Billing Platform infrastructure in a remote-first, globally distributed team. You will design IaC patterns, build robust observability, and drive reliability for AI‑assisted deployments.

The role emphasizes secure, scalable infrastructure and collaboration across multiple teams. With 8+ years in platform or SRE roles and deep AWS/Terraform expertise, you will influence engineering practices, mentor peers, and

Qualifications

  • 8+ years in platform, infra, or SRE roles supporting production SaaS systems at scale.
  • Deep AWS expertise: ECS or EKS, RDS, networking, IAM, cost management.
  • Expert-level Terraform; reusable module patterns and standards.
  • Experience building observability stacks (Datadog, Grafana, or similar) at scale.
  • Strong familiarity with CI/CD systems (Jenkins, GitHub Actions) and pipeline ownership.
  • Kubernetes at an operational and architectural level.
  • Track record of identifying systemic risks across teams and driving improvements.
  • Security-first mindset: threat modeling, blast radius, least-privilege, audit trails as design requirements.
  • Strong written English; ability to scale influence via written communication.
  • Bachelor’s degree in CS, Engineering, or related field, or equivalent experience.

Responsibilities

  • Own and evolve the infrastructure supporting Billing Platform services: compute, storage, networking, CI/CD, and observability.
  • Design and maintain IaC (Terraform) for billing system infrastructure on AWS; set module patterns and standards.
  • Build and own observability systems — metrics, logging, alerting — with a focus on billing accuracy and payment reliability.
  • Define deployment patterns and runbooks for AI-agent-assisted development workflows: safe rollback and automated validation.
  • Partner with software engineers on service design, bringing infra constraints into planning before code is written.
  • Identify systemic risks and drive improvements across teams or org boundaries.
  • Lead incident response for billing system issues; participate in on-call rotations as needed.
  • Mentor engineers to raise the overall technical floor.

Skills

8+ years experience
AWS expertise
Terraform expert
Observability stacks
CI/CD pipelines
Kubernetes operator
Security‑first mindset
Strong English writing

Education

Bachelor's degree in CS/Engineering

Tools

Datadog
Grafana
Jenkins
GitHub Actions
Terraform

Job description

About Docker

Docker has been one of the most loved brands in developer tooling, trusted by more than 20 million monthly users and over 20 billion container image pulls. From solo founders to the world's largest companies, developers rely on Docker to build, share, and run their applications across our suite of products including Docker Desktop, Docker Hub, and Docker Scout.
We are a globally distributed, remote-first team building the tools that define how software gets built and delivered. As AI agents redefine software development, Docker is at the center of that shift, providing the sandboxed environments, verified images, and secure infrastructure that make autonomous workflows trustworthy by default.

_______________________________________________________________________

We're building AI-native development practices into how this team works at a foundational level. That means infrastructure design needs to account for a new kind of collaborator: AI agents that generate, deploy, and operate software. The Staff Infrastructure Engineer on this team will keep systems running, define what safe, observable, AI-assisted infrastructure operations look like in practice, and set the standard for how the broader engineering organization follows.

What you'll work on

The Billing Platform Engineering team owns the systems that make Docker's commercial model real. You'll work on problems like:

  • How do we design infrastructure that makes AI-generated deployments safe to ship and easy to roll back?

  • How do we instrument billing systems so that failures — billing miscalculations, entitlement gaps, payment errors — are detected immediately and unambiguously?

  • How do we build infrastructure that scales with usage-based billing workloads without manual intervention?

  • How do we make the developer experience on this team faster and more reliable — local environments, CI/CD pipelines, deployment tooling?

Responsibilities
  • Own and evolve the infrastructure supporting Billing Platform services: compute, storage, networking, CI/CD, and observability

  • Design and maintain IaC (Terraform) for billing system infrastructure on AWS; set module patterns and standards for the team

  • Build and own observability systems — metrics, logging, alerting — with a focus on billing accuracy and payment reliability

  • Define deployment patterns and runbooks that work well in an AI-agent-assisted development workflow: clear rollback procedures, safe promotion gates, automated validation

  • Partner with software engineers on service design — bringing infrastructure constraints and operational requirements into the conversation before code is written

  • Identify systemic risks and drive improvements that span team or organizational boundaries

  • Lead incident response for billing system issues. This role may require participation in an on‑call rotation to provide support outside of standard business hours, including evenings, weekends, and holidays, as needed.

  • Mentor engineers across the team; your technical judgment should raise the floor for everyone

Qualifications
  • 8+ years in platform, infrastructure, or SRE roles supporting production SaaS systems at scale

  • Deep AWS expertise: ECS or EKS, RDS (Postgres preferred), networking, IAM, cost management — you've operated these systems under real load and real incidents

  • Expert‑level Terraform; you've designed reusable module patterns and set standards others follow

  • Experience building and owning observability stacks (Datadog, Grafana, or similar) at an organizational level — not just using them

  • Strong familiarity with CI/CD systems — Jenkins, GitHub Actions, or equivalent — including pipeline design and developer experience ownership

  • Kubernetes at an operational and architectural level

  • A track record of identifying systemic risks and driving improvements that span team or organizational boundaries

  • Security‑first mindset: threat modeling, blast radius analysis, least‑privilege by default, audit trails as a design requirement

  • Strong written English; at Staff level, written communication is how you scale your influence across teams

  • Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent practical experience

What sets you apart

You don't wait for problems to be handed to you — you find them, frame them, and drive the solution. You've operated at a scope where your decisions affected multiple teams or systems, and you know how to build consensus and move work forward without direct authority. You've thought seriously about what infrastructure needs to look like when AI agents are generating and shipping code — safe deployment patterns, strong observability, clean rollback — and you want to help define that standard here. Experience with billing, payments, or financial systems infrastructure is a meaningful plus.

What to Expect
First 30 Days

You will ship code in your first week. We run an agent‑first development workflow — infrastructure changes start with a plan, specifications are written before generation, and every change is reviewed before it merges — and onboarding is no exception. You will get hands‑on with the infrastructure supporting Billing Platform services early, shadow on‑call, and build a clear picture of the system before you start making bigger changes. By the end of 30 days you will have shipped real work and know where the most important problems are.

First 90 Days

You will have taken ownership of one or more infrastructure components and delivered an improvement from design to production with measurable impact. You will be an active participant in deployment and reliability discussions, bringing infrastructure constraints and operational requirements into the conversation early — before code is written. You will be a full participant in the on‑call rotation and have begun shaping the team's technical direction.

One Year Outlook

You will be the team's trusted authority on billing infrastructure. You will have driven meaningful improvements to observability, deployment safety, or platform reliability — and your work will be directly visible in the resilience and correctness of systems that handle real financial transactions for millions of Docker users. You will have helped define what AI‑agent‑assisted infrastructure operations look like done right, and that standard will be visible beyond this team.

Docker considers sponsorship on a case‑case basis based on business needs.

Compensation & Equity

United States: $170,350 – $275,550 + equity

______________________________________________________________________

Posting Information
  • Open vacancy: This posting is for an existing open role.

  • AI in hiring: Docker may use AI‑assisted tools during our recruiting process.

  • Interview recordings: Candidates will be invited to opt in to interview recordings to support interviewer calibration and consistent evaluations. Recordings are optional and require explicit consent.

______________________________________________________________________

Perks & Benefits
  • Remote-first by design – Work from your home, with offices in Seattle and Paris for connection and collaboration.

  • Flexibility that fits your life – We trust you to manage your schedule while delivering great work.

  • Time to recharge – Generous PTO, designated quarterly Whaleness Days, and a designated end‑of‑year Whaleness break.

  • Home office support – Set up your workspace for comfort and success.

  • Technology stipend – Equivalent to US$100 net per month to help support your work.

  • Learning & development – Annual stipend for conferences, courses, certifications, and continued learning.

  • Parental leave – 16 weeks of paid parental leave after six months of employment.

  • Equity for all full‑time employees – Share in Docker's long‑term success as we continue to grow.

  • Comprehensive benefits – Medical, retirement, and paid holidays vary by country.

  • Docker swag – Because representing the whale never gets old.

Docker is proud to be an equal opportunity employer. We are committed to building a team that reflects a broad range of backgrounds, experiences, and perspectives. We believe diverse teams build better products, make better decisions, and better serve our global community.

#LI-REMOTE

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Software Engineer, Billing
Software Engineer, Billing

Docker, Inc. • Seattle (WA)

On-site
USD 120,000 - 160,000
Freedom & flexibility
Home office setup
16 weeks of paid parental leave
+3
Software Engineer, Billing
Software Engineer, Billing

Docker • Seattle (WA)

On-site
USD 120,000 - 150,000
Freedom & flexibility
Quarterly Whaleness Days
Home office setup stipend
+7
Senior Software Engineer, Billing
Senior Software Engineer, Billing

Docker, Inc • Seattle (WA)

On-site
USD 136,750 - 222,750
Home office setup
16 weeks of paid parental leave
Technology stipend
+3
Staff Sales Engineer, Strategic Accounts (EMEA)
Staff Sales Engineer, Strategic Accounts (EMEA)

Docker • Seattle (WA)

Hybrid
GBP 150,000 - 210,000
Remote-first by design
Time to recharge
Home office support
+6
Senior Solutions Engineer, Mid-Enterprise
Senior Solutions Engineer, Mid-Enterprise

Docker • United States

Remote
USD 171,000 - 244,000
Remote-first
Flexible schedule
PTO
+7
Senior Principal Software Engineer, Networking
Senior Principal Software Engineer, Networking

Docker • United States

Remote
USD 219,000 - 352,000
Remote‑first by design
Technology stipend
Learning & development stipend
+4
Principal Solutions Architect, Professional Services
Principal Solutions Architect, Professional Services

Docker • Seattle (WA)

On-site
GBP 102,000 - 146,000
Remote-first
Flexible schedule
Generous PTO
+7
Manager, Engineering, Secure Build and SCS Services (East Coast Preferred)
Manager, Engineering, Secure Build and SCS Services (East Coast Preferred)

Docker, Inc. • Northern (KY)

On-site
USD 175,000 - 283,000
Remote-first
Flexible schedule
Generous PTO
+7
Senior Software Engineer, Infrastructure
Senior Software Engineer, Infrastructure

Docker, Inc • United States

On-site
CAD 115,210 - 197,841
Freedom & flexibility
Quarterly Whaleness Days
Home office setup
+7
Senior Principal Engineer, Infrastructure
Senior Principal Engineer, Infrastructure

Docker, Inc. • Seattle (WA)

On-site
USD 180,000 - 220,000
Freedom & flexibility; fit your work around your life
16 weeks of paid Parental leave
Technology stipend equivalent to $100 net/month
+3