Staff Software Engineer, Infrastructure

Docker

United States

Remote

USD 170,000 - 276,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Freedom & flexibility
Designated quarterly Whaleness Days
Home office setup
16 weeks of parental leave
Technology stipend
Equity
Remote-first culture

Job summary

Docker is seeking a Staff Engineer to set technical direction and lead production adoption for a global, multi-region platform. You’ll stay hands-on while guiding cross‑team architecture and delivering self‑service provisioning and secure defaults.

This role emphasizes reliability, scalable infrastructure, and measurable impact, with opportunities to shape RFCs, design APIs in Go, and drive deployment pipelines using Terraform and GitOps tools, all in a remote‑first, globally distributed team.

Qualifications

  • 8+ years of professional, hands‑on software engineering experience in backend, infrastructure, or platform engineering.
  • Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent practical experience
  • Strong software engineering in Go or a similar language: design, testing, debugging, review, long-term maintainability
  • A track record designing, shipping, and operating cloud services or infrastructure platforms in production
  • Deep expertise in at least one of: Kubernetes, networking, cloud platforms, reliability engineering, or developer platforms
  • Experience setting technical direction and leading work that needs cross‑team alignment
  • Clear written and verbal communication in a remote environment (RFCs, design docs, incident writeups)

Responsibilities

  • Take ambiguous infrastructure problems and turn them into proposals the org can rally around, then drive them through RFCs and architecture reviews across teams
  • Design self-service capabilities and platform APIs (primarily in Go) for onboarding, provisioning, deployment, observability defaults, and day-2 operations, with contracts and docs teams actually use
  • Set delivery standards using Terraform, GitOps with Argo CD, progressive rollout, and good testing, including building the continuous-deployment flow we’re missing today
  • Evolve the multi-tenant EKS foundations toward better reliability, security, scale, and cost: Envoy Gateway ingress, traffic routing, and the multi-region, cross-account connectivity we need
  • Improve SLOs, alerting, and incident follow-up on Grafana Cloud so production gets safer and less dependent on heroics

Skills

8+ years experience
Go
Kubernetes
Networking
Cloud platforms
Reliability engineering
Developer platforms
Linux

Education

Bachelor's degree in Computer Science, Engineering, or related field

Tools

Terraform
GitOps with Argo CD
Envoy Gateway

Job description

Docker has been one of the most loved brands in developer tooling, trusted by more than 20 million monthly users and over 20 billion container image pulls. From solo founders to the world’s largest companies, developers rely on Docker to build, share, and run their applications across our suite of products including Docker Desktop, Docker Hub, and Docker Scout.

We are a globally distributed, remote-first team building the tools that define how software gets built and delivered. As AI agents redefine software development, Docker is at the center of that shift, providing the sandboxed environments, verified images, and secure infrastructure that make autonomous workflows trustworthy by default.

Docker is shipping a wave of new products this year, with R&D initiatives likely to lead to more, and we’re investing heavily in the platform underneath all of it. That platform supports hundreds of engineers across many development teams and carries high-scale production traffic and data transfer every day. It has grown faster than its foundations, and this year is about closing that gap.

Today, much of that work still leans on a handful of experts unblocking the same provisioning and operational workflows by hand. The top priority for this role is moving that work from expert-driven support to paved roads : self-service systems with clear ownership, safe defaults, strong guardrails, and adoption we can measure. The goal is a platform teams trust enough to stop thinking about it, one that just works, so they can focus on their own products instead of ours.

The concrete version sits on this year’s roadmap: spinning up a new global region or application environment should take hours, not days. Right now it takes days. Getting there means building the foundations underneath it. We need a real multi-region, cross-account network architecture and a testing and continuous-deployment flow teams can trust, then a self-service layer on top.

We’re the container company building our own internal platform, so the bar for “the easy path is also the safe path” is high. You’d be joining a team of four, growing to seven this year (this is one of those hires), and we’re looking for a Staff engineer to set technical direction and lead it through real production adoption.

Responsibilities

This is a Staff-level role, so success is measured by leverage rather than just your own commits. On a team this size you’ll stay hands-on in the codebase while also setting direction, aligning teams on pragmatic standards, and carrying platform investments through to adoption. Concretely, you will:

  • Take ambiguous infrastructure problems and turn them into proposals the org can rally around, then drive them through RFCs and architecture reviews across teams.

  • Design self-service capabilities and platform APIs (primarily in Go ) for onboarding, provisioning, deployment, observability defaults, and day-2 operations, with contracts and docs teams actually use.

  • Set delivery standards using Terraform , GitOps with Argo CD , progressive rollout, and good testing, including building the continuous-deployment flow we’re missing today.

  • Evolve the multi-tenant EKS foundations toward better reliability, security, scale, and cost: Envoy Gateway ingress, traffic routing, and the multi-region, cross-account connectivity we need.

  • Improve SLOs, alerting, and incident follow-up on Grafana Cloud so production gets safer and less dependent on heroics.

We judge this work by outcomes the consuming teams feel: how fast they can provision and ship, how much they can do without us, and how reliably it all runs.

AI-assisted operations

We’re actively investing in AI-assisted and agentic workflows to cut operational toil. We care that they stay safe, auditable, and human-reviewed. You’ll help shape where these earn their place and where they don’t. Early targets include:

  • Alert enrichment and incident context-gathering : assembling the relevant signals, history, and runbook so the on-call engineer starts with context instead of a blank page.

  • Runbook-assisted diagnosis and remediation recommendations , with a human in the loop on anything that changes production.

  • Onboarding and readiness assistants that answer the questions our experts answer today.

If you’ve built operational automation and have a healthy skepticism about where automation belongs, this is a place to put both to work.

On-call

Operational ownership is part of the job. You’ll join the rotation after onboarding and shadowing. As a Staff engineer, you’ll also improve the health of on-call itself, with better alerts, stronger runbooks, less toil, and blameless postmortems aimed at prevention.

Qualifications
  • 8+ years of professional, hands‑on, full‑time software engineering experience in backend, infrastructure, or platform engineering.

  • Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent practical experience

  • Strong software engineering in Go or a similar language: design, testing, debugging, review, long-term maintainability.

  • A track record designing, shipping, and operating cloud services or infrastructure platforms in production. We hire for skill and impact, not years.

  • Deep expertise in at least one of: Kubernetes, networking, cloud platforms, reliability engineering, or developer platforms, plus solid Linux, networking, and production‑ops fundamentals.

  • Experience setting technical direction and leading work that needs cross‑team alignment.

  • Clear written and verbal communication in a remote environment (RFCs, design docs, incident writeups).

Nice to have: EKS and ingress/CNI/service‑mesh experience; observability with OpenTelemetry/Prometheus/Grafana; CI/CD and progressive delivery (GitHub Actions, Argo CD, canaries); experience leading migrations or adoption programs across teams.

You don’t need every item here. We value deep expertise in one area, strong systems judgment, and curiosity across the rest.

What to Expect
First 30 days
  • Build context, meet partner teams, ship your first change, shadow on-call.
First 90 days
  • Own a strategic platform problem with a clear plan and metrics; lead an improvement from design to production.
One Year Outlook
  • Lead a major cross‑team initiative (for example, self‑service provisioning of new regions and environments, or the multi‑region networking and CD foundations behind it) and establish durable patterns that change how Docker engineers build and operate services.

Docker considers visa sponsorship on a case‑by‑case basis based on business needs.

Compensation & Equity

Canada: CA$238,250 – CA$382,250 + equity

United States: $170,350 – $275,550 + equity

Perks

  • Freedom & flexibility; fit your work around your life

  • Designated quarterly Whaleness Days plus end of year Whaleness break

  • Home office setup; we want you comfortable while you work

  • 16 weeks of paid Parental leave (after 6 months of employment)

  • Technology stipend equivalent to $100 USD net/month

  • PTO plan that encourages you to take time to do the things you enjoy

  • Training stipend for conferences, courses and classes

  • Equity; we are a growing start‑up and want all employees to have a share in the success of the company

  • Docker Swag

  • Medical benefits, retirement and holidays vary by country

  • Remote‑first culture, with offices in Seattle and Paris

Docker embraces diversity and equal opportunity. We are committed to building a team that represents a variety of backgrounds, perspectives, and skills. The more inclusive we are, the better our company will be.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Software Engineer, Infrastructure
Senior Software Engineer, Infrastructure

Docker, Inc • United States

On-site
CAD 115,210 - 197,841
Freedom & flexibility
Quarterly Whaleness Days
Home office setup
+7
Manager, Engineering, Secure Build and SCS Services (East Coast Preferred)
Manager, Engineering, Secure Build and SCS Services (East Coast Preferred)

Docker, Inc • Seattle (WA)

On-site
GBP 120,000 - 180,000
Remote‑first by design
Flexibility to manage schedule
Generous PTO
+6
Staff Software Engineer, Developer Experience
Staff Software Engineer, Developer Experience

Docker, Inc. • England (AR), Northern (KY)

On-site
USD 110,000 - 180,000
Remote-first culture
Equity
Office locations Seattle and Paris
Principal Software Engineer, Docker Hardened Images
Principal Software Engineer, Docker Hardened Images

Docker • United States

Remote
USD 181,000 - 292,000
Freedom & flexibility
Whaleness Days
Home office setup
+4
Senior Principal Engineer, Central Ops & Systems (Seattle Preferred)
Senior Principal Engineer, Central Ops & Systems (Seattle Preferred)

Docker • United States

Remote
USD 219,000 - 352,000
Remote-first by design
Flexibility that fits your life
Time to recharge
+2
Manager, Engineering, Secure Build and SCS Services (East Coast Preferred)
Manager, Engineering, Secure Build and SCS Services (East Coast Preferred)

Docker, Inc. • Northern (KY)

On-site
USD 175,000 - 283,000
Remote-first
Flexible schedule
Generous PTO
+7
Staff Software Engineer, Cloud Sandboxes (Seattle or SF/Bay Area)
Staff Software Engineer, Cloud Sandboxes (Seattle or SF/Bay Area)

Docker • United States

Remote
USD 170,000 - 276,000
Freedom & flexibility
Whaleness Days & year-end break
Home office setup
+1
Staff Software Engineer, Cloud Sandboxes (Seattle or SF Bay Area)
Staff Software Engineer, Cloud Sandboxes (Seattle or SF Bay Area)

Docker • San Francisco (CA), Seattle (WA)

On-site
USD 170,350 - 275,550
Freedom & flexibility
Whaleness Days
Home office setup
+5
Staff Sales Engineer, Strategic Accounts (EMEA)
Staff Sales Engineer, Strategic Accounts (EMEA)

Docker • Seattle (WA)

Hybrid
GBP 150,000 - 210,000
Remote-first by design
Time to recharge
Home office support
+6
Senior Solutions Engineer, Mid-Enterprise
Senior Solutions Engineer, Mid-Enterprise

Docker, Inc • United States

On-site
USD 170,000 - 244,000
Remote-first
Generous PTO
Whaleness Days
+7