Senior Engineer - Production Cloud and Container Services

United States Digital Space LLC

Los Angeles (CA)

Hybrid

USD 181,000 - 217,000

Full time

2 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Medical insurance
Dental insurance
Vision insurance
401(k) with company match
Employee Stock Purchase Program
Flexible vacation policy
Paid sick leave
Paid holidays

Job summary

United States Digital Space LLC is seeking a Senior Engineer to scale and operate our Kubernetes-based platform for control plane services across multi-cloud environments. You will expand platform features, onboard services, and drive efficient, secure operations.

Ideal candidates have 5+ years in distributed infrastructure, experience with CNCF tools, Terraform, and cloud providers like AWS and GCP. This hybrid US role offers remote possibilities with strong travel and on-call responsibilities.

Qualifications

  • Experience building and operating production-grade Kubernetes clusters in multiple regions, clouds or data centers.
  • Experience with infrastructure and configuration management tooling, Terraform used to manage infrastructure.
  • Experience provisioning and managing users and resources with cloud providers such as AWS and GCP.

Responsibilities

  • Design, build, and operate infrastructure (cloud, the company datacenter) to enable reliable and rapid deployment.
  • Diagnose and resolve performance and reliability issues across the stack.
  • Deploy and support complex 3rd party and internally developed applications.
  • Write tools to automate maintenance and deployment of servers, services, and applications.
  • Collaborate with internal users and evolve the platform and its operations.
  • Drive projects across time zones.

Skills

Go
Python
Kubernetes
Terraform
Linux
CI/CD
GitOps
Prometheus
Datadog
Grafana
AWS
GCP
multi-cloud
on-call

Tools

Kubernetes
Terraform
GitHub
Prometheus
Flux
Helm
Docker
EC2
S3
GKE

Job description

the company helps people stay better connected with the things they love. the company’s edge cloud platform enables customers to create great digital experiences quickly, securely, and reliably by processing, serving, and securing our customers’ applications as close to their end-users as possible — at the edge of the Internet. The platform is designed to take advantage of the modern internet, to be programmable, and to support agile software development. the company’s customers include many of the world’s most prominent companies, including GitHub, Yelp, Paramount, and JetBlue.

We're building a more trustworthy Internet. Come join us.

**Posting Open Date: Sept. 25, 2026Anticipated Posting Close Date*: Oct. 23, 2026Job posting may close early due to the volume of applicants.*

Senior Engineer - Production Cloud and Container Services

Fleet Operations and Production Engineering at the company is looking for a Senior Engineer to join our Cloud and Container Services team. This role is focused on helping to scale and manage the company’s Kubernetes based platform for control plane services. This platform is built on top of multiple public cloud services and contains many Kubernetes ecosystem components. We’re working to scale out our platform to support growth and at the same time evolve to address new business priorities. A successful candidate will help expand our platform feature set, support existing users and onboard new services, and drive efficiency while maintaining a secure platform.

What You'll Do:
  • Design, build, and operate infrastructure (cloud, the company datacenter) to enable reliable and rapid deployment. This includes effective monitoring and resilient operations in a large-scale multi-cloud and hybrid cloud / on-premise environment. The majority of workloads are containerized but some are using native cloud services such as compute and storage.
  • Diagnose and resolve performance and reliability issues across the stack: application, operating system, network, 3rd party services and APIs, including cross-application dependencies.
  • Deploy and support complex 3rd party and internally developed applications.
  • Write tools to automate maintenance and deployment of servers, services, and applications.
  • Collaborate with internal users and continually evolve the platform and its operations using solid engineering practices.
  • Drive projects sometimes independently and sometimes collaboratively across time zones.
  • Configure access and manage operations within a multi-cloud environment.
What We're Looking For:
  • Experience running high availability systems and supporting distributed infrastructure. Most Senior level Engineers at the company have more than 5 years of related experience.
  • You have designed services with fault tolerance and across geographies.
  • You have deployed and managed multi-tiered services.
  • Understanding of Linux systems, high and low level. You have used tcpdump and tracing tools.
  • Experience building and operating production-grade kubernetes clusters in multiple regions, clouds or data centers. You have experience with tooling in the CNCF space, such as: prometheus, flux, helm, etc.
  • Experience with programming languages such as Go and Python. You can read code and reason about what it does. You can write code within an existing large code base such as adding features. You can create medium-sized programs from scratch such as custom kubernetes controllers, custom prometheus exporters, and building tooling to help manage infrastructure.
  • Experience with infrastructure and configuration management tooling. You have used Terraform to manage infrastructure.
  • Experience provisioning and managing users and resources with cloud providers such as AWS and GCP. You have provisioned users and accounts using both graphical user interfaces and infrastructure as code frameworks. You have experience provisioning components on public cloud and understand how they work together in creating a multi-tiered service. You have deployed and managed services built on top of public cloud components such as EC2, S3, and GKE.
  • Experience with CI/CD and GitOps tooling. You can iterate infrastructure via pipelines through code changes. You have experience using Github including creating and reviewing pull requests.
  • Experience with monitoring tools such as Prometheus, Datadog and Grafana.
  • Experience working on a distributed team. You have experience collaborating across timezones. You can articulate challenges of distributed teams and how you mitigate them.
  • Experience leveraging AI tooling to accelerate your work and enable new capabilities. You have used tooling like Claude Code and created skills or agents with demonstrated success in unlocking you, your team, and/or the company to achieve more.
Work Hours:
  • This position will require you to be available during core business hours and occasional nights and weekends as necessary to support on-call coverage.
Work Location(s) & Travel Requirements:

The preferred locations for this position are:

  • San Francisco, CA, USA
  • Denver, CO, USA
  • New York, NY, USA

the company currently embraces a largely hybrid model for most roles which allows employees flexibility to split their time between the office and home. There is a strong preference for Hybrid near a local office. However, we may be willing to consider remote candidates within the US, UK or Canada.

This position may require travel as required by your role or requested by your manager.

Pursuant to the San Francisco Fair Chance Ordinance and the Los Angeles Fair Chance Initiative for Hiring Ordinance, we will consider for employment qualified applicants with arrest and conviction records.

Salary:

The estimated salary range for this position when hired in the US is $181,220 to $217,464.

Starting salary may vary based on permissible, non-discriminatory factors such as experience, skills, qualifications, and location.

This role may be eligible to participate in the company’s equity and discretionary bonus programs.

Benefits:

We care about you. the company works hard to create a positive environment for our employees, and we think your life outside of work is important too. We support our teams with great benefits that start on the first day of your employment with the company. Curious about our offerings?

In the US, we offer a comprehensive benefits package including medical, dental, and vision insurance. Family planning, mental health support along with Employee Assistance Program, Insurance (Life, Disability, and Accident), a Flexible Vacation policy and up to 18 days of accrued paid sick leave are there to help support our employees. We also offer 401(k) (including company match) and an Employee Stock Purchase Program. For 2026, we offer 11 paid local holidays, 12 paid company wellness days.

Why the company?

-

We have a huge impact.

the company is a small company with a big reach. Not only doour customershave a tremendous user base, but we also support a growing number ofopen source projects and initiat

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Engineer - Network Management Platform
Senior Engineer - Network Management Platform

United States Digital Space LLC • Los Angeles (CA)

Hybrid
USD 181,000 - 217,000
Equity
Discretionary bonus
Comprehensive benefits
Senior Engineer - Production Cloud and Container Services
Senior Engineer - Production Cloud and Container Services

Fastly, Inc. • San Francisco (CA)

Hybrid
USD 181,000 - 217,000
Medical, dental, vision insurance
401(k) with company match
Employee Stock Purchase Program
+2
Senior Engineer – Production Cloud and Container Services
Senior Engineer – Production Cloud and Container Services

WebHosting • Denver (CO), Northern (KY)

Hybrid
USD 181,000 - 217,000
401(k) match
Employee Stock Purchase Program
Paid time off
+1
Senior Manager, Engineering - Platform Security
Senior Manager, Engineering - Platform Security

United States Digital Space LLC • Los Angeles (CA)

On-site
USD 228,000 - 274,000
Staff Engineer - Control Systems
Staff Engineer - Control Systems

United States Digital Space LLC • Los Angeles (CA)

Hybrid
USD 211,000 - 254,000
Medical Insurance
Dental Insurance
Vision Insurance
+3
Senior Program Manager, Network Infrastructure
Senior Program Manager, Network Infrastructure

United States Digital Space LLC • Los Angeles (CA)

On-site
USD 143,000 - 195,000
Medical, dental, and vision insurance
401(k) with company match
Employee Stock Purchase Program
+2
Founding Senior Software Engineer (Infrastructure)
Founding Senior Software Engineer (Infrastructure)

Stealth Startup • San Carlos (CA)

Hybrid
USD 215,000 - 290,000
Unlimited breakfast, lunch, dinner, &
Gym and daily commute fee
Internet and phone bill covered
Senior Engineer - Security Products (Backend APIs)
Senior Engineer - Security Products (Backend APIs)

United States Digital Space LLC • Denver (CO)

Hybrid
USD 181,000 - 217,000
Medical coverage
Dental coverage
Vision coverage
+2
Senior Backend Software Engineer (APIs)
Senior Backend Software Engineer (APIs)

United States Digital Space LLC • Los Angeles (CA)

Hybrid
USD 181,000 - 217,000
Medical insurance
Dental insurance
Vision insurance
+4
Senior DataCenter Provisioning Engineer
Senior DataCenter Provisioning Engineer

United States Digital Space LLC • United States

Hybrid
USD 122,000 - 173,000