Senior Site Reliability Engineer

United States Digital Space LLC

Bellevue (CA)

On-site

USD 180,000 - 230,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Amazing Benefits
Making Social Impact
Fostering Diversity, Equity, Inclusion

Job summary

United States Digital Space LLC is seeking a Senior Site Reliability Engineer to help build, improve, and maintain our cloud platform services. You will design scalable cloud environments, enforce security policies, and enable corporate engineering teams to operate securely at scale.

Reporting to the Manager, Site Reliability Engineering, the role emphasizes automation, testing, and operational excellence with experience in AWS multi-account governance, Terraform, Python, and Kubernetes.

Qualifications

  • 5+ years of experience in SRE, DevOps, or Systems Engineering roles.
  • Expert in building AWS multi-account environments (AWS Orgs, IAM, Identity Center, StackSets).
  • Infrastructure as code (Terraform) and secure automation tools in Python; Git-based CI/CD workflows.
  • Container orchestration (Kubernetes) and observability tools (Splunk, CloudWatch, Grafana).
  • Security/compliance in highly secure, regulated environments (FedRAMP) and access to federal data.
  • Networking & AWS infrastructure: VPCs, TGWs, VPC endpoints.
  • Solid Linux system administration knowledge.

Responsibilities

  • Secure Cloud Infrastructure & Pipelines: Design, build, and modernize scalable cloud environments and development tools while strictly enforcing security policies and standards for regulated environments.
  • Cross-Functional Collaboration & Advocacy: Partner with software engineering teams to champion DevOps and SRE best practices, deliver internal customer service, and contribute to Agile workflows.
  • Technical Documentation & Operations: Create and maintain comprehensive technical documentation, including runbooks and disaster recovery procedures to ensure system reliability.

Skills

5+ years of experience in SRE/DevOps

Tools

Terraform
Python
GitLab/GitHub Actions
Kubernetes
Splunk
CloudWatch
Grafana

Job description

**Secure Every Identity, from AI to Human

**Identity is the key to unlocking the potential of AI. the company secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence.

This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk.

The Technology, Data and Intelligence Team Message

the company’s Technology, Data and Intelligence (TDI) team delivers the systems, tools, and services that power internal operations across the company. From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology.

The Senior Site Reliability Engineer Opportunity

Reporting to the Manager, Site Reliability Engineering, this role will help build, improve, and maintain our cloud platform services by designing and implementing complex cloud-based engineering enablement systems. With a strong focus on automation, testing, and operational excellence, you will deliver foundational infrastructure capabilities that enable corporate engineering teams to operate securely, reliably, and at scale.

What you'll be doing
  • Secure Cloud Infrastructure & Pipelines: Design, build, and modernize scalable cloud environments and development tools while strictly enforcing security policies and standards for regulated environments.
  • Cross-Functional Collaboration & Advocacy: Partner with software engineering teams to champion DevOps and SRE best practices, deliver excellent internal customer service, and actively contribute to Agile workflows (e.g., demos, architecture sessions).
  • Technical Documentation & Operations: Create and maintain comprehensive technical documentation, including network diagrams, runbooks, and disaster recovery procedures to ensure system reliability and knowledge sharing.
What you'll bring to the role
  • Professional Experience & Scale: 5+ years of experience in SRE, DevOps, or Systems Engineering roles with a proven track record of delivering complex, large-scale infrastructure projects.
  • AWS Expertise & Centralized Governance: Expert in building and managing AWS multi-account environments (spanning hundreds of accounts), with deep proficiency in authentication, governance, and organization management (AWS Orgs, IAM, Identity Center, StackSets).
  • Automation & CI/CD Pipelines: Highly skilled in infrastructure as code (Terraform), writing secure automation tools in Python, and building Git-based CI/CD workflows (GitLab, GitHub Actions).
  • Containerization & Observability: Strong hands-on experience managing container orchestration environments (Kubernetes) and utilizing monitoring and logging tools (Splunk, CloudWatch, Grafana stack).
Extra credit if you have experience in the following
  • Networking & AWS Infrastructure: Hands-on experience with general networking concepts (BGP and IPsec management) and leveraging core AWS networking services (VPCs, TGWs, and VPC endpoints).
  • System Administration: Solid foundational knowledge and experience in Linux system administration.
  • Security & Compliance: Proven experience operating within highly secure, regulated environments (e.g., FedRAMP), with a strong understanding of FIPS, STIGs, and data boundary implementations.
Additional requirements:
  • Federal Access & Eligibility: This position requires the ability to access federal environments and/or have access to protected federal data. As a condition of employment for this position, the successful candidate must be able to submit documentation establishing U.S. Person status (e.g. a U.S. Citizen, National, Lawful Permanent Resident, Refugee, or Asylee. 22 CFR 120.15) upon hire.
What you can look forward to as an the company employee!
  • Amazing Benefits
  • Making Social Impact
  • Fostering Diversity, Equity, Inclusion and Belonging at the company

the company is an Equal Opportunity Employer/Affirmative Action Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, marital status, age, physical or mental disability, or status as a protected veteran. We also consider for employment qualified applicants with arrest and convictions records, consistent with applicable laws. If reasonable accommodation is needed to participate in the job application or interview process, please use this Form to request an accommodation.

the company is committed to complying with applicable data privacy and security laws and regulations. For more information, please see our Privacy Policy at https://www.the company.com/privacy-policy/.

P14598_3509133

Below is the annual base salary range for candidates located in San Francisco Bay Area. Your actual base salary will depend on factors such as your skills, qualifications, experience, and work location. In addition, the company offers equi

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Site Reliability Engineer (FedRAMP)
Staff Site Reliability Engineer (FedRAMP)

United States Digital Space LLC • Washington

On-site
USD 174,000 - 239,000
Health insurance
Dental insurance
Vision insurance
+2
Manager, Site Reliability Engineering
Manager, Site Reliability Engineering

United States Digital Space LLC • San Francisco (CA)

Hybrid
USD 204,000 - 306,000
Equity
Bonus
Health insurance
+6
Staff TDI Site Reliability Engineer, Okta Federal
Staff TDI Site Reliability Engineer, Okta Federal

United States Digital Space LLC • San Francisco (CA)

Hybrid
USD 174,000 - 239,000
Senior Manager, Site Reliability Engineering - Infrastructure Platform
Senior Manager, Site Reliability Engineering - Infrastructure Platform

United States Digital Space LLC • San Francisco (CA)

Hybrid
USD 232,000 - 319,000
Equity
Bonus
Health insurance
+2
Staff Site Reliability Engineer, Federal (TS/SCI)
Staff Site Reliability Engineer, Federal (TS/SCI)

United States Digital Space LLC • Washington

On-site
USD 170,000 - 260,000
Staff TDI Site Reliability Engineer, Okta Federal
Staff TDI Site Reliability Engineer, Okta Federal

United States Digital Space LLC • Washington

Hybrid
USD 174,000 - 239,000
Equity
Bonus
Health insurance
+5
Senior DevOps/SRE Engineer
Senior DevOps/SRE Engineer

SEI • Chicago (IL)

Hybrid
USD 140,000 - 170,000
Comprehensive healthcare benefits
401(k) match
Paid Time Off (PTO)
+2
Senior DevOps/SRE Engineer
Senior DevOps/SRE Engineer

SEI • Oaks (PA)

Hybrid
USD 140,000 - 170,000
Comprehensive healthcare coverage
401(k) matching
Tuition reimbursement
+1
Senior Manager, Site Reliability Engineering (Federal)
Senior Manager, Site Reliability Engineering (Federal)

United States Digital Space LLC • Washington

On-site
USD 207,000 - 285,000
Health insurance
401(k) plan
Paid leave
Staff Software Engineer, Core Infrastructure
Staff Software Engineer, Core Infrastructure

United States Digital Space LLC • San Francisco (CA)

On-site
USD 194,000 - 267,000
Equity
Bonus
Health insurance
+6