Staff Software Engineer, Replication Foundations

Engg

United States

On-site

USD 210,000 - 260,000

Full time

7 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Temporal Technologies is seeking a Staff Software Engineer to lead the Replication Foundations team within Temporal’s Cloud Global Services (CGS) organization. You will shape the technical direction of Temporal’s OSS replication stack, spanning architecture, protocols, rollout, and operations.

You will mentor engineers, drive reliability and scalability initiatives across multi-cluster setups, and partner with product and cloud teams to align OSS with Temporal Cloud needs.

Qualifications

  • Experience designing production distributed systems with high availability and reliability.
  • Ability to define architecture, invariants, and trade-offs in evolving problems.
  • Proficient in debugging complex production issues, including concurrency bugs.

Responsibilities

  • Set the technical direction for OSS replication stack from design to rollout.
  • Lead design reviews and mentor engineers across teams.
  • Collaborate with CGS, Product, and Cloud teams to align needs.
  • Triage production incidents and drive follow-up improvements.
  • Identify risks and shape the replication roadmap for cloud products.
  • Mentor engineers, lead design reviews, and improve testing practices.

Skills

Distributed systems
Go programming
Concurrency
Performance optimization
System architecture
Communication
Mentoring

Tools

Go
C++
Java

Job description

ROLE SUMMARY

We’re hiring a Staff Software Engineer to join the Replication Foundations team within Temporal’s Cloud Global Services (CGS) organization. Replication Foundations owns and evolves Temporal’s core replication stack in Temporal OSS—the distributed systems backbone behind key Temporal Cloud capabilities such as High Availability namespaces, cross-cluster and cross-region failover, and migration products that enable customers to move workloads between self-hosted Temporal and Temporal Cloud. The team also builds foundational scalability and reliability mechanisms that support Temporal at scale. In this role, you’ll help set the technical direction for Temporal’s distributed replication systems. You’ll lead complex, correctness-critical initiatives spanning architecture, design, implementation, rollout, and operations. You’ll work across teams to evolve reliable and scalable replication capabilities that support both the open source project and Temporal Cloud.

WHAT YOU'LL DO
  • Set the technical direction and evolve the architecture of Temporal’s OSS replication stack, from problem definition through rollout and operational support.
  • Lead the design and implementation of replication protocols that power:
    • High Availability namespaces
    • Cross-cluster and cross-region replication
    • Migration between Temporal clusters, including cloud-to-self-hosted and cloud-to-cloud scenarios
  • Drive scalability and reliability initiatives such as:
    • Multi-cell namespaces
    • Enabling a namespace to span multiple clusters
    • Improving load distribution and handling hot spots
  • Define and communicate system-level guarantees, including consistency models, ordering, idempotency, failure recovery, performance, and operational behavior.
  • Identify architectural risks and opportunities, and shape the technical roadmap for replication capabilities that support current and future cloud products.
  • Partner with Cloud Enablement, CGS, Product, and other engineering teams to align OSS replication foundations with customer and product needs.
  • Lead design reviews, raise the quality of implementation and testing practices, mentor engineers, and provide technical guidance across the organization.
  • Lead or contribute to debugging complex production issues, incident response, and follow-up improvements related to replication and core system behavior.
WHAT YOU'LL BRING
  • A track record of designing and delivering complex production distributed systems, including systems where correctness, availability, and performance are critical.
  • Deep understanding of distributed systems fundamentals such as replication, partitioning, consistency, fault tolerance, durability, concurrency, and failure recovery.
  • Experience defining system architecture, protocol behavior, invariants, and trade-offs in ambiguous or evolving problem spaces.
  • Experience debugging complex production issues, including concurrency bugs, data inconsistencies, partial failures, and performance bottlenecks.
  • Proficiency writing production-quality concurrent code in Go; experience with Java, C++, or similar systems languages is also welcome.
  • Strong written and verbal communication skills, including the ability to explain complex designs and trade-offs to both technical and cross-functional audiences.
  • Demonstrated ability to influence technical direction across teams, build alignment without direct authority, and mentor engineers.
  • A thoughtful and curious approach to understanding how systems behave under load, failure, and changing workload conditions.
NICE TO HAVE
  • Experience designing or maintaining replication protocols or data-plane infrastructure.
  • Experience with multi-cluster or multi-region architectures, including active-active or active-passive systems.
  • Familiarity with database internals, log-based replication, or event-sourced systems.
  • Prior contributions to large open source projects or distributed systems infrastructure.

Temporal Technologies is an Equal Opportunity Employer. Temporal Technologies does not discriminate on the basis of race, religion, color, sex, gender identity, sexual orientation, age, non-disqualifying physical or mental disability, national origin, veteran status, or any other basis covered by appropriate law. All employment is decided on the basis of qualifications, merit, and business need. We embrace and celebrate differences and diversity. Temporal is committed to providing access, equal opportunity, and reasonable accommodation for individuals with disabilities in employment, its services, programs, and activities. If you need to request a reasonable accommodation, please let your Recruiter know so we can assist.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff Software Engineer, Replication Foundations
Staff Software Engineer, Replication Foundations

Temporal • Seattle (WA)

On-site
USD 190,000 - 260,000
Staff Software Engineer, Replication Foundations
Staff Software Engineer, Replication Foundations

Engg • United States

Remote
USD 180,000 - 270,000
Staff Engineer: Distributed Replication Systems
Staff Engineer: Distributed Replication Systems

Temporal • Seattle (WA)

On-site
USD 190,000 - 260,000
Staff Software Engineer, Replication Foundations
Staff Software Engineer, Replication Foundations

United States Digital Space LLC • United States

On-site
USD 140,000 - 210,000
Software Engineer II, OSS
Software Engineer II, OSS

Engg • Seattle (WA)

On-site
USD 130,000 - 190,000
Senior Software Engineer, Open Source Server
Senior Software Engineer, Open Source Server

Temporal • United States

On-site
USD 140,000 - 190,000
Software Engineer II, Open Source Server
Software Engineer II, Open Source Server

Engg • Seattle (WA)

On-site
USD 120,000 - 180,000
Software Engineer II, Open Source Server
Software Engineer II, Open Source Server

Temporal • United States

On-site
USD 140,000 - 190,000
Software Engineer II, Open Source Server
Software Engineer II, Open Source Server

Temporal • Seattle (WA), San Francisco (CA)

Hybrid
USD 120,000 - 180,000
Senior Software Engineer, Cloud Platform Foundations
Senior Software Engineer, Cloud Platform Foundations

Engg • Seattle (WA)

On-site
USD 140,000 - 210,000