Staff Distributed Systems Engineer — Real-Time Infra

Alexander Chapman

New York (NY)

On-site

USD 180,000 - 240,000

Full time

11 days ago
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Alexander Chapman is seeking a Staff Distributed Systems Engineer to own and scale critical production infrastructure. You will lead design and operation of high-throughput, multi-region services, ensuring reliability, performance, and data integrity across regions.

The role emphasizes building failover, disaster recovery, and observability capabilities while mentoring engineers and raising the team's technical bar for operating systems at scale.

Qualifications

  • 10+ years in backend, platform, or distributed systems engineering.
  • Deep experience with high-throughput, distributed production systems.
  • Strong PostgreSQL performance, replication, and failure mode knowledge.
  • Experience with AWS, Terraform, and reliability-focused architectures.

Responsibilities

  • Design and operate high-throughput, multi-region backend services.
  • Own reliability, scalability, and performance of shared infrastructure.
  • Improve datastore and cache performance, replication, and failure handling.
  • Implement backpressure, rate limiting, circuit breakers, and bounded retries.
  • Reduce cross-region latency and improve data locality.
  • Test failover procedures and disaster recovery strategies.
  • Define and validate RTO/RPO objectives for critical services.
  • Enhance observability and operational tooling across the platform.
  • Mentor engineers and raise the team's reliability bar.

Skills

Distributed systems
Backend engineering
Go / TypeScript
Observability
Failover / DR design
Performance tuning
Mentoring

Tools

Terraform
PostgreSQL
Redis
AWS (ECS/RDS/ElastiCache)
Datadog APM
NATS JetStream / Kafka

Job description

Alexander Chapman is seeking a Staff Distributed Systems Engineer to own and scale critical production infrastructure. You will lead design and operation of high-throughput, multi-region services, ensuring reliability, performance, and data integrity across regions.

The role emphasizes building failover, disaster recovery, and observability capabilities while mentoring engineers and raising the team's technical bar for operating systems at scale.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Backend Engineer — Real-Time, Scalable Systems
Backend Engineer — Real-Time, Scalable Systems

Alexander Chapman • California (MO)

On-site
USD 140,000 - 190,000
Staff Distributed Systems Engineer
Staff Distributed Systems Engineer

Alexander Chapman • New York (NY)

On-site
USD 180,000 - 240,000
Distributed Systems Engineer – Scale & Reliability
Distributed Systems Engineer – Scale & Reliability

Joinimagine • San Francisco (CA)

On-site
USD 180,000 - 220,000
Staff Distributed Systems Engineer - Scale & Resilience
Staff Distributed Systems Engineer - Scale & Resilience

SOLANA FOUNDATION • New York (NY), Northern (KY)

Hybrid
USD 180,000 - 270,000
Staff Systems Engineer: Low-Level, Distributed Infra
Staff Systems Engineer: Low-Level, Distributed Infra

Staffworx • New York (NY)

On-site
USD 280,000 - 380,000
Relocation assistance
Visa transfers considered
On-site in New York City
Senior Distributed Systems Engineer - Multi-Region Scale
Senior Distributed Systems Engineer - Multi-Region Scale

FOMO Labs Inc. • New York (NY), Northern (KY)

Hybrid
USD 190,000 - 270,000
Staff Engineer, Distributed Systems (Remote, Flex Hours)
Staff Engineer, Distributed Systems (Remote, Flex Hours)

BairesDev • Peru (IL)

On-site
USD 140,000 - 190,000
Remote work 100%
USD or local currency compensation
Home office setup
+4
Remote Senior Systems Engineer, High-Performance Linux
Remote Senior Systems Engineer, High-Performance Linux

Andromeda • San Francisco (CA)

Hybrid
USD 190,000 - 260,000
Infrastructure Engineer, Distributed Systems & Reliability
Infrastructure Engineer, Distributed Systems & Reliability

Perplexity • Seattle (WA)

On-site
USD 220,000 - 405,000
Staff Elixir Engineer: Lead High-Throughput Systems
Staff Elixir Engineer: Lead High-Throughput Systems

Stord • Atlanta (GA)

On-site
USD 180,000 - 240,000