Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Alexander Chapman is seeking a Staff Distributed Systems Engineer to own and scale critical production infrastructure. You will lead design and operation of high-throughput, multi-region services, ensuring reliability, performance, and data integrity across regions.
The role emphasizes building failover, disaster recovery, and observability capabilities while mentoring engineers and raising the team's technical bar for operating systems at scale.
We are representing a rapidly growing financial technology company building a next-generation trading platform. Our client is developing a consumer-focused platform that combines real-time trading with social discovery, giving users a streamlined way to discover market activity, follow traders, receive real-time insights, and execute trades across a growing range of financial products. Behind the consumer product is a sophisticated, highly distributed backend platform supporting real-time market data, trading activity, social features, financial data, and user-facing services across multiple regions. With significant growth ahead, the engineering organization is expanding its core infrastructure capabilities and is looking for a Staff Distributed Systems Engineer to take ownership of reliability, scalability, and performance across the platform.
This is a hands-on Staff-level position with direct ownership of critical production infrastructure. You will own shared systems including datastores, caches, messaging infrastructure, and regional application services. The focus will be on designing systems that remain predictable and resilient through traffic surges, dependency failures, infrastructure changes, and partial regional outages. You will also lead the development of new failover and disaster-recovery capabilities, including defining recovery objectives and implementing the systems and testing required to safely recover services and data.
This is not a purely architectural Staff position. You will be deeply involved in production systems, making decisions around databases, caching, replication, regional architecture, failure recovery, traffic management, and system performance. The role is particularly well suited to an engineer who has spent years dealing with the realities of distributed systems: overloaded services, database contention, hot keys, failed deployments, network failures, dependency degradation, regional outages, and unpredictable traffic. It is an opportunity to have broad technical ownership over the distributed systems layer of a rapidly scaling financial platform, while helping shape the architecture and engineering practices as the platform grows.