Site Reliability Engineer – Low-Latency Trading Systems — Ondo Finance

The Bitcoin Street Journal

Northern (KY)

Hybrid

USD 140,000 - 210,000

Full time

10 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Remote-first team
Competitive compensation
Full benefits
Flexible vacation policy

Job summary

Ondo Finance seeks an experienced SRE to own reliability and observability of a 24/7 real-time trading platform built with Rust and Go, deployed on a multi-region Kubernetes cluster on AWS. You will read and modify Go and Rust code, debug latency regressions, run on-call during market hours, and build automation to keep the system healthy with a small team.

This is a remote-first role with competitive compensation, full benefits, and opportunities to shape observability, deployment safety, and

Qualifications

  • 5+ years in SRE, production engineering, or infrastructure roles for latency-sensitive systems.
  • Strong programming ability in Go or Rust, and willingness to work in both; this role changes application code.
  • Deep, hands-on Kubernetes and AWS experience: stateful, latency-sensitive workloads in production.
  • Strong observability instincts: fluent PromQL, structured-log analysis, high-signal alerts.
  • Solid Linux internals and networking fundamentals: chase p99 regressions through kernel, NIC, or GC.
  • Sound judgment under pressure and clear written communication during incidents.

Responsibilities

  • Debug production incidents end to end: stale data feeds, rate limits, WebSocket disconnects, latency regressions in trading path.
  • Harden market data ingestion from providers and venue feeds, including staleness detection, failover, and replay.
  • Build reconciliation and data-integrity tooling across live gauges, Postgres, and data lake to ensure positions and PnL always agree.
  • Participate in an on-call rotation covering US equity market hours and 24/7 crypto venues.

Skills

Go
Rust
Kubernetes
AWS
PromQL
Linux networking

Tools

GitOps (Flux)
SOPS secrets
Datadog

Job description

  • Location: United States
  • Sector: Blockchain
  • Source: web3.career
About the company

Hi, we’re Ondo Finance. Our mission is to provide institutional-grade, blockchain-enabled investment products and services. We have both a technology arm that develops decentralized finance technology, and an asset management arm that creates and manages tokenized funds. We are the global leader in tokenized treasuries, tokenized stocks and ETFs, and are building the future of institutional-grade financial services onchain.

Founded by folks from Goldman Sachs Digital Assets Team, we’re backed by some of the best investors in the world including Founders Fund, Coinbase Ventures, Pantera Capital, Tiger Global, and more. We are currently the leaders in the space in terms of AUM and are well capitalized to continue growing the firm. We’re fully remote, with team members across the U.S.

About the role

Ondo operates real-time trading systems that run around the clock across traditional and crypto venues. The platform spans low-latency Rust engines, a fleet of Go services for trading, execution, and PnL accounting, and a multi-region Kubernetes footprint on AWS.

We are looking for an SRE with strong systems programming skills to own the reliability, observability, and performance of this platform. This is a hands-on role: you will read and modify Go and Rust code, debug latency regressions down to the feed handler, run incident response during market hours, and build the automation that keeps a 24/7 trading system healthy with a small team.

Target outcomes
  • Own production reliability for real-time trading services: trading engines, execution gateways, market data ingestion, and PnL/reconciliation pipelines
  • Operate and evolve our multi-region Kubernetes clusters on AWS (EKS), deployed via GitOps (Flux) with SOPS-encrypted secrets
  • Build and refine observability: Prometheus metrics and alerting, Datadog logs and dashboards, and the SLOs that catch degradation before it costs money
  • Improve deploy safety: progressive rollouts, config-reload behavior, and guardrails that prevent a bad push from touching live trading
Responsibilities
  • Debug production incidents end to end: stale market data feeds, exchange rate limits, WebSocket disconnects, order-lifecycle desyncs, and latency regressions in the trading path
  • Harden market data ingestion from providers such as Databento and venue-native feeds (REST and WebSocket), including staleness detection, failover, and replay
  • Build reconciliation and data-integrity tooling across live gauges, Postgres, and our S3 parquet data lake, so positions, fills, and PnL always agree
  • Participate in an on-call rotation covering US equity market hours and 24/7 crypto venues
Requirements
  • 5+ years in SRE, production engineering, or infrastructure roles, with meaningful time supporting real-time or latency-sensitive systems
  • Strong programming ability in Go or Rust, and willingness to work in both; this role changes application code, not just infrastructure
  • Deep, hands-on Kubernetes and AWS experience: you have run stateful, latency-sensitive workloads in production, not just stateless web services
  • Strong observability instincts: fluent PromQL, structured-log analysis, and experience designing alerts with high signal and low noise
  • Solid Linux internals and networking fundamentals: you can chase a p99 regression through the kernel, the NIC, or the GC
  • Sound judgment under pressure and clear written communication during and after incidents
Nice to haves
  • Experience operating trading systems, execution infrastructure, or market data infrastructure at a trading firm, exchange, or broker
  • Familiarity with market microstructure and order lifecycle (order books, order types, fills and reconciliation)
  • Experience with market data providers and protocols (Databento, SIP/prop equity feeds, venue WebSocket APIs)
  • Exposure to crypto venues and on-chain trading
  • Python for operational tooling and data analysis (pandas, parquet, BigQuery)
  • Experience with GitOps workflows, infrastructure as code, and secrets management at scale
What we offer
  • Competitive compensation including but not limited to salary, future token rights, and/or equity (according to your preferences) - We are well-funded and believe that great talent deserves great compensation.
  • Full benefits (medical, vision, and dental) and flexible vacation policy (PTO).
  • Remote-first team across many countries - You will be an early team member helping shape our vision, culture, and design practices.
  • A+ colleagues - Our team includes alumni from: Goldman Sachs, Blackrock, Two Sigma, Bridgewater, SpaceX, AWS, Meta, Google, McKinsey, Coinbase, Circle, Uniswap.
  • Best-in-class investors - We are proud to be backed by leading crypto experts and VCs, including Pantera Capital, Founders Fund, and Coinbase Ventures.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer - Low-Latency Trading Systems New Remote (US)
Site Reliability Engineer - Low-Latency Trading Systems New Remote (US)

Ondo • Northern (KY)

Remote
USD 180,000 - 260,000
Remote-first team
Competitive compensation
Full benefits
Senior Backend Engineer (Trading Infrastructure)
Senior Backend Engineer (Trading Infrastructure)

SOLANA FOUNDATION • Northern (KY)

On-site
USD 120,000 - 210,000
Competitive pay
Remote-first team
Benefits package
+2
Senior Backend Engineer (Trading Infrastructure) — Ondo Finance
Senior Backend Engineer (Trading Infrastructure) — Ondo Finance

The Bitcoin Street Journal • Northern (KY)

Hybrid
USD 140,000 - 230,000
Salary + token rights
Full benefits (medical, vision, and/or
Remote-first team
+1
Remote SRE — Real-Time Low-Latency Trading Systems
Remote SRE — Real-Time Low-Latency Trading Systems

SOLANA FOUNDATION • United States

On-site
USD 150,000 - 210,000
Remote-first team
Healthcare benefits
Flexible PTO
+3
Real-Time SRE: Low-Latency Trading Systems (Remote)
Real-Time SRE: Low-Latency Trading Systems (Remote)

Ondo • Northern (KY)

Remote
USD 180,000 - 260,000
Remote-first team
Competitive compensation
Full benefits
Senior Backend Engineer (Trading Infrastructure)
Senior Backend Engineer (Trading Infrastructure)

Far Coder • Northern (KY)

Hybrid
USD 150,000 - 230,000
Competitive salary
Full benefits
Remote-friendly team
+1
Low-Latency Trading SRE — Real-Time Go/Rust, Kubernetes
Low-Latency Trading SRE — Real-Time Go/Rust, Kubernetes

The Bitcoin Street Journal • Northern (KY)

Hybrid
USD 140,000 - 210,000
Remote-first team
Competitive compensation
Full benefits
+1
Infrastructure Engineer (DevOps) — Ondo Finance
Infrastructure Engineer (DevOps) — Ondo Finance

The Bitcoin Street Journal • Northern (KY)

Hybrid
USD 120,000 - 180,000
Competitive compensation
Full benefits
Flexible PTO
+2
Infrastructure Engineer — Ondo Finance
Infrastructure Engineer — Ondo Finance

The Bitcoin Street Journal • Northern (KY)

Hybrid
USD 120,000 - 170,000
Compensation & token rights
Full benefits (medical, vision, dental
PTO
+3
Lead Site Reliability Engineer (SRE)
Lead Site Reliability Engineer (SRE)

Optimal Market Technologies • Chicago (IL)

Hybrid
USD 175,000 - 200,000