Site Reliability Engineer – Real-Time FinTech Platform

Onyxodds

New York (NY)

Vor Ort

USD 145.000 - 190.000

Vollzeit

Vor 5 Tagen
Sei unter den ersten Bewerbenden
Bewerbungsgenerator

Hebe dich für diese Rolle von der Masse ab — erstelle in etwa einer Minute einen maßgeschneiderten Lebenslauf und ein Anschreiben.

Schaffe es an den ATS-Filtern vorbei

Benefits dieser Stelle

Equity ownership
Fully covered health benefits

Zusammenfassung

Onyx is hiring a Site Reliability Engineer for our NoHo, Manhattan office to build and scale reliable production systems on AWS. You’ll own infrastructure, deployment pipelines, and observability, collaborating with engineers across product and security to keep latency low and uptime high.

You’ll manage Terraform-based infrastructure, strengthen monitoring with Datadog, and lead incident response, post-incident reviews, and runbooks while balancing long-term improvements with immediate

Qualifikationen

  • Experience operating production systems in AWS.
  • Proficient with Terraform and infrastructure as code.
  • Hands-on observability, incident response, and production troubleshooting.
  • Understanding of distributed systems, networking, databases, and cloud architecture.
  • Built automation using Python, TypeScript, Go, Bash, or similar languages.
  • Familiar with deployment pipelines, release strategies, and rollback procedures.
  • Ability to balance long-term infra improvements with immediate production needs.
  • Calm under pressure, ownership from alert to fix.

Aufgaben

  • Build, operate, and improve scalable production infrastructure in AWS.
  • Manage cloud infrastructure through Terraform and IaC.
  • Strengthen monitoring, logging, tracing, alerting, and observability using Datadog.
  • Improve availability, performance, resilience, and security of services.
  • Build tooling and automation to make infrastructure safer and easier to operate.
  • Improve deployment systems, release processes, environment management, and rollback capabilities.
  • Partner with engineers to design reliable, observable services.
  • Monitor production systems and lead incident investigation and resolution.
  • Develop runbooks, escalation processes, and post-incident reviews.
  • Improve database reliability, backups, DR, and capacity planning.
  • Participate in on-call rotation and support during high-volume events.

Kenntnisse

AWS production systems
Terraform
Observability
Incident response
Distributed systems
Networking
Databases
Deployment pipelines
Python/Go/TypeScript
On-call ownership

Tools

Datadog
Kubernetes
Docker
ECS/EKS

Jobbeschreibung

Onyx is hiring a Site Reliability Engineer for our NoHo, Manhattan office to build and scale reliable production systems on AWS. You’ll own infrastructure, deployment pipelines, and observability, collaborating with engineers across product and security to keep latency low and uptime high.

You’ll manage Terraform-based infrastructure, strengthen monitoring with Datadog, and lead incident response, post-incident reviews, and runbooks while balancing long-term improvements with immediate

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

SITE RELIABILITY ENGINEER
SITE RELIABILITY ENGINEER

Onyxodds • New York (NY)

Vor Ort
USD 145.000 - 190.000
Equity ownership
Fully covered health benefits
Full-Stack FinTech Engineer — Trading Platform (Equity)
Full-Stack FinTech Engineer — Trading Platform (Equity)

Onyxodds • New York (NY)

Vor Ort
USD 150.000 - 200.000
Medical, dental, and vision insurance
Meaningful equity ownership
Senior SRE, Software Engineering (AWS / Scaling Infrastructure)
Senior SRE, Software Engineering (AWS / Scaling Infrastructure)

PulseRise Technologies • New York (NY)

Vor Ort
USD 130.000 - 160.000
Senior Site Reliability Engineer - Fintech Cloud & Resilience
Senior Site Reliability Engineer - Fintech Cloud & Resilience

Longbridge • New York (NY)

Vor Ort
USD 130.000 - 190.000
Frontend Engineer - Real-Time Trading UI (Equity Eligible)
Frontend Engineer - Real-Time Trading UI (Equity Eligible)

Onyxodds • New York (NY)

Vor Ort
USD 150.000 - 200.000
Equity ownership
Bonus potential
Health, dental, vision insurance
Senior Site Reliability Engineer — Cloud & Automation
Senior Site Reliability Engineer — Cloud & Automation

London Stock Exchange Group • St. Louis (MO), Northern (KY)

Hybrid
USD 140.000 - 180.000
Site Reliability Engineer III — Scale, Observability & Automation
Site Reliability Engineer III — Scale, Observability & Automation

onXmaps, Inc. • Bozeman (MT)

Hybrid
USD 130.000 - 153.000
Health benefits including no monthly–$
401(k) matching
Parental leave
+2
Senior Site Reliability Engineer - Hybrid NYC/SF, Equity
Senior Site Reliability Engineer - Hybrid NYC/SF, Equity

Prowork • New York (NY)

Hybrid
USD 180.000 - 200.000
Site Reliability Engineer
Site Reliability Engineer

Stott and May • New York (NY)

Vor Ort
USD 120.000 - 140.000
DevOps Engineer (Cloud)
DevOps Engineer (Cloud)

Chris Baily • New York (NY)

Hybrid
USD 150.000 - 185.000