Site Reliability Engineer

Longbridge

New York (NY)

On-site

USD 140,000 - 190,000

Full time

27 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Longbridge is seeking a hands-on Site Reliability Engineer to design, scale, and safeguard the reliability of our next-generation financial platforms. You will partner with product and engineering teams across the globe to build resilient systems.

You will own system reliability, implement automation at scale, lead incident response, and promote cloud-native tooling (Kubernetes, Prometheus, AWS/GCP). The role requires 5+ years in SRE/DevOps, strong Linux, and CI/CD expertise.

Qualifications

  • 5+ years in SRE/DevOps or production engineering roles.
  • Strong background in cloud platforms (AWS/GCP/Azure) and container orchestration.
  • Proficiency in at least one programming language (Python or Go) for automation.
  • Solid Linux administration skills and experience with CI/CD pipelines.
  • Proven ability in incident management and troubleshooting distributed systems.

Responsibilities

  • Own system reliability: design and operate highly available distributed systems.
  • Build automation at scale: monitoring, alerting, and infrastructure as code.
  • Partner globally: collaborate with design through deployment across teams.
  • Lead incident response: on-call, root-cause analysis, reduce MTTR.
  • Future-proof our stack: evaluate cloud-native tech (Kubernetes, Prometheus, AWS/GCP).
  • Stress-test and safeguard: disaster recovery, chaos testing, capacity planning for wealth services.

Skills

Cloud platforms
Container orchestration
Automation scripting
Linux administration
CI/CD
Incident management
Cross-functional collaboration
Fintech knowledge
Chinese language (bonus)

Tools

Docker
Kubernetes
Terraform
Ansible
Helm
Prometheus

Job description

Longbridge is a fast-growing online brokerage platform on a mission to make investing smarter, simpler, and more accessible for everyone.

As part of our global expansion, we’re looking for a hands-on Site Reliability Engineer (SRE) to design, scale, and safeguard the reliability of our next-generation financial platforms. This is a high-impact role where you’ll partner closely with product and engineering teams across the globe.

What You’ll Do
  • Own system reliability: Design, implement, and operate highly available, secure distributed systems to meet strict uptime and performance targets.
  • Build automation at scale: Develop and enforce best practices in monitoring, alerting, and infrastructure-as-code (e.g., Terraform, Ansible, Helm).
  • Partner globally: Work with development teams from design through deployment, ensuring reliability and resiliency are built in from day one.
  • Lead incident response: Drive on-call processes, conduct root-cause analysis, and continuously reduce MTTR and failure recurrence.
  • Future-proof our stack: Evaluate and adopt modern cloud-native technologies (e.g., Kubernetes, Prometheus, AWS/GCP) to keep systems secure and scalable.
  • Stress-test and safeguard: Lead disaster recovery, chaos testing, and capacity planning for critical wealth management services.
What We’re Looking For
  • 5+ years of experience in SRE, DevOps, or production engineering roles.
  • Strong background in AWS (or GCP/Azure) and container orchestration (Docker, Kubernetes).
  • Proficiency in at least one programming language (Python, Go, or similar) for automation and tooling.
  • Solid Linux administration skills and experience with CI/CD pipelines.
  • Proven ability in incident management and troubleshooting distributed systems.
  • Strong collaboration and communication skills across remote/global teams.
  • Comfortable working in a fast-moving fintech/tech startup environment.
  • Bonus: Experience supporting regulated financial systems; ability to communicate in Chinese to collaborate with Asia-based colleagues is a plus
Why Join Us
  • Shape the reliability foundation of our U.S. product launch in wealth tech.
  • Opportunity to build systems from the ground up and influence technical direction.
  • Competitive compensation package and growth opportunities.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Longbridge Securities • Town of Texas (WI)

On-site
USD 100,000 - 130,000
Competitive compensation package
Growth opportunities
Site Reliability Engineer
Site Reliability Engineer

Socket.dev • New York (NY)

On-site
USD 140,000 - 190,000
Site Reliability Engineer
Site Reliability Engineer

Socket.dev • Dallas (TX)

On-site
USD 120,000 - 190,000
Site Reliability Engineer
Site Reliability Engineer

Longbridge Singapore • New York (NY)

On-site
USD 140,000 - 190,000
Competitive compensation
Growth opportunities
Site Reliability Engineer (Remote - United States)
Site Reliability Engineer (Remote - United States)

Longbridge Singapore • California (MO)

On-site
USD 90,000 - 120,000
Medical insurance
Vision insurance
401(k)
+1
Site Reliability Engineer
Site Reliability Engineer

Longbridge Singapore • Dallas (TX)

On-site
USD 120,000 - 190,000
Senior Site Reliability Engineer — FinTech & Cloud Native
Senior Site Reliability Engineer — FinTech & Cloud Native

Longbridge • New York (NY)

On-site
USD 140,000 - 190,000
FinTech SRE: Build Scalable, Reliable Cloud Platforms
FinTech SRE: Build Scalable, Reliable Cloud Platforms

Socket.dev • Dallas (TX)

On-site
USD 120,000 - 190,000
Global FinTech SRE: Build Reliable, Scalable Cloud Systems
Global FinTech SRE: Build Reliable, Scalable Cloud Systems

Socket.dev • New York (NY)

On-site
USD 140,000 - 190,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Mission Staffing • New York (NY)

Hybrid
USD 140,000 - 200,000