Senior Site Reliability Engineer

Cross River

United States

Remote

USD 160,000 - 200,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Cross River is seeking a Senior Site Reliability Engineer to own the reliability and performance of mission-critical systems. You will drive DevOps guardrails, CI/CD, and IaC initiatives across engineering teams to enable scalable, secure infrastructure in a fast-paced fintech environment.

The ideal candidate has 8+ years in SRE/DevOps, deep AWS experience, Terraform, Docker, and strong observability skills with tools like New Relic, ELK, Prometheus, or Grafana.

Qualifications

  • 8+ years in SRE, DevOps, or Infra Engineering roles.
  • 5+ years with AWS; multi-cloud is a plus.
  • Terraform proficiency.
  • CI/CD design, build and automation.
  • Docker and container orchestration (ECS preferred).
  • Observability with New Relic, ELK, CloudWatch, Prometheus, Datadog, or Grafana.
  • Proficiency in .NET, PowerShell, Python, Go, or Bash.
  • Linux and Windows administration.
  • DNS, load balancing, CDNs, and network security.
  • Strong written and verbal communication.

Responsibilities

  • Define and enforce DevOps guardrails, standards, and best practices.
  • Enable teams to design, implement, and maintain CI/CD pipelines.
  • Co-develop IaC with Terraform.
  • Establish deployment strategies (blue/green, canary, rolling).
  • Build developer self-service tooling and internal platforms.
  • Shift-left reliability, security, observability.
  • Define and monitor SLOs, SLIs, and error budgets.
  • Maintain observability stacks with logging, metrics, tracing, and alerting.
  • Lead incident response and on-call rotations.
  • Capacity planning and performance engineering.
  • Eliminate toil through automation.
  • Conduct reliability reviews and chaos engineering.
  • Manage and optimize cloud infrastructure.
  • Collaborate to improve system architecture and resiliency.
  • Support .NET-based services on Windows/Linux.

Skills

SRE / DevOps
AWS
IaC
CI/CD
Containers & Orchestration
Observability
Scripting / Programming
Linux & Windows
Networking
Communication

Tools

Terraform
Docker
Kubernetes / ECS
New Relic
ELK Stack
Prometheus
Grafana
CloudWatch
Datadog
GitOps

Job description

Who We Are

Cross River builds the infrastructure behind the world's most innovative financial products. Our technology and capital solutions power payments, cards, lending, and digital asset capabilities that move money safely, instantly, and inclusively — trusted by leading fintechs, enterprises, and disruptors across the globe.

Our mission is simple: to build the financial infrastructure that expands access and opportunity for all. Guided by a culture of collaboration, curiosity, and purpose, Cross River has been named one of American Banker'sBest Places to Work in Fintech year after year. Whether you're designing code, solving regulatory puzzles, or developing strategy, you'll join a team where innovation and integrity drive everything we do — and where your work helps shape the future of finance.

What We're Looking For

We are seeking a highly skilled and motivated Senior Site Reliability Engineer with 8+ years of hands-on experience ensuring the reliability, scalability, and performance of mission-critical systems. The ideal candidate brings deep expertise in building and maintaining production infrastructure, establishing DevOps best practices, and driving operational excellence across engineering teams. We're looking for someone who takes ownership of system reliability, thrives in a collaborative and fast-paced environment, and is passionate about building resilient financial infrastructure.

Responsibilities
  • Define and enforce DevOps guardrails , standards, and best practices to ensure consistency, security, and compliance across engineering teams
  • Enable Engineering teams to design, implement, and maintain CI/CD pipelines best to enable fast, safe, and repeatable deployments across all environments
  • Co-Develop and maintain Infrastructure as Code (IaC) with Application teams using tools such as Terraform
  • Establish and govern deployment strategies including blue/green, canary, and rolling deployments
  • Build and maintain developer self-service tooling and internal platforms that accelerate delivery while maintaining governance
  • Champion a "shift-left" culture by embedding reliability, security, and observability practices early in the software development lifecycle
  • Help define, implement, and monitor Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets for critical services
  • Build and maintain comprehensive observability stacks including centralized logging, metrics, distributed tracing, and alerting using tools such as New Relic, ELK, Prometheus, and Grafana
  • Lead incident response and management , including on-call rotations, root cause analysis (RCA), and blameless post-mortems
  • Perform capacity planning and performance engineering to ensure systems scale efficiently with business growth
  • Identify and eliminate toil through automation, reducing manual operational overhead
  • Conduct reliability reviews and chaos engineering exercises to proactively identify and mitigate failure modes
  • Manage and optimize cloud infrastructure to balance reliability, cost, and performance
  • Collaborate with software engineering teams to improve system architecture, resiliency patterns , and fault tolerance
  • Support and maintain .NET-based services running across Windows and Linux environments
Qualifications
  • SRE / DevOps Experience: 8+ years in SRE, DevOps, or Infrastructure Engineering roles
  • Cloud Platforms: 5+ years with AWS (preferred); experience with multi-cloud is a plus
  • Infrastructure as Code: Strong proficiency with Terraform
  • CI/CD: Deep experience designing, building, and maintaining CI/CD pipelines and automation workflows
  • Containers & Orchestration: Strong experience with Docker and container orchestration (ECS preferred)
  • Observability: Proficiency with tools such as New Relic, ELK Stack, CloudWatch, Prometheus, Datadog, or Grafana
  • Scripting / Programming: Proficiency in .NET, PowerShell, Python, Go, or Bash
  • Operating Systems: Strong Linux and Windows systems administration skills
  • Networking: Solid understanding of DNS, load balancing, CDNs, and network security
  • Communication: Strong written and verbal communication skills
Preferred / Nice-to-Have
  • Experience implementing and managing service mesh technologies (e.g., Istio, Linkerd, AWS App Mesh)
  • Familiarity with SRE frameworks as outlined in Google's SRE handbook
  • Experience with secrets management (e.g., HashiCorp Vault, AWS Secrets Manager)
  • Understanding of compliance and regulatory requirements in financial services (SOC 2, PCI-DSS, etc.)
  • Experience with chaos engineering tools (e.g., Gremlin, Litmus, AWS Fault Injection Simulator)
  • Experience supporting .NET applications in production environments
  • Financial industry / banking infrastructure experience is helpful, but not required
  • Crypto / blockchain infrastructure experience is helpful, but not required
  • Experience with GitOps workflows and patterns
  • Familiarity with cost optimization and FinOps practices in cloud environments
Key Metrics of Success
  • System uptime and availability targets consistently met or exceeded
  • Reduction in mean time to detect (MTTD) and mean time to resolve (MTTR)
  • Adoption and adherence to DevOps guardrails across engineering teams
  • Measurable reduction in operational toil through automation
  • Healthy error budget management across critical services

#LI-Remote

Salary Range: $160,000.00 - $200,000.00

Cross River is an Equal Opportunity Employer. Cross River does not discriminate on the basis of race, religion, color, sex, gender identity, sexual orientation, age, non-disqualifying physical or mental disability, national origin, veteran status or any other basis covered by appropriate law. All employment is decided on the basis of qualifications, merit, and business need.

By submitting your application, you give Cross River permission to email, call, or text you using the contact details provided. We will only contact you with job related information.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Engineer II, IT Infrastructure
Engineer II, IT Infrastructure

Cross River • Fort Lee (NJ)

On-site
USD 80,000 - 110,000
Systems Administrator, IT Infrastructure
Systems Administrator, IT Infrastructure

Cross River • Fort Lee (NJ)

On-site
USD 80,000 - 110,000
Senior Site Reliability Engineer — Remote Infra & CI/CD
Senior Site Reliability Engineer — Remote Infra & CI/CD

Cross River • United States

Remote
USD 160,000 - 200,000
Senior Software Engineer
Senior Software Engineer

Cross River • Fort Lee (NJ)

On-site
USD 160,000 - 200,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Hobbsnews • Jersey City (NJ)

On-site
USD 153,000 - 192,000
Benefits eligible
Discretionary incentive plan
Senior Site Reliability Engineer
Senior Site Reliability Engineer

United States Digital Space LLC • Charlotte (TX)

On-site
USD 153,000 - 192,000
Discretionary incentive eligible
Benefits package
Senior Site Reliability Engineer
Senior Site Reliability Engineer

National Black MBA Association • Jersey City (NJ)

On-site
USD 153,000 - 192,000
Benefits eligible
Annual discretionary plan
Senior DevOps/SRE Engineer
Senior DevOps/SRE Engineer

SEI • Chicago (IL)

On-site
USD 140,000 - 170,000
Comprehensive healthcare benefits
401(k) match
Paid Time Off (PTO)
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Mission Staffing • New York (NY)

On-site
USD 140,000 - 200,000
Site Reliability Engineer
Site Reliability Engineer

TalentDome Staffing • United States

On-site
USD 140,000 - 210,000