Sr. Site Reliability Engineer

Far Coder

Northern (KY)

Hybrid

USD 25,000 - 40,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

MeridianLink is seeking a Senior Site Reliability Engineer to join our cloud engineering team in a remote-capable role serving USA-based fintech applications. You will own reliability, scalability, and observability across multi-cloud infrastructure, applying IaC and automated remediation to prevent outages.

You will design SLOs/SLIs, lead an observability program with modern tooling (Prometheus, Grafana, ELK, Datadog), and partner with developers to build resilient services on AWS or Azure,

Qualifications

  • 7+ years in SRE, DevOps, or platform engineering with production systems.
  • Expert in Azure and/or AWS cloud platforms.
  • Experience with observability: monitoring, logging, tracing.
  • Strong knowledge of SLOs/SLIs/SLAs and error budgets.
  • Proficient in Python and scripting for automation.
  • Experience with IaC tools like Terraform.
  • Familiarity with incident management and on-call.

Responsibilities

  • Design and maintain SLOs/SLIs for critical systems.
  • Lead observability strategy with monitoring/logging/tracing.
  • Own runbooks, incident response, postmortems.
  • Architect cloud infra on AWS/Azure with IaC.
  • Develop AIOps capabilities and automated remediation.

Skills

Kubernetes
Terraform
CI/CD
Python
AWS
Azure

Tools

Bash
Compliance
DevOps
Excel
MEAN
Predictive Analytics

Job description

# Remote Sr. Site Reliability Engineer Job at MeridianLinkUSA5 hours agoFull TimeUSA$25000 - $40000 USDKubernetesTerraformCI/CDPythonAWSAzureBashComplianceDevOpsExcelMEANPredictive Analytics“When applying, mention the word FarCoder to show you’ve read the job post completely. Employers can look for these words to identify genuine, thoughtful applicants and avoid spam.”## Job OverviewMeridianLink is hiring a remote candidate for Sr. Site Reliability Engineer. This is a full time position. Work location: USA.The role typically involves technologies such as Kubernetes, Terraform, CI/CD, Python, AWS, Azure.## Required Skills### Primary Skills* Kubernetes* Terraform* CI/CD* Python* AWS* Azure### Secondary Skills* Bash* Compliance* DevOps* Excel* MEAN* Predictive AnalyticsSkills required for this role include Kubernetes, Terraform, CI/CD, and related tools for day-to-day development.## Job Details* **Employment Type:** Full Time* **Location:** USA* **Salary:** $25000 – $40000 USD## Tech StackKubernetes, Terraform, CI/CD, Python, AWS, Azure, Bash, Compliance, DevOps, Excel, MEAN, Predictive Analytics## Role detailsAbout the RoleWe are seeking a Senior Site Reliability Engineer to join our cloud engineering team. You will own the reliability, scalability, and observability of our critical financial SaaS applications and infrastructure, working across cloud platforms to ensure our customers experience is seamless, secure, and performant services. This is a high-impact role for someone who is passionate about building resilient systems and preventing outages before they happen.Key Responsibilities* Design, implement, and maintain Service Level Objectives (SLOs) and Service Level Indicators (SLIs) across all critical systems; ensure we meet or exceed targets consistently* Lead observability strategy by designing comprehensive monitoring, logging, and tracing architectures; select and deploy observability tools that provide deep visibility into system behavior* Build and own runbooks, incident response procedures, and post-incident review processes; mentor the team on incident management and blameless postmortems* Architect and deploy cloud infrastructure on AWS or Azure; implement infrastructure-as-code practices and ensure high availability, disaster recovery, and business continuity* Develop automation and AIOps capabilities to reduce toil, accelerate incident detection, and enable self-healing systems; implement intelligent alerting to minimize false positives* Drive reliability improvements through load testing, chaos engineering, and failure scenario analysis; identify and eliminate single points of failure* Partner with application and backend teams to design reliable systems from inception; conduct architecture reviews and reliability assessments* Write production-grade Python tooling for automation, metrics collection, alert management, and operational workflows* Champion security and compliance in infrastructure; implement defense-in-depth principles for a regulated fintech environmentRequired Qualifications* 7+ years in Site Reliability Engineering, DevOps, platform engineering, or closely related roles with significant responsibility for production systems* Expert-level experience with Azure or AWS (or both); deep knowledge of compute, networking, storage, and managed services; experience managing infrastructure at scale* Demonstrated expertise in observability: designing and implementing monitoring, alerting, logging, and distributed tracing solutions; hands-on with observability platforms (e.g., Prometheus, Grafana, ELK, Datadog, New Relic, or similar)* Strong background in SLOs, SLIs, and SLAs; experience defining meaningful objectives and building systems to meet them; understanding of error budgets and their role in prioritization* Proven experience designing and troubleshooting highly available, resilient, and scalable systems; deep understanding of distributed systems concepts and failure modes* Proficiency in Python, PowerShell, bash, etc. scripting languages for production automation, tooling, and systems programming; ability to write clean, maintainable code for operational workflows* Hands-on experience with AIOps practices: event correlation, intelligent alerting, predictive analytics, and automated remediation; familiarity with AIOps platforms is a plus* Experience with infrastructure-as-code tools (e.g., Terraform, CloudFormation, Ansible); version control and CI/CD pipeline design* Track record of incident management and on-call ownership; comfort with incident response and the ability to remain calm under pressure* Excellent communication skills; ability to work cross-functionally and influence without authority; comfort mentoring junior engineersPreferred Qualifications* Experience in the fintech, payments, banking, or other regulated industries; understanding of compliance requirements (SOC 2, PCI-DSS, etc.)* Experience with Kubernetes and container orchestration; deep knowledge of containerized application deployment and management* Proficiency with observability as code; experience building custom metrics, dashboards, and alerts programmatically* Background in chaos engineering or reliability testing; experience using tools like Gremlin or similar platforms* Contribution to open-source observability or infrastructure projects* Expertise in network security, application security, or infrastructure hardening* Experience with database optimization, query performance tuning, and backup/recovery strategies
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Security Engineer
Security Engineer

Far Coder • Northern (KY)

Hybrid
USD 25,000 - 40,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

MeridianLink, Inc. • Northern (KY)

Hybrid
USD 140,000 - 210,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

MeridianLink • United States

Remote
USD 140,000 - 190,000
Senior DevOps Engineer
Senior DevOps Engineer

Far Coder • New York (NY)

On-site
USD 160,000 - 200,000
Equity participation
Sr. Backend Engineer
Sr. Backend Engineer

Fuel Talent LLC • Seattle (WA)

On-site
USD 200,000 - 225,000
Site Reliability Engineer
Site Reliability Engineer

Jobot • Akron (OH)

Remote
USD 100,000 - 150,000
Comprehensive health insurance
Vision insurance
Dental insurance
+3
Remote Senior Site Reliability Engineer-Scale & Resilience
Remote Senior Site Reliability Engineer-Scale & Resilience

Far Coder • Northern (KY)

Hybrid
USD 25,000 - 40,000
Manager, Software Engineering
Manager, Software Engineering

CV in • Northern (KY)

Hybrid
USD 110,000 - 160,000
Manager, Engineering - DevOps
Manager, Engineering - DevOps

Insightsoftware • United States

Remote
USD 139,000 - 174,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Luxoft • Buffalo (NY)

On-site
USD 140,000 - 190,000