Senior Platform & Reliability Engineer - Remote-First Lead

Contabo

United States

On-site

USD 104,000 - 150,000

Full time

3 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Remote or hybrid with flexible hours
Workation across the EU
EGYM Wellpass
Corporate discounts program
30 days vacation + holidays off
Extra day off for social

Job summary

Contabo seeks a Senior Platform & Reliability Engineer to own core infrastructure services, including API gateways, storage, identity, and observability. You will design on-call runbooks, improve distributed tracing and dashboards, and advise across teams on shared infrastructure.

The role emphasizes pragmatic stability and knowledge sharing across a remote-first organization. You will work with Kong, Nginx Ingress, Ceph, Longhorn, Vault, Keycloak, Cloudflare, Prometheus, Grafana, and

Qualifications

  • 7+ years in platform, infrastructure, or SRE roles with end-to-end responsibility.
  • Hands-on storage (Ceph) and Kubernetes persistent storage (Longhorn).
  • Experience with API gateways/ingress (Kong, Nginx Ingress).
  • CDN/edge security, DDoS mitigation, WAF config (Cloudflare).
  • Strong system-design fundamentals and trade-off judgment.
  • Observability experience (Prometheus, Grafana, OpenTelemetry) desirable.
  • Secrets/identity infra (Vault, Keycloak); on-call process drafting a plus.

Responsibilities

  • Design and establish an on-call process with runbooks.
  • Mature observability: tracing, SLOs/SLIs, dashboards across the stack.
  • Advise across teams on shared infrastructure services.
  • Document knowledge and distribute it to prevent single points of failure.
  • Evaluate aging components and decide fixes, replacements, or retirements.

Skills

SRE experience
English fluency
On-call design
Documentation
Cross-team collaboration

Tools

Kong
Nginx Ingress
Ceph
Longhorn
Vault
Keycloak
Cloudflare
Prometheus
Grafana
OpenTelemetry
Kubernetes
Vault/Identity infra

Job description

Contabo seeks a Senior Platform & Reliability Engineer to own core infrastructure services, including API gateways, storage, identity, and observability. You will design on-call runbooks, improve distributed tracing and dashboards, and advise across teams on shared infrastructure.

The role emphasizes pragmatic stability and knowledge sharing across a remote-first organization. You will work with Kong, Nginx Ingress, Ceph, Longhorn, Vault, Keycloak, Cloudflare, Prometheus, Grafana, and

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Platform Reliability Engineer - Kubernetes & Observability
Senior Platform Reliability Engineer - Kubernetes & Observability

PLP Group • New York (NY)

On-site
USD 75,000 - 130,000
Senior Platform Infrastructure Engineer - Reliability & Security
Senior Platform Infrastructure Engineer - Reliability & Security

DaParrot Ltd • Northern (KY)

Hybrid
USD 140,000 - 210,000
Competitive compensation with equity
Flexible PTO
401k
+3
Senior Platform Reliability Engineer — Remote
Senior Platform Reliability Engineer — Remote

Alkami Technology • United States

On-site
USD 145,000 - 165,000
Remote-first environment
Unlimited paid time off
401(k) with employer match
Remote Platform Reliability Engineer
Remote Platform Reliability Engineer

Redwood Logistics LLC • Northern (KY)

Hybrid
USD 160,000 - 175,000
Health insurance
401k with match
Paid time off
+1
Senior Platform & Reliability Engineer (all genders)
Senior Platform & Reliability Engineer (all genders)

Contabo • United States

Hybrid
USD 104,000 - 150,000
Remote or hybrid with flexible hours
Workation across the EU
EGYM Wellpass
+3
Senior Production Engineer Platform Reliability (Remote)
Senior Production Engineer Platform Reliability (Remote)

GitHub • United States

Remote
USD 255,000 - 425,000
Senior SRE & Backend Engineer - Reliability at Scale(Remote)
Senior SRE & Backend Engineer - Reliability at Scale(Remote)

Nabla • Buffalo (NY)

Hybrid
USD 140,000 - 210,000
Competitive salary
Stock options
Medical coverage
+5
Senior Platform Engineer - Remote, Reliability & Automation
Senior Platform Engineer - Remote, Reliability & Automation

Service Management Group (SMG) • United States

Remote
USD 120,000 - 180,000
Unlimited PTO
Parental leave
Company-issued equipment and tools
+1
Senior Platform Reliability Engineer - Kubernetes & Cloud
Senior Platform Reliability Engineer - Kubernetes & Cloud

Modal • New York (NY)

On-site
USD 150,000 - 190,000
Senior Platform Engineer — Real-Time, Remote High-Perf Infra
Senior Platform Engineer — Real-Time, Remote High-Perf Infra

Remotiee • United States

On-site
USD 100,000 - 130,000