Integration Reliability Engineer

Claryo, Inc.

San Francisco (CA)

On-site

USD 150,000 - 170,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Top-tier medical, dental, and vision coverage
401k with employer matching
Equity
Parental leave
Unlimited vacation

Job summary

Claryo, Inc. is seeking an Integration Reliability Engineer in San Francisco, CA, responsible for ensuring the reliability of systems across cloud and edge environments. The candidate will build and maintain observability tools and improve incident response processes. Qualifications include 3+ years of experience in SRE, strong Linux and networking skills, and familiarity with Kubernetes and cloud platforms. This role offers a competitive salary range of $150K - $170K along with comprehensive benefits including medical coverage and unlimited vacation.

Qualifications

  • 3+ years of experience in SRE, infrastructure, or distributed systems.
  • Strong Linux and networking fundamentals.
  • Experience operating systems in production environments.
  • Comfortable working in real-world, imperfect environments.

Responsibilities

  • Own reliability of systems across cloud (Kubernetes), edge compute, and on‑site deployments.
  • Build and maintain monitoring, alerting, and observability systems.
  • Improve incident response and on-call processes.
  • Diagnose issues across infrastructure, networking, and distributed systems.

Skills

SRE, infrastructure, or distributed systems
Linux
Networking fundamentals
Cloud platforms (GCP, AWS, or Azure)
Kubernetes and containerized systems
Observability tools (Prometheus, Grafana, OpenTelemetry)
Debugging issues across multiple layers
Event-driven systems (Kafka or similar)

Job description

We’re looking for an Integration Reliability Engineer to own the reliability of our system across cloud, edge, and real‑world environments. Our platform runs across distributed infrastructure—connecting cloud services, on‑site compute, and live video/data pipelines inside warehouses. This role is responsible for making systems observable, diagnosable, and repeatable as we scale across deployments, working closely with engineering and deployment teams to ensure the system performs reliably in production—not just in ideal conditions.

What You’ll Own
  • Own reliability of systems across cloud (Kubernetes), edge compute, and on‑site deployments
  • Build and maintain monitoring, alerting, and observability systems
  • Define and improve incident response, severity levels, and on‑call processes
  • Improve deployment and bring‑up workflows across facilities
  • Diagnose issues across infrastructure, networking, and distributed systems
  • Partner with engineering to identify root causes and prevent recurring issues
  • Improve system visibility, debugging, and operational toolingHelp make deployments repeatable and scalable across sites
Required Qualifications
  • 3+ years of experience in SRE, infrastructure, or distributed systems
  • Strong Linux and networking fundamentals
  • Experience operating systems in production environments
  • Experience working with networking in constrained or distributed environments (e.g., VPNs, secure tunnels, on‑site networking)
  • Experience with:
    • Kubernetes and containerized systems
    • Cloud platforms (GCP, AWS, or Azure)
    • Observability tools (Prometheus, Grafana, OpenTelemetry, etc.)
  • Ability to debug issues across multiple layers of the stack (infra → services → network)
  • Comfortable working in real‑world, imperfect environments (not just clean cloud systems)
  • Strong ownership and ability to drive issues to resolution
Preferred Qualifications
  • Experience with multi‑site or edge deployments
  • Experience with event‑driven systems (Kafka or similar)
  • Familiarity with video or streaming systems (RTSP, WebRTC)
  • Experience working with hardware‑integrated systems
  • Exposure to security/compliance frameworks (SOC2, ISO27001, etc.)
  • US citizen/ permanent resident
  • Located in SFBAY or NY area
Why This Role Matters
  • Systems that work outside ideal environments
  • Fast, reliable diagnosis and recovery when things break
  • Repeatable deployments across real‑world facilities
Equal Opportunity Statement

We’re an equal opportunity employer that values diversity and inclusion. We welcome teammates of all backgrounds and don’t discriminate based on race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or veteran status.

Benefits

At Claryo, we offer a competitive benefits package that supports your health and well‑being, including top‑tier medical, dental, and vision coverage, 401k with employer matching, equity, parental leave, and unlimited vacation.

Compensation Range: $150K - $170K

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Integration Reliability Engineer
Integration Reliability Engineer

Claryo, Inc. • New York (NY)

On-site
USD 150,000 - 170,000
Top-tier medical coverage
401k with employer matching
Unlimited vacation
Systems Reliability Engineer (SRE)
Systems Reliability Engineer (SRE)

Claryo • San Francisco (CA)

On-site
USD 140,000 - 190,000
Systems Reliability Engineer (SRE)
Systems Reliability Engineer (SRE)

Claryo • New York (NY)

Hybrid
USD 120,000 - 180,000
Medical, dental, vision coverage
401k with employer matching
Equity
+2
Systems Reliability Engineer (SRE)
Systems Reliability Engineer (SRE)

Claryo, Inc. • New York (NY)

On-site
USD 150,000 - 170,000
Medical, Dental, Vision
401k with employer matching
Equity
+2
Systems Reliability Engineer (SRE)
Systems Reliability Engineer (SRE)

Claryo, Inc. • San Francisco (CA)

On-site
USD 150,000 - 170,000
Medical coverage
Dental coverage
Vision coverage
+4
Reliability Engineer: Cloud, Edge & On-site Deployments
Reliability Engineer: Cloud, Edge & On-site Deployments

Claryo, Inc. • San Francisco (CA)

On-site
USD 150,000 - 170,000
Systems Reliability Engineer – Edge & Cloud Observability
Systems Reliability Engineer – Edge & Cloud Observability

Claryo • New York (NY)

Hybrid
USD 120,000 - 180,000
Medical, dental, vision coverage
401k with employer matching
Equity
+2
Edge & Cloud SRE: Reliability & Real-World Deployments
Edge & Cloud SRE: Reliability & Real-World Deployments

Claryo • San Francisco (CA)

On-site
USD 140,000 - 190,000
SRE: Edge-to-Cloud Reliability & Observability
SRE: Edge-to-Cloud Reliability & Observability

Claryo, Inc. • New York (NY)

On-site
USD 150,000 - 170,000
Medical, Dental, Vision
401k with employer matching
Equity
+2
Edge & Cloud SRE — Reliability & Observability
Edge & Cloud SRE — Reliability & Observability

Claryo, Inc. • San Francisco (CA)

On-site
USD 150,000 - 170,000
Medical coverage
Dental coverage
Vision coverage
+4