Senior Site Reliability Engineer: Scalable Infra & Observability

Early Warning

Chicago (IL)

Hybrid

USD 106,000 - 130,000

Full time

12 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Healthcare Coverage
401(k) Matching
Paid Time Off
Parental Leave (12 weeks)

Job summary

Early Warning is seeking an experienced engineer to help build and maintain scalable application infrastructure with a focus on reliability and automation. You will implement tooling around deployments, disaster recovery, and observability while ensuring security and high performance.

You will work with development teams to apply modern microservice design patterns and contribute to on-call rotation, incident management, and cross-team collaboration in a hybrid work environment.

Qualifications

  • Bachelor’s degree in Computer Science or related field.
  • 3+ years managing large complex projects in a technical environment.
  • Experience in Incident and Problem Management.
  • Proven work experience in medium to large-scale enterprise.
  • Strong scripting language skills.
  • Hands-on with observability solutions.
  • Linux systems administration.
  • Knowledge of Git and security/encryption protocols.
  • Strong communication and collaboration across teams.

Responsibilities

  • Implement software and tools to improve performance, availability, scalability, and latency while meeting security standards.
  • Build automation and tooling for deployments, configuration changes and disaster recovery.
  • Implement and evangelize Observability and monitoring systems to detect problems.
  • Evaluate application capacity and provide stats to product/business teams; plan scalable paths.
  • Identify performance bottlenecks and troubleshoot with cross-functional teams.
  • Standardize practices across disciplines to improve delivery.
  • Collaborate with development teams on lifecycle feedback and microservice design patterns.
  • Serve as a technical liaison and provide runbooks to Level 1/2 teams.
  • Participate in 24x7 on-call rotation.
  • Develop repeatable patterns and reusable work across teams.
  • Support data/system confidentiality and integrity.

Skills

Incident Management
Problem Management
Linux
Git
Scripting
Observability
Security protocols
CI/CD
Collaboration
Communication

Education

Bachelor’s Degree in Computer Science or related field

Tools

AWS
Docker
Kubernetes
Swarm

Job description

Early Warning is seeking an experienced engineer to help build and maintain scalable application infrastructure with a focus on reliability and automation. You will implement tooling around deployments, disaster recovery, and observability while ensuring security and high performance.

You will work with development teams to apply modern microservice design patterns and contribute to on-call rotation, incident management, and cross-team collaboration in a hybrid work environment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer - Scale & Observability
Senior Site Reliability Engineer - Scale & Observability

Early Warning Services LLC • United States

Hybrid
USD 118,000 - 183,000
Healthcare Coverage
401(k) Retirement Plan with company-mn
Paid Time Off
+1
Senior SRE: Scalable Infra, Observability & Automation
Senior SRE: Scalable Infra, Observability & Automation

Early Warning • Scottsdale (AZ)

Hybrid
USD 106,000 - 156,000
Healthcare Coverage
401(k) Company Match
Paid Time Off
+2
Senior Site Reliability Engineer: Observability & Resiliency
Senior Site Reliability Engineer: Observability & Resiliency

Early Warning • San Francisco (CA)

Hybrid
USD 139,000 - 174,000
Discretionary incentive plan
Comprehensive benefits package
Senior Site Reliability Engineer: Observability & Resiliency
Senior Site Reliability Engineer: Observability & Resiliency

Early Warning • Chicago (IL)

Hybrid
USD 116,000 - 145,000
Discretionary incentive plan
Health benefits
Senior Site Reliability Engineer: Scalable Hybrid Infra
Senior Site Reliability Engineer: Scalable Hybrid Infra

Redwood Materials • Nevada (IA)

On-site
USD 140,000 - 180,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

O.C. Tanner • Salt Lake City (UT)

On-site
USD 130,000 - 180,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Axiom Pursuits • San Francisco (CA)

On-site
USD 150,000 - 180,000
Staff SRE: Scale, Resilience & Observability
Staff SRE: Scale, Resilience & Observability

Early Warning Services LLC • San Francisco (CA)

On-site
USD 130,000 - 160,000
Healthcare Coverage
401(k) Retirement Plan
Paid Time Off
+2
Senior SRE & Platform Engineer — Observability & Automation
Senior SRE & Platform Engineer — Observability & Automation

Techunting • United States

On-site
USD 120,000 - 150,000
Staff SRE: Scale, Resilience & Observability
Staff SRE: Scale, Resilience & Observability

Early Warning Services LLC • Scottsdale (AZ)

On-site
USD 120,000 - 160,000
Healthcare Coverage
401(k) Retirement Plan
Flexible Time Off
+2