SRE

Benton Partners

Chicago, New York (IL, NY)

On-site

USD 175,000 - 225,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Benton Partners in Chicago/New York seeks an experienced Site Reliability Engineer to join our Platform Engineering team, shaping SRE practices across cloud and on-prem environments. You will drive reliability standards and collaborate with our DevOps and Cloud teams to scale trading systems.

You will implement observability, automate deployments, and define SLOs to improve performance, resilience, and efficiency across diverse platforms.

Qualifications

  • 8+ years of experience in SRE or similar roles within distributed systems.
  • Expert with Prometheus, Grafana, Loki, Tempo, and OpenTelemetry.
  • Deep knowledge of Kubernetes and Docker containerization.
  • Experience with cloud (AWS preferred) and on-prem hosting.
  • Scripting in Python, Bash, or Go to automate tasks.
  • Strong CI/CD, agile methods, and DevOps culture.

Responsibilities

  • Develop and promote our SRE philosophy and scalable infrastructure practices.
  • Implement end-to-end observability using Prometheus, Grafana, Loki, Tempo.
  • Participate in on-call rotation (~1 week per month).
  • Define standards for reliability in Kubernetes environments; optimize configurations.
  • Develop automation to improve deployment pipelines, health checks, and recovery.
  • Collaborate with development teams to improve service stability and SLOs.

Skills

SRE experience
Prometheus
Grafana
Loki
Tempo
OpenTelemetry
Kubernetes
Docker
AWS
On-premises
Python
Bash
Go
CI/CD
DevOps

Job description

Senior Site Reliability Engineer - Platform

ChicagoNew York

We arelookingfora Site Reliability Engineer, to join our growing Platform Engineering team,who can cultivate our SRE philosophy, processes, and technologies from the ground up.This roleentailsdriving standards and fostering adoption across our technology teams, whilstcloselypartnering with our DevOps and Cloud teams.

With a hands-on approach,you'llwork across both cloud and on-premises hosting platforms, ensuring the reliability and scalability of ourtradingsystemsand production environments. This is a chance to play a pivotal role in transforming our operational capabilities and enhancing performance across a wide array of environments and platforms.

Key Responsibilities:

  • Develop and promote our SRE philosophy,establishingbest practices and processes that will be instrumental in scaling our infrastructure.
  • Implementand scaleend-to-end observability and monitoring solutions using Prometheus, Grafana, Loki, andTempo, ensuring high visibility into application performance and infrastructure health.
  • Participate in on-call rotation with approximately 1 week per month of on-call time shared equally across members of the team
  • Review and define standards forapplication reliability requirements within ourKubernetesenvironment, ensuringapplication configuration isoptimizedfor performance,costand reliability.
  • Develop automation and tooling to improve efficiency and reliability of deployment pipelines, system health checks, and recovery procedures.
  • Collaborate with development teams to enhance service stability, scalability, and fault tolerance through SRE best practices like blameless post-mortems and service levelobjectives(SLOs).

To be considered a good fit, you musthave:

  • 8+ years of experience in SRE or similar roles within complex, distributed systems environments.
  • SMEwith key SRE technologies such as Prometheus, Grafana, Loki,Tempo,andOpenTelemetry.
  • Extensive knowledge of container orchestration using Kubernetes and containerization with Docker.
  • Hands-on experience with both cloud (AWS preferred) and on-premiseshosting platforms.
  • Proven ability to script in languages like Python, Bash, or Go, to automate routine tasks and deployment pipelines.
  • Strong understanding of CI/CD principles, agile methodologies, and DevOps culture.
  • High levelof initiative, passion for reliability engineering, detail orientation, and follow-through capabilities.
  • Exceptional interpersonal and communication skills, with the ability to explain complex technical concepts to a diverse audience.

With respect to NY, CA, and IL based applicants, the starting base pay range for this role is between USD 175000 and USD 225000 annually. The actual base pay is dependent upon several factors, including, but not limited to, relevant experience, business needs and market demands. This role may also be eligible for bonus compensation and employee benefits.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineering (SRE)
Site Reliability Engineering (SRE)

Weekday (YC W21) • New York (NY)

On-site
USD 150,000 - 250,000
Equity or bonus opportunities
Health benefits
Paid time off
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Mission Staffing • New York (NY)

Hybrid
USD 140,000 - 200,000
Senior DevOps Engineer/Site Reliability Engineer-East Coast
Senior DevOps Engineer/Site Reliability Engineer-East Coast

Stellar Cyber • North Carolina

On-site
USD 165,000 - 215,000
Pre‑IPO Stock Options
Medical, Dental & Vision care
401(k)
+2
Senior DevOps Engineer/Site Reliability Engineer-East Coast
Senior DevOps Engineer/Site Reliability Engineer-East Coast

Stellar Cyber • New York (NY)

Hybrid
USD 165,000 - 215,000
Pre-IPO Stock Options
Medical, Dental & Vision care
401(k)
+1
Senior DevOps Engineer/Site Reliability Engineer-East Coast
Senior DevOps Engineer/Site Reliability Engineer-East Coast

Stellar Cyber • New Jersey

On-site
USD 165,000 - 215,000
Pre-IPO Stock Options
Medical, Dental & Vision care
401(k)
+2
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Methodic • San Francisco (CA)

On-site
USD 140,000 - 210,000
Senior SRE Engineer
Senior SRE Engineer

Compunnel, Inc. • Alpharetta (GA)

On-site
USD 140,000 - 190,000
Senior Site Reliability Engineer - Banking & Finance
Senior Site Reliability Engineer - Banking & Finance

Hamilton Barnes Associates Limited • New York (NY)

Hybrid
USD 360,000 - 440,000
Strong compensation and bonus potential
Collaborative engineering culture
Work on mission-critical systems
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Hamilton Barnes ? • New York (NY)

Hybrid
USD 340,000 - 400,000
Strong compensation and bonus potential
Hybrid working environment
Collaborative engineering culture
Site Reliability Engineer
Site Reliability Engineer

Harrison Clarke • New York (NY)

On-site
USD 120,000 - 160,000