Staff Site Reliability Engineer

Pearl Street Technologies

Pittsburgh (Allegheny County)

On-site

USD 130,000 - 180,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Enverus is seeking a Staff Site Reliability Engineer to join our Cloud Engineering team, managing the companys AWS presence and driving reliability across scale. The role emphasizes automation, collaboration with development teams, and maintaining high uptime as the platform rapidly evolves.

Youll work with Kubernetes, Terraform/CloudFormation, and multiple cloud providers to enable zero-downtime deployments and robust operational practices.

Qualifications

  • 5+ years of Windows and Linux server administration.
  • 3+ years of AWS administration.
  • 3+ years in a high‑performance DevOps/SysOps/Operations team.
  • Excellent communication and collaboration skills.
  • Experience working with global teams (NA, Europe, Asia).
  • Experience with Kubernetes, IaC tools (Terraform/CloudFormation), and cloud providers.

Responsibilities

  • Manage our global AWS presence and ensure infrastructure reliability.
  • Keep infrastructure humming during releases and maintenance.
  • Organize, secure, and automate existing infrastructure and deployments.
  • Collaborate with developers to improve operational performance.
  • Ensure platform stability and balance under growth.
  • Maintain high site uptime while embracing rapid change.
  • Scale infrastructure to meet demand and evolving tech.
  • Support zero-downtime deployments with code updates.
  • Develop and improve operational practices and procedures.
  • Implement, monitor, and maintain CI/CD frameworks.
  • Coordinate on-call rotations and automate repetitive tasks.

Skills

Windows & Linux administration
AWS administration
DevOps / SysOps
Communication skills
Global collaboration

Tools

Kubernetes
Terraform
CloudFormation
AWS
Python
C#
Golang
Azure

Job description

Description

Senior Site Reliability Engineer

Why YOU want this position:

At Enverus, we’re committed to empowering the global quality of life by helping our customers make energy affordable and accessible to the world.

We are the most trusted energy-dedicated SaaS company, with a platform built to maximize value from generative AI, and our innovative solutions are reshaping the way energy is consumed and managed. By offering anytime, anywhere access to analytics and insights, we’re helping our customers make better decisions that help provide communities around the world with clean, affordable energy.

The energy industry is changing fast. But we’ve continued to lead the way in energy technology, creating intelligent connections across the entire energy ecosystem, from renewables, power and utilities, to oil and gas and financial institutions. Our solutions create more efficient production and distribution, capital allocation, renewable energy development, investment and sourcing, and help reduce costs by automating crucial business operations. Of course, this wouldn’t be possible without our people, which is why we have built a team of individuals from a diverse range of backgrounds.

Are you ready to help power the global quality of life? Join Enverus, and be a part of creating a brighter, more sustainable tomorrow.

We are currently seeking a Staff Site Reliability Engineer to join our Cloud Engineering team that manages our entire AWS presence. This role will be based in Canada.

This role offers the opportunity to join a rapidly growing company delivering industry-leading solutions to customers in the world’s most dynamic and fastest growing sector.

Performance Objectives

  • Work on a team that manages our entire global AWS presence
  • The team you will be working with is responsible for keeping our infrastructure humming as new releases and maintenance updates are rolled out
  • You will help organize, secure, and automate existing infrastructure and deployments
  • You will work closely with developers to provide feedback and drive operational improvements within our products and operations infrastructure
  • You will be responsible for ensuring that our platform is stable and balanced
  • Maintain high site up time, while embracing rapid change and growth
  • Scale infrastructure to meet increasing demand and evolving technology
  • Help the dev teams working on our code bases realize zero down-time deployments
  • Develop and improve operational practices and procedures
  • Implement, monitor, and maintain CI/CD frameworks
  • You will coordinate and participate in on-call rotations (day shifts)
  • Automate, automate, automate…

Competitive Candidate Profile

  • 5+ years of professional Windows and Linux server administration
  • 3+ years of Amazon Web Services (AWS) administration
  • 3+ years of experience within a high-performance, 24x7, DevOps, SysOps, or Operations team
  • You have excellent communication and collaboration skills
  • You demonstrate the ability to succeed in a high-pressure environment with rapidly changing priorities
  • You are an excellent problem solver, and willing to roll up your sleeves to take on any issue thrown your way
  • You have a desire not just to resolve problems, but to fully understand them and prevent them in the future
  • You seek out opportunities to improve, fix bugs, and challenge assumptions
  • You have experience working with global teams (North America, Europe, Asia)
  • You have experience with the following technologies:
    • Kubernetes (Container Orchestration)
    • Infrastructure as Code (Terraform, Cloudformation)
    • AWS (or other major cloud providers)
    • Python, C#, or Golang programming experience is a plus
    • Azure and Azure local experience is a plus
    • You prefer to lead the charge, not just keep up with it
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Site Reliability Engineer - Scale Global AWS Infra
Staff Site Reliability Engineer - Scale Global AWS Infra

Pearl Street Technologies • Pittsburgh

On-site
USD 130,000 - 180,000
Sr Software Engineer - 26255
Sr Software Engineer - 26255

Enverus • United States

On-site
USD 120,000 - 180,000
Competitive compensation package
Medical, Dental, Vision plan
401(k) retirement plan
Site Reliability Engineer
Site Reliability Engineer

FORT • United States

Hybrid
USD 150,000 - 180,000
Healthcare benefits
Flexible work environment
Large-scale cloud platform project
+1
Sr Software Engineer
Sr Software Engineer

Pearl Street Technologies • Spokane (WA)

On-site
USD 120,000 - 160,000
Medical, Dental, and Vision Plan
401(k) Retirement Savings Plan
Competitive Compensation Package
+1
Sr Software Engineer - 26255
Sr Software Engineer - 26255

Enverus • Houston (TX)

On-site
USD 120,000 - 180,000
Fast-paced environment
ME00528-Site Reliability Engineer 3
ME00528-Site Reliability Engineer 3

Momentum Engineering, Inc • Annapolis (MD)

On-site
USD 120,000 - 170,000
11 paid holidays
Minimum of 3 weeks PTO
Company-sponsored group medical plan
+4
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Mike Albert Fleet Solutions • Cincinnati (OH)

Hybrid
USD 100,000 - 135,000
Site Reliability Engineer
Site Reliability Engineer

TalentDome Staffing • United States

On-site
USD 140,000 - 210,000
Application Support Analyst - 26194
Application Support Analyst - 26194

Pearl Street Technologies • Pittsburgh

On-site
USD 70,000 - 95,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Jobgether • United States

Remote
USD 150,000 - 200,000
Competitive salary
Comprehensive healthcare coverage
401(k) plan with company matching
+3